Jietong Huasheng – a provider of artificial intelligence technologies and services
Free value-added services
AI office tools AI improves efficiency

Jietong Huasheng – a provider of artificial intelligence technologies and services

The development strategy of Lingyun Technology: “Rooted in Tsinghua, serving the world”

Tags:

What is Jietong Huasheng?

Beijing Jietong Huasheng Technology Co., Ltd. was established in October 2000; it has been dedicated to the research and development of technologies related to intelligent voice, intelligent semantics, intelligent vision, and big data analysis. The company provides AI capabilities, industry-specific products, and customized solutions to developers, enterprises, and public institutions.

The core technology platform of Jettone HiSound is named “Lingyun”; it offers an open platform, cloud services, local SDKs, options for private deployment, all-in-one solutions, as well as industry-specific solutions. It is not a simple individual chat tool.

Lingyun Platform Capability Matrix

AbilityMain inputsGenerate resultsTypical tasks
Speech recognitionReal-time audio or recordingsText and voice eventsCustomer service, meetings, transcripts, and in-vehicle use
Speech synthesisText and pronunciation parametersNatural speechBroadcasting, audio reading, and voice assistants
Semantic understandingUser speech and contextIntent, entities, and question-answer resultsCustomer service, navigation, and device control
Machine translationText or audio contentTarget language translation or speechInterlingual communication and document translation
OCRLicenses, invoices, bank cards, and document imagesText and structured fieldsAccount opening, finance, records management, and government affairs
Handwriting recognitionHandwritten trajectories or imagesElectronic textInput methods, terminals, and forms
Voice print recognitionSpeaker’s voiceVerification or identification resultsRemote identity authentication and security
Faces and fingerprintsFacial image or fingerprint featuresIdentity comparison resultsAuthentication, access control, and business verification
Data miningBusiness and interaction dataAnalysis, classification, and insightsOperations, quality control, and risk analysis

Speech Recognition V10.2

Lingyun Speech Recognition V10.2 was released in 2026; it utilizes self-developed large models to reengineer audio processing, language handling, end-to-end decoding, and noise suppression techniques. This version relies on collaborative reasoning between small and large models – the small models handle tasks that require low latency, while the large models are responsible for language modeling, resolving ambiguity, and ensuring the accuracy of meaning in long texts.

  • Improve the recognition of spoken language, dialect accents, and specialized terminology.
  • Improves handling of multiple speakers speaking at once, homophones, and sentence segmentation in long texts.
  • Combines microphone arrays, noise reduction, echo cancellation, and reverb removal.
  • It supports stream-based recognition and high-concurrency deployment.
  • It covers Mandarin, Chinese-English mixtures, multiple languages, and various dialects.
  • Models in the fields of finance, law, healthcare, and government services are provided.

The accuracy and real-time performance shown on the public page are based on specific testing criteria, and they do not constitute guarantees under any conditions related to equipment, noise, or application areas. The purchaser should conduct re-evaluation using actual recordings, industry-specific terminology, and the target hardware.

Voice recognition delivery method

Delivery formatInternet dependenceSuitable scenariosFocus on confirmation
Cloud APIInternet connection is required.Fast access and flexible invocationConcurrency, latency, billing, and data transmission
SDK componentsBased on cloud or on-premises capabilitiesMobile, desktop, and embedded applicationsPlatform, resources, and authorization period
Private deploymentCan run in an institutional environmentFinance, government affairs, and sensitive operationsHardware, upgrades, operation and maintenance, and data boundaries
All-in-one machineAccording to the equipment planStandardized local voice tasksCapacity, redundancy, and after-sales support

V10.2 also provides explicit support for the aarch64 architecture, as well as domestic hardware platforms such as Hygon and Ascend. The specific models, drivers, throughput rates, and precision levels need to be determined through testing specific to each project.

Speech synthesis

Lingyun text-to-speech converts text into male, female, childlike, and other voices; it allows for adjustment of speech speed, pitch, and volume. The capabilities page lists 21 languages, and it supports mixed Chinese and English speech, as well as the use of specialized vocabulary lists and customized voice profiles.

  • Suitable for use at airports, stations, hospitals, and for public information announcements.
  • It can generate voices for customer service, call centers, and voice assistants.
  • It supports audio reading, navigation, and device alerts.
  • The pronunciation can be customized for characters, phrases, and sentences.
  • Cloud and offline engines can be combined based on network conditions.

The fact that a particular timbre is available does not mean that unlimited commercial use is permitted; custom timbres may also involve issues related to the rights of real people’s voices. The contract should specify consent for data collection, the purposes for which it will be used, the duration of its validity, rules regarding sub-licensing, and procedures for deletion.

Semantic understanding and multi-turn dialogue

Semantic understanding capabilities are used to identify user intentions, entity attributes, and context, and they enable multi-turn conversations as well as intelligent question-answering. By combining speech recognition with speech synthesis, these capabilities can be utilized for interactions in customer service, automotive systems, home devices, or robots.

  • Interpret intents in areas such as weather, navigation, flights, hotels, and news.
  • Handles phone calls, text messages, address books, and application control commands.
  • Supports devices, music, videos, and in-vehicle controls.
  • Dictionaries, intents, and models can be customized for vertical businesses.
  • Supports a combination of cloud and on-premises capabilities.

Traditional NLU and large-model agents operate at different levels of capability; the former focuses on controlled intent recognition and field parsing, while the latter is better suited for handling complex knowledge, workflow orchestration, and generative interactions.

OCR and document recognition

Lingyun OCR can process IDs, invoices, bank cards, business cards, and ordinary documents, converting images into text or structured fields. Paper documents can also be turned into searchable dual-layer PDFs.

Identification typeEnterOutputUse cases
License recognitionID cards, passports, and vehicle documentsIdentity and document fieldsAccount opening, border control, and government services
Bill recognitionInvoices, bank drafts, and custom-made draftsAmount, number, and business fieldsFinance and Auditing
Bank card recognitionPicture of bank cardCard number, bank, and card typeFinancial data entry
Document recognitionForms, books, newspapers, and screenshotsText or double-layer PDFDigitalization of archives
Business card recognitionBusiness card photoContact fieldCustomer management

The restrictions on images in online experiences may differ from those applied to commercial SDKs. Licenses, bank cards, and invoices contain sensitive information; therefore, public demonstrations should not be used to process real customer data.

Machine translation

Machine translation supports text and voice translation, covering various language combinations such as Chinese-English, Uyghur-Chinese, Chinese-Japanese, and Chinese-Korean. Developers can integrate this translation functionality into third-party applications through APIs.

Contracts, patents, medical and compliance documents should be reviewed by professionals, with particular attention paid to numbers, negations, terminology, and legal implications. Multilingual lists do not ensure that all language pairs have the same quality.

Voice print, facial recognition, and fingerprint recognition

Voiceprint recognition carries out one-to-one or one-to-many verification by comparing the acoustic characteristics of speakers, and it can be combined with facial recognition and fingerprint scanning to create a multi-dimensional authentication method. Voiceprints can also be used for separating different speakers and analyzing their emotional tendencies.

  • Suitable for remote financial identity verification and account protection.
  • It can be used for social security, security systems, and identity verification in smart devices.
  • It supports verification methods such as free input and fixed text.
  • It allows for distinguishing between speakers in multi-person conversations.
  • Algorithms’ scoring cannot be used as the sole basis for determining identity.

Voice prints, facial features, and fingerprints are highly sensitive biometric data; errors in processing them can lead to false rejections or incorrect identifications. Projects of this kind must include mechanisms for liveness detection, manual verification, threshold validation, and appeal processes.

Government AI Agent Cluster

Jietong Huasheng combines large models, intelligent speech technology, semantic understanding, and data analysis to create intelligent agent clusters for government hotline services; these clusters handle task acceptance, support for agents, ticket processing, as well as monitoring and optimization. It is a solution designed for organizations, and it is not a robot that individuals can purchase directly.

AgentMain tasksEnterOutput
Text customer serviceMultiple channels for consultation and processingCitizen issues and knowledge baseAnswers, guidance, and business entry points
Voice navigationMultiple rounds of querying and filling in fieldsPhone voiceIntent, fields, and navigation results
Knowledge AssistantManaging government affairs knowledgeMulti-format policies and business documentsKnowledge entries and recommendations
Form filler assistantExtract ticket elementsPublic demands and callsStructured work order
Order distribution assistanceSelect the appropriate organizing entitySubject, region, and matterSuggestions for categorization and task assignment
Agent assistanceSupport before, during, and after speechPortraits, conversations, and rulesPhrases, alerts, and summaries
Outbound follow-up callsFollow-up, notifications, and remindersLists, scripts, and ticketsOutbound call results and statistics
Quality inspection and supervisionComprehensive quality inspection and ticket monitoringRecordings, tickets, and rulesRisks, early warnings, and reporting

The efficiency and accuracy values mentioned in the product documentation represent the parameters defined by the solution provider; they should not be used directly as the criteria for acceptance upon purchase. For each project, it is necessary to establish new parameters regarding the dataset, baseline, allowable errors, and acceptance methods.

Intelligent agent for quality inspection of financial dual-recording

The dual-recording quality inspection agent integrates recording functions, business rules, speech analysis, computer vision, and large-scale models to serve the banking, insurance, securities, trust, and consumer finance sectors. It is capable of processing both historical videos and newly generated high-definition content.

  • Verify obstruction, framing, facial features, and card/document information.
  • Detect exaggerated marketing, aggressive sales tactics, and sensitive wording.
  • It can recognize actions such as nodding, shaking the head, signing, and speaking aloud.
  • Identify emotional changes and mark the controversial segments.
  • It supports online signing and handwriting recognition.
  • Use the results of manual review to optimize subsequent quality checks.

Dual-recording deployment mode

Deployment modeSuitable for institutionsPrimary valueKey points of procurement
Unified deployment by the centerUnified data centerCentralized management model and rulesNetwork, throughput, and disaster recovery
Branch cache deploymentMulti-location groupTaking into account both central management and local processing.Synchronization, caching, and version consistency
Private deploymentHighly sensitive financial institutionsThe data remains in a controlled environment.Hardware, upgrades, auditing, and deletion
Multi-tenant managementGroup and service platformIsolate data and permissions of different organizationsTenant boundaries and administrator permissions

Intelligent customer service and voice analysis

Lingyun’s products also include intelligent customer service, automated outbound calls, voice navigation, as well as offline or real-time voice analysis. Companies can combine transcription, intent recognition, knowledge-based Q&A, process execution, and quality control to create a comprehensive call center solution.

  1. Access via phone, web, apps, or terminal channels.
  2. Use speech recognition to convert calls into text.
  3. Analyze intent, entities, emotions, and risk words.
  4. Query the knowledge base or trigger business processes.
  5. Generate a response using text-to-speech or connect to a human agent.
  6. Save the necessary logs and perform quality checks.
  7. Optimize rules and models based on the results of manual review.

Developer onboarding process

  1. Register for a Lingyun developer account and link it to a mobile phone or email address.
  2. Create the application and fill in the necessary description.
  3. Select the cloud or on-premises capability and the corresponding capKey.
  4. Download the SDK of the target system as well as local resources.
  5. Configure appKey, developerKey, cloudUrl, and capKey.
  6. Run the examples included with the SDK to verify authorization and the resulting output.
  7. Test accuracy, latency, and anomaly handling on real data.
  8. After the testing is complete, contact sales to switch to a commercial version.

Test application rules

ProjectCurrent rulesImpact
SDK DownloadAvailable for free downloadThis does not mean that commercial services will remain free forever.
Testing periodEach test application is used for half a year.It can be recreated if testing is conducted after the expiration date.
Number of applicationsUp to 10 per developerPlanning capabilities and a platform are required.
Apply deletionCreated applications cannot be deleted.Avoid occupying spots arbitrarily.
Example programSome platform SDKs come built-in.Facilitates the verification of keys and capabilities
Official commercial useContact BusinessAuthorization, billing, and service level need to be confirmed.

SDKs, APIs, and interfaces

The Lingyun Open Platform utilizes application keys, developer keys, cloud addresses, and capability identifiers to carry out authentication and to select the appropriate capabilities. Cloud-based capabilities require an internet connection, while local capabilities need resources to be downloaded and authorization to be obtained.

Access projectConfirmabilityPrecautions
Android SDKDownloads and examples are available.Check the system version, architecture, and permissions.
iOS SDKDownloads and examples are available.Check the signature, privacy policy, and approval rules.
Windows C++Downloads and examples are available.Verify the runtime library and bitness.
Windows JavaDownloads and examples are available.Verify JDK and native dependencies
Linux and domestic solutionsPartial capability supportObtain the compatibility list by chip and system.
ASR interfaceWeb requests, Sockets, WebServices, MRCPDifferent deliveries may support different protocols.
Translation APIConfirm that interface calls will be provided.Pairs of languages, characters, and concurrency require confirmation via contract.

Price and procurement methods

Package or versionPriceBilling cycleCore benefits or quotaSuitable for users
Development testingFree SDK downloadTest the application for half a yearCreate an application, download the SDK, and conduct capability testingDevelopers and prototype team
Commercial use of cloud APIsContact salesBy call or contractCloud capabilities, concurrency, and technical supportNetworked applications and platforms
Commercial use of local SDKCustom quoteAccording to the scope of authorizationAuthorization for local engines, resources, and devicesTerminals and offline applications
Private deploymentCustom quoteBy project and maintenanceModels, servers, deployment, and operation and maintenanceFinance, government affairs, and large enterprises
All-in-one machineCustom quoteProcurement and Service ContractsHardware, software, and standard capabilitiesLocal standardization tasks
Industry agentsContact salesBy module, scale, and service periodKnowledge, processes, models, and integrationHotlines, financial, and industry clients

The public page does not provide information regarding the unified production price, the amount of free usage, the tiering system for concurrent tasks, renewal procedures, refund policies, or the rules for handling data after service suspension. The pricing for the project shall be determined based on the final contract, the configuration list, and the scope of acceptance.

How to select and verify

  1. Identify the types of data: audio, images, text, and identity data.
  2. Choose cloud, on-premises, private, or hybrid deployment.
  3. Prepare a test set that includes real-world noise, accents, and business-related scenarios.
  4. Agree on accuracy, recall, latency, throughput, and availability.
  5. Test network disconnection, timeouts, error responses, and manual takeover.
  6. Review keys, permissions, logs, backups, and deletion.
  7. Confirm compatibility with upgrades, model iterations, and domestic solutions.
  8. Include the acceptance criteria, metrics, and procedures for handling breaches in the contract.

Privacy and data security

Currently, the enterprise site and the open platform do not provide a comprehensive and unified privacy policy, data processing agreements, a list of sub-processors, information on retention periods, or rules regarding the discontinuation of model training. When it comes to calls, certificates, and biometric data, it is not sufficient to rely solely on the product’s feature pages.

  • Confirm the roles of the customer, Jietong Huasheng, and third parties.
  • Confirm the cloud transmission location, storage location, and backup scope.
  • Confirm the retention period for recordings, texts, images, and feature templates.
  • Verify whether business data is used for model training and optimization.
  • Verify user permissions, audit logs, and asset recovery upon departure.
  • Confirmation of proof of deletion, return, and destruction.
  • Confirm the notification procedures for security incidents and the time limits for emergency response.

Special attention to biometric recognition

Voice prints, facial features, fingerprints, identification documents, and financial account information are highly sensitive; therefore, a lawful, adequate, and clear basis for processing these data is required before they can be collected. It cannot be assumed that such data can be collected just because an interface is capable of recognizing them.

  • Explain the purpose and necessity of identity authentication.
  • Provide non-biometric alternatives.
  • Distinguish between the original sample and the feature template.
  • Prevents voice print replay, facial photo, and forgery attacks.
  • Provide options for manual review and user appeals.
  • After the service is completed, the data and copies are deleted as agreed.

Copyright, timbre, and commercial use

The fact that the SDK can be downloaded does not mean that the models, sound libraries, recognition resources, and synthesized voices can be used for commercial purposes without any restrictions. Companies must comply with the terms of the commercial licensing agreement regarding the number of devices, the amount of usage, geographic location, duration, and rights to redistribute the materials.

For voice replication and the use of customized speakers, it is necessary to obtain consent from the person whose voice is being used, as well as permission for recording; additionally, considerations must be given to the appropriate scenarios in which such voices can be used and to mechanisms for withdrawing them. When using these voices for broadcasting, it is important to avoid impersonating real people, providing false endorsements, or using synthetic content without proper identification.

GitHub and the boundaries of open source

ProjectPublic statusLicenseCorrect understanding
Lingyun Intelligent Input Method Android ExamplePublic codeMITOpen-source client examples
Example of Android wake-up using Lingyun voice recognitionPublic codeMITOpen-source client examples
Lingyun SDKAvailable for downloadCommercial terms to be confirmed separately.Being downloadable does not mean it is open source.
Cloud service codeNot disclosedNot yet made publicEvaluate based on closed-source services
Model weights and training dataNot disclosedNot yet made publicIt cannot be considered an open-source model.
Industry AI agent productsUnpublished codeIn accordance with the contractIt belongs to commercial solutions.

Supported platforms

Platform or environmentConfirm statusExplanation
Web open platformAvailable nowRegister, create applications, documents, and experience various capabilities
AndroidSDK supportSuitable for mobile and smart devices
iOSSDK supportSuitable for mobile apps
WindowsC++ and Java SDKsSuitable for desktops and business systems
LinuxPartial capability supportConfirm by product and architecture
Domestic chips and acceleration cardsSome new features are supported.A specific compatibility list is required.
Private cloud and all-in-one devicesDeliverableConfigure by project

The fact that an SDK is supported on a certain platform does not mean that there exists a unified, consumer-oriented cloud-native application; personal applications, enterprise components, and industry-specific systems should be considered separately.

Which users are it suitable for

  • Application developers who need voice, OCR, and semantic capabilities.
  • Companies that develop customer service, outbound calling, and quality control systems.
  • Organizations that need the digitization of meetings, transcripts, and documents.
  • Financial services related to the handling of licenses, documents, and identity verification.
  • Public departments that are building intelligent agent clusters for government service hotlines.
  • Large customers that require local and domestically-based deployment.
  • Manufacturers that develop products for use in vehicles, home appliances, speakers, and robots.

Main advantages

  • It covers a variety of capabilities including voice, semantics, vision, and biometrics.
  • Supports cloud, on-premises, private, and all-in-one delivery.
  • SDKs, documentation, and examples assist developers in creating prototypes.
  • It can be customized for industry-specific terminology, timbre, templates, and models.
  • Government and financial agents cover the entire business process.
  • The new version of speech recognition takes into account both the semantic capabilities of large models and low latency.

Capacity boundaries

  • The accuracy of promotional materials cannot replace actual business validation.
  • The free download of the SDK does not mean that its use in production or for commercial purposes is also free.
  • The testing applications are available for only half a year and their number is limited.
  • The commercial pricing, concurrency limits, SLAs, and refund policies are not publicly available in a unified manner.
  • Biometric recognition involves errors and poses a high risk to privacy.
  • The rules for complete data preservation and model training must be confirmed through a contract.
  • Different capabilities, systems, and chip support ranges vary.
  • The two open-source examples do not imply that the platform and the models are open source.

Frequently Asked Questions

Is iFlytek suitable for individual direct use?

It is primarily aimed at developers and organizations. Individuals can experience some of its features or use it in their own applications, but the core value of Lingyun lies in its SDKs, APIs, privacy-focused solutions, and industry-specific solutions.

Is the Lingyun SDK permanently free?

The SDK can be downloaded for free, with a trial period of six months for testing purposes. For full-scale production, cloud-based usage, local resources, and commercial licensing, further arrangements are required.

How many test applications can each developer create?

The current development guidelines specify that a maximum of 10 applications can be created, and the applications that have already been created cannot be deleted. It is necessary to plan the number of applications in advance before selecting capabilities and platforms.

Does Lingyun support private deployment?

Supported. Voice recognition, dual-recording quality inspection, and other enterprise solutions can be implemented in a private or on-premises format; the hardware requirements, number of concurrent tasks, upgrade options, and costs are determined based on each project.

What voice interfaces does Lingyun offer?

The speech recognition page confirms that it supports web requests, Socket, WebService, MRCP, and other methods; the specific protocols depend on the capabilities and the version in use.

Does Lingyun’s GitHub code represent an open-source product?

This is not the case. To date, only two Android example projects under the MIT license have been identified; cloud services, SDKs, model weights, and industry-specific products have not been made open source as a result.

Can voice print recognition serve as the sole means of identity verification?

It is not recommended. Live detection, as well as risk control for devices and business operations, should be employed, and manual review along with other verification methods should be available in case of misjudgments.

Does Jietong Huasheng disclose its prices?

The development and testing rules are made public, but there are no unified publicly available prices for the production APIs, offline SDKs, private versions, all-in-one solutions, and agents; formal quotes are required to obtain such information.

How to add a FAQ Schema

This page already provides FAQs that are visible to users; the site template can use these questions and answers to generate structured FAQ data. Script tags are not inserted directly into the text in order to meet the security requirements for content fields.

  1. Only use the questions and answers that are visible directly to the user on the page.
  2. The page object is set to FAQPage.
  3. Each question is set as Question.
  4. The answer is written to acceptedAnswer and set as Answer.
  5. The testing period, quantity, and open-source boundaries must be consistent with the main text.
  6. Update the FAQ simultaneously when there are changes in the product version or commercial terms.
  7. After publishing, check for grammar, duplicates, and search engine guidelines.

Summary

Jietong Huasheng’s strength lies in its ability to combine voice technology, semantics, OCR, biometric recognition, and large-model agents to create enterprise-level solutions, offering delivery options via the cloud, on-premises, in a private environment, or as domestically developed systems. It is suitable for projects that require deep integration with industry-specific processes.

When making a selection, do not rely solely on the accuracy shown in demonstrations or on free SDKs; instead, verify the precision, latency, and stability using actual data. Commercial licensing terms, handling of sensitive data, model training, deletion processes, SLAs, and upgrade options must all be specified in the contract.

©️Copyright notice: Unless otherwise specified, all articles on this site are copyrighted bySharing of AI toolsAll content on this site is original; without permission, no individual, media outlet, website, or organization may reproduce, copy, or otherwise distribute it, nor may they create mirrors of it on servers that are not owned by this site. Otherwise, we reserve the right to take legal action against such parties in accordance with the law.

Tools similar to iFlytek – an AI technology and services provider.