Jietong Huasheng – a provider of artificial intelligence technologies and services
The development strategy of Lingyun Technology: “Rooted in Tsinghua, serving the world”
Tags:AI improves efficiencyWhat is Jietong Huasheng?
Beijing Jietong Huasheng Technology Co., Ltd. was established in October 2000; it has been dedicated to the research and development of technologies related to intelligent voice, intelligent semantics, intelligent vision, and big data analysis. The company provides AI capabilities, industry-specific products, and customized solutions to developers, enterprises, and public institutions.
The core technology platform of Jettone HiSound is named “Lingyun”; it offers an open platform, cloud services, local SDKs, options for private deployment, all-in-one solutions, as well as industry-specific solutions. It is not a simple individual chat tool.
Lingyun Platform Capability Matrix
| Ability | Main inputs | Generate results | Typical tasks |
|---|---|---|---|
| Speech recognition | Real-time audio or recordings | Text and voice events | Customer service, meetings, transcripts, and in-vehicle use |
| Speech synthesis | Text and pronunciation parameters | Natural speech | Broadcasting, audio reading, and voice assistants |
| Semantic understanding | User speech and context | Intent, entities, and question-answer results | Customer service, navigation, and device control |
| Machine translation | Text or audio content | Target language translation or speech | Interlingual communication and document translation |
| OCR | Licenses, invoices, bank cards, and document images | Text and structured fields | Account opening, finance, records management, and government affairs |
| Handwriting recognition | Handwritten trajectories or images | Electronic text | Input methods, terminals, and forms |
| Voice print recognition | Speaker’s voice | Verification or identification results | Remote identity authentication and security |
| Faces and fingerprints | Facial image or fingerprint features | Identity comparison results | Authentication, access control, and business verification |
| Data mining | Business and interaction data | Analysis, classification, and insights | Operations, quality control, and risk analysis |
Speech Recognition V10.2
Lingyun Speech Recognition V10.2 was released in 2026; it utilizes self-developed large models to reengineer audio processing, language handling, end-to-end decoding, and noise suppression techniques. This version relies on collaborative reasoning between small and large models – the small models handle tasks that require low latency, while the large models are responsible for language modeling, resolving ambiguity, and ensuring the accuracy of meaning in long texts.
- Improve the recognition of spoken language, dialect accents, and specialized terminology.
- Improves handling of multiple speakers speaking at once, homophones, and sentence segmentation in long texts.
- Combines microphone arrays, noise reduction, echo cancellation, and reverb removal.
- It supports stream-based recognition and high-concurrency deployment.
- It covers Mandarin, Chinese-English mixtures, multiple languages, and various dialects.
- Models in the fields of finance, law, healthcare, and government services are provided.
The accuracy and real-time performance shown on the public page are based on specific testing criteria, and they do not constitute guarantees under any conditions related to equipment, noise, or application areas. The purchaser should conduct re-evaluation using actual recordings, industry-specific terminology, and the target hardware.
Voice recognition delivery method
| Delivery format | Internet dependence | Suitable scenarios | Focus on confirmation |
|---|---|---|---|
| Cloud API | Internet connection is required. | Fast access and flexible invocation | Concurrency, latency, billing, and data transmission |
| SDK components | Based on cloud or on-premises capabilities | Mobile, desktop, and embedded applications | Platform, resources, and authorization period |
| Private deployment | Can run in an institutional environment | Finance, government affairs, and sensitive operations | Hardware, upgrades, operation and maintenance, and data boundaries |
| All-in-one machine | According to the equipment plan | Standardized local voice tasks | Capacity, redundancy, and after-sales support |
V10.2 also provides explicit support for the aarch64 architecture, as well as domestic hardware platforms such as Hygon and Ascend. The specific models, drivers, throughput rates, and precision levels need to be determined through testing specific to each project.
Speech synthesis
Lingyun text-to-speech converts text into male, female, childlike, and other voices; it allows for adjustment of speech speed, pitch, and volume. The capabilities page lists 21 languages, and it supports mixed Chinese and English speech, as well as the use of specialized vocabulary lists and customized voice profiles.
- Suitable for use at airports, stations, hospitals, and for public information announcements.
- It can generate voices for customer service, call centers, and voice assistants.
- It supports audio reading, navigation, and device alerts.
- The pronunciation can be customized for characters, phrases, and sentences.
- Cloud and offline engines can be combined based on network conditions.
The fact that a particular timbre is available does not mean that unlimited commercial use is permitted; custom timbres may also involve issues related to the rights of real people’s voices. The contract should specify consent for data collection, the purposes for which it will be used, the duration of its validity, rules regarding sub-licensing, and procedures for deletion.
Semantic understanding and multi-turn dialogue
Semantic understanding capabilities are used to identify user intentions, entity attributes, and context, and they enable multi-turn conversations as well as intelligent question-answering. By combining speech recognition with speech synthesis, these capabilities can be utilized for interactions in customer service, automotive systems, home devices, or robots.
- Interpret intents in areas such as weather, navigation, flights, hotels, and news.
- Handles phone calls, text messages, address books, and application control commands.
- Supports devices, music, videos, and in-vehicle controls.
- Dictionaries, intents, and models can be customized for vertical businesses.
- Supports a combination of cloud and on-premises capabilities.
Traditional NLU and large-model agents operate at different levels of capability; the former focuses on controlled intent recognition and field parsing, while the latter is better suited for handling complex knowledge, workflow orchestration, and generative interactions.
OCR and document recognition
Lingyun OCR can process IDs, invoices, bank cards, business cards, and ordinary documents, converting images into text or structured fields. Paper documents can also be turned into searchable dual-layer PDFs.
| Identification type | Enter | Output | Use cases |
|---|---|---|---|
| License recognition | ID cards, passports, and vehicle documents | Identity and document fields | Account opening, border control, and government services |
| Bill recognition | Invoices, bank drafts, and custom-made drafts | Amount, number, and business fields | Finance and Auditing |
| Bank card recognition | Picture of bank card | Card number, bank, and card type | Financial data entry |
| Document recognition | Forms, books, newspapers, and screenshots | Text or double-layer PDF | Digitalization of archives |
| Business card recognition | Business card photo | Contact field | Customer management |
The restrictions on images in online experiences may differ from those applied to commercial SDKs. Licenses, bank cards, and invoices contain sensitive information; therefore, public demonstrations should not be used to process real customer data.
Machine translation
Machine translation supports text and voice translation, covering various language combinations such as Chinese-English, Uyghur-Chinese, Chinese-Japanese, and Chinese-Korean. Developers can integrate this translation functionality into third-party applications through APIs.
Contracts, patents, medical and compliance documents should be reviewed by professionals, with particular attention paid to numbers, negations, terminology, and legal implications. Multilingual lists do not ensure that all language pairs have the same quality.
Voice print, facial recognition, and fingerprint recognition
Voiceprint recognition carries out one-to-one or one-to-many verification by comparing the acoustic characteristics of speakers, and it can be combined with facial recognition and fingerprint scanning to create a multi-dimensional authentication method. Voiceprints can also be used for separating different speakers and analyzing their emotional tendencies.
- Suitable for remote financial identity verification and account protection.
- It can be used for social security, security systems, and identity verification in smart devices.
- It supports verification methods such as free input and fixed text.
- It allows for distinguishing between speakers in multi-person conversations.
- Algorithms’ scoring cannot be used as the sole basis for determining identity.
Voice prints, facial features, and fingerprints are highly sensitive biometric data; errors in processing them can lead to false rejections or incorrect identifications. Projects of this kind must include mechanisms for liveness detection, manual verification, threshold validation, and appeal processes.
Government AI Agent Cluster
Jietong Huasheng combines large models, intelligent speech technology, semantic understanding, and data analysis to create intelligent agent clusters for government hotline services; these clusters handle task acceptance, support for agents, ticket processing, as well as monitoring and optimization. It is a solution designed for organizations, and it is not a robot that individuals can purchase directly.
| Agent | Main tasks | Enter | Output |
|---|---|---|---|
| Text customer service | Multiple channels for consultation and processing | Citizen issues and knowledge base | Answers, guidance, and business entry points |
| Voice navigation | Multiple rounds of querying and filling in fields | Phone voice | Intent, fields, and navigation results |
| Knowledge Assistant | Managing government affairs knowledge | Multi-format policies and business documents | Knowledge entries and recommendations |
| Form filler assistant | Extract ticket elements | Public demands and calls | Structured work order |
| Order distribution assistance | Select the appropriate organizing entity | Subject, region, and matter | Suggestions for categorization and task assignment |
| Agent assistance | Support before, during, and after speech | Portraits, conversations, and rules | Phrases, alerts, and summaries |
| Outbound follow-up calls | Follow-up, notifications, and reminders | Lists, scripts, and tickets | Outbound call results and statistics |
| Quality inspection and supervision | Comprehensive quality inspection and ticket monitoring | Recordings, tickets, and rules | Risks, early warnings, and reporting |
The efficiency and accuracy values mentioned in the product documentation represent the parameters defined by the solution provider; they should not be used directly as the criteria for acceptance upon purchase. For each project, it is necessary to establish new parameters regarding the dataset, baseline, allowable errors, and acceptance methods.
Intelligent agent for quality inspection of financial dual-recording
The dual-recording quality inspection agent integrates recording functions, business rules, speech analysis, computer vision, and large-scale models to serve the banking, insurance, securities, trust, and consumer finance sectors. It is capable of processing both historical videos and newly generated high-definition content.
- Verify obstruction, framing, facial features, and card/document information.
- Detect exaggerated marketing, aggressive sales tactics, and sensitive wording.
- It can recognize actions such as nodding, shaking the head, signing, and speaking aloud.
- Identify emotional changes and mark the controversial segments.
- It supports online signing and handwriting recognition.
- Use the results of manual review to optimize subsequent quality checks.
Dual-recording deployment mode
| Deployment mode | Suitable for institutions | Primary value | Key points of procurement |
|---|---|---|---|
| Unified deployment by the center | Unified data center | Centralized management model and rules | Network, throughput, and disaster recovery |
| Branch cache deployment | Multi-location group | Taking into account both central management and local processing. | Synchronization, caching, and version consistency |
| Private deployment | Highly sensitive financial institutions | The data remains in a controlled environment. | Hardware, upgrades, auditing, and deletion |
| Multi-tenant management | Group and service platform | Isolate data and permissions of different organizations | Tenant boundaries and administrator permissions |
Intelligent customer service and voice analysis
Lingyun’s products also include intelligent customer service, automated outbound calls, voice navigation, as well as offline or real-time voice analysis. Companies can combine transcription, intent recognition, knowledge-based Q&A, process execution, and quality control to create a comprehensive call center solution.
- Access via phone, web, apps, or terminal channels.
- Use speech recognition to convert calls into text.
- Analyze intent, entities, emotions, and risk words.
- Query the knowledge base or trigger business processes.
- Generate a response using text-to-speech or connect to a human agent.
- Save the necessary logs and perform quality checks.
- Optimize rules and models based on the results of manual review.
Developer onboarding process
- Register for a Lingyun developer account and link it to a mobile phone or email address.
- Create the application and fill in the necessary description.
- Select the cloud or on-premises capability and the corresponding capKey.
- Download the SDK of the target system as well as local resources.
- Configure appKey, developerKey, cloudUrl, and capKey.
- Run the examples included with the SDK to verify authorization and the resulting output.
- Test accuracy, latency, and anomaly handling on real data.
- After the testing is complete, contact sales to switch to a commercial version.
Test application rules
| Project | Current rules | Impact |
|---|---|---|
| SDK Download | Available for free download | This does not mean that commercial services will remain free forever. |
| Testing period | Each test application is used for half a year. | It can be recreated if testing is conducted after the expiration date. |
| Number of applications | Up to 10 per developer | Planning capabilities and a platform are required. |
| Apply deletion | Created applications cannot be deleted. | Avoid occupying spots arbitrarily. |
| Example program | Some platform SDKs come built-in. | Facilitates the verification of keys and capabilities |
| Official commercial use | Contact Business | Authorization, billing, and service level need to be confirmed. |
SDKs, APIs, and interfaces
The Lingyun Open Platform utilizes application keys, developer keys, cloud addresses, and capability identifiers to carry out authentication and to select the appropriate capabilities. Cloud-based capabilities require an internet connection, while local capabilities need resources to be downloaded and authorization to be obtained.
| Access project | Confirmability | Precautions |
|---|---|---|
| Android SDK | Downloads and examples are available. | Check the system version, architecture, and permissions. |
| iOS SDK | Downloads and examples are available. | Check the signature, privacy policy, and approval rules. |
| Windows C++ | Downloads and examples are available. | Verify the runtime library and bitness. |
| Windows Java | Downloads and examples are available. | Verify JDK and native dependencies |
| Linux and domestic solutions | Partial capability support | Obtain the compatibility list by chip and system. |
| ASR interface | Web requests, Sockets, WebServices, MRCP | Different deliveries may support different protocols. |
| Translation API | Confirm that interface calls will be provided. | Pairs of languages, characters, and concurrency require confirmation via contract. |
Price and procurement methods
| Package or version | Price | Billing cycle | Core benefits or quota | Suitable for users |
|---|---|---|---|---|
| Development testing | Free SDK download | Test the application for half a year | Create an application, download the SDK, and conduct capability testing | Developers and prototype team |
| Commercial use of cloud APIs | Contact sales | By call or contract | Cloud capabilities, concurrency, and technical support | Networked applications and platforms |
| Commercial use of local SDK | Custom quote | According to the scope of authorization | Authorization for local engines, resources, and devices | Terminals and offline applications |
| Private deployment | Custom quote | By project and maintenance | Models, servers, deployment, and operation and maintenance | Finance, government affairs, and large enterprises |
| All-in-one machine | Custom quote | Procurement and Service Contracts | Hardware, software, and standard capabilities | Local standardization tasks |
| Industry agents | Contact sales | By module, scale, and service period | Knowledge, processes, models, and integration | Hotlines, financial, and industry clients |
The public page does not provide information regarding the unified production price, the amount of free usage, the tiering system for concurrent tasks, renewal procedures, refund policies, or the rules for handling data after service suspension. The pricing for the project shall be determined based on the final contract, the configuration list, and the scope of acceptance.
How to select and verify
- Identify the types of data: audio, images, text, and identity data.
- Choose cloud, on-premises, private, or hybrid deployment.
- Prepare a test set that includes real-world noise, accents, and business-related scenarios.
- Agree on accuracy, recall, latency, throughput, and availability.
- Test network disconnection, timeouts, error responses, and manual takeover.
- Review keys, permissions, logs, backups, and deletion.
- Confirm compatibility with upgrades, model iterations, and domestic solutions.
- Include the acceptance criteria, metrics, and procedures for handling breaches in the contract.
Privacy and data security
Currently, the enterprise site and the open platform do not provide a comprehensive and unified privacy policy, data processing agreements, a list of sub-processors, information on retention periods, or rules regarding the discontinuation of model training. When it comes to calls, certificates, and biometric data, it is not sufficient to rely solely on the product’s feature pages.
- Confirm the roles of the customer, Jietong Huasheng, and third parties.
- Confirm the cloud transmission location, storage location, and backup scope.
- Confirm the retention period for recordings, texts, images, and feature templates.
- Verify whether business data is used for model training and optimization.
- Verify user permissions, audit logs, and asset recovery upon departure.
- Confirmation of proof of deletion, return, and destruction.
- Confirm the notification procedures for security incidents and the time limits for emergency response.
Special attention to biometric recognition
Voice prints, facial features, fingerprints, identification documents, and financial account information are highly sensitive; therefore, a lawful, adequate, and clear basis for processing these data is required before they can be collected. It cannot be assumed that such data can be collected just because an interface is capable of recognizing them.
- Explain the purpose and necessity of identity authentication.
- Provide non-biometric alternatives.
- Distinguish between the original sample and the feature template.
- Prevents voice print replay, facial photo, and forgery attacks.
- Provide options for manual review and user appeals.
- After the service is completed, the data and copies are deleted as agreed.
Copyright, timbre, and commercial use
The fact that the SDK can be downloaded does not mean that the models, sound libraries, recognition resources, and synthesized voices can be used for commercial purposes without any restrictions. Companies must comply with the terms of the commercial licensing agreement regarding the number of devices, the amount of usage, geographic location, duration, and rights to redistribute the materials.
For voice replication and the use of customized speakers, it is necessary to obtain consent from the person whose voice is being used, as well as permission for recording; additionally, considerations must be given to the appropriate scenarios in which such voices can be used and to mechanisms for withdrawing them. When using these voices for broadcasting, it is important to avoid impersonating real people, providing false endorsements, or using synthetic content without proper identification.
GitHub and the boundaries of open source
| Project | Public status | License | Correct understanding |
|---|---|---|---|
| Lingyun Intelligent Input Method Android Example | Public code | MIT | Open-source client examples |
| Example of Android wake-up using Lingyun voice recognition | Public code | MIT | Open-source client examples |
| Lingyun SDK | Available for download | Commercial terms to be confirmed separately. | Being downloadable does not mean it is open source. |
| Cloud service code | Not disclosed | Not yet made public | Evaluate based on closed-source services |
| Model weights and training data | Not disclosed | Not yet made public | It cannot be considered an open-source model. |
| Industry AI agent products | Unpublished code | In accordance with the contract | It belongs to commercial solutions. |
Supported platforms
| Platform or environment | Confirm status | Explanation |
|---|---|---|
| Web open platform | Available now | Register, create applications, documents, and experience various capabilities |
| Android | SDK support | Suitable for mobile and smart devices |
| iOS | SDK support | Suitable for mobile apps |
| Windows | C++ and Java SDKs | Suitable for desktops and business systems |
| Linux | Partial capability support | Confirm by product and architecture |
| Domestic chips and acceleration cards | Some new features are supported. | A specific compatibility list is required. |
| Private cloud and all-in-one devices | Deliverable | Configure by project |
The fact that an SDK is supported on a certain platform does not mean that there exists a unified, consumer-oriented cloud-native application; personal applications, enterprise components, and industry-specific systems should be considered separately.
Which users are it suitable for
- Application developers who need voice, OCR, and semantic capabilities.
- Companies that develop customer service, outbound calling, and quality control systems.
- Organizations that need the digitization of meetings, transcripts, and documents.
- Financial services related to the handling of licenses, documents, and identity verification.
- Public departments that are building intelligent agent clusters for government service hotlines.
- Large customers that require local and domestically-based deployment.
- Manufacturers that develop products for use in vehicles, home appliances, speakers, and robots.
Main advantages
- It covers a variety of capabilities including voice, semantics, vision, and biometrics.
- Supports cloud, on-premises, private, and all-in-one delivery.
- SDKs, documentation, and examples assist developers in creating prototypes.
- It can be customized for industry-specific terminology, timbre, templates, and models.
- Government and financial agents cover the entire business process.
- The new version of speech recognition takes into account both the semantic capabilities of large models and low latency.
Capacity boundaries
- The accuracy of promotional materials cannot replace actual business validation.
- The free download of the SDK does not mean that its use in production or for commercial purposes is also free.
- The testing applications are available for only half a year and their number is limited.
- The commercial pricing, concurrency limits, SLAs, and refund policies are not publicly available in a unified manner.
- Biometric recognition involves errors and poses a high risk to privacy.
- The rules for complete data preservation and model training must be confirmed through a contract.
- Different capabilities, systems, and chip support ranges vary.
- The two open-source examples do not imply that the platform and the models are open source.
Frequently Asked Questions
Is iFlytek suitable for individual direct use?
It is primarily aimed at developers and organizations. Individuals can experience some of its features or use it in their own applications, but the core value of Lingyun lies in its SDKs, APIs, privacy-focused solutions, and industry-specific solutions.
Is the Lingyun SDK permanently free?
The SDK can be downloaded for free, with a trial period of six months for testing purposes. For full-scale production, cloud-based usage, local resources, and commercial licensing, further arrangements are required.
How many test applications can each developer create?
The current development guidelines specify that a maximum of 10 applications can be created, and the applications that have already been created cannot be deleted. It is necessary to plan the number of applications in advance before selecting capabilities and platforms.
Does Lingyun support private deployment?
Supported. Voice recognition, dual-recording quality inspection, and other enterprise solutions can be implemented in a private or on-premises format; the hardware requirements, number of concurrent tasks, upgrade options, and costs are determined based on each project.
What voice interfaces does Lingyun offer?
The speech recognition page confirms that it supports web requests, Socket, WebService, MRCP, and other methods; the specific protocols depend on the capabilities and the version in use.
Does Lingyun’s GitHub code represent an open-source product?
This is not the case. To date, only two Android example projects under the MIT license have been identified; cloud services, SDKs, model weights, and industry-specific products have not been made open source as a result.
Can voice print recognition serve as the sole means of identity verification?
It is not recommended. Live detection, as well as risk control for devices and business operations, should be employed, and manual review along with other verification methods should be available in case of misjudgments.
Does Jietong Huasheng disclose its prices?
The development and testing rules are made public, but there are no unified publicly available prices for the production APIs, offline SDKs, private versions, all-in-one solutions, and agents; formal quotes are required to obtain such information.
How to add a FAQ Schema
This page already provides FAQs that are visible to users; the site template can use these questions and answers to generate structured FAQ data. Script tags are not inserted directly into the text in order to meet the security requirements for content fields.
- Only use the questions and answers that are visible directly to the user on the page.
- The page object is set to FAQPage.
- Each question is set as Question.
- The answer is written to acceptedAnswer and set as Answer.
- The testing period, quantity, and open-source boundaries must be consistent with the main text.
- Update the FAQ simultaneously when there are changes in the product version or commercial terms.
- After publishing, check for grammar, duplicates, and search engine guidelines.
Summary
Jietong Huasheng’s strength lies in its ability to combine voice technology, semantics, OCR, biometric recognition, and large-model agents to create enterprise-level solutions, offering delivery options via the cloud, on-premises, in a private environment, or as domestically developed systems. It is suitable for projects that require deep integration with industry-specific processes.
When making a selection, do not rely solely on the accuracy shown in demonstrations or on free SDKs; instead, verify the precision, latency, and stability using actual data. Commercial licensing terms, handling of sensitive data, model training, deletion processes, SLAs, and upgrade options must all be specified in the contract.
Guigong Network Security Registration No. 45132202000164