TIX released the latest-generation voice recognition large model

On 23 September, the latest generation of the mega-map of voice recognition was officially launched by TIX. It is understood that the model continues and deepens the technical capability of the Spark-Audio-1.0-Preview speech base model and achieves a significant improvement in overall speech recognition through key technologies such as enhanced self-regression synergy with LLM, a combination of Chinese and English text and acoustic enhancement, and dynamic context injection. At the same time, the reasoned cost of Spark-ASR-2.0 increased only 10% compared to the previous generation model. Spark-ASR-2.0 will start a step-by-step go-live feed from September 24th and will provide API services through the open platform of the e-mail, which will then be applied over time to products such as e-mail-AI glasses, e-mail smarts, e-mails. (36kr)

Search