CosyVoice
CosyVoice is a family of open-source multilingual neural text-to-speech (TTS) and voice cloning models developed by the Tongyi Speech Lab (Tongyi SpeechTeam) at Alibaba Group and released under the Apache 2.0…
Explore Speech & Audio AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Speech & Audio AI.
Showing 1-5 of 5 articles
CosyVoice is a family of open-source multilingual neural text-to-speech (TTS) and voice cloning models developed by the Tongyi Speech Lab (Tongyi SpeechTeam) at Alibaba Group and released under the Apache 2.0…
GLM-4-Voice is an open-weights end-to-end speech-to-speech large language model released in October 2024 by Zhipu AI together with the Knowledge Engineering Group (KEG) at Tsinghua University.
Kai-Fu Lee (Chinese: 李開復; born December 3, 1961) is a Taiwanese-American computer scientist, venture capitalist, and author who is the founder and CEO of 01.AI, the Beijing large language model startup behind…
Qwen2-Audio is an audio-language model developed by the Qwen team at Alibaba Cloud, released in August 2024 .
iFlytek (Chinese: 科大讯飞, formally Anhui USTC iFlytek Co., Ltd.) is a partially state-owned Chinese artificial intelligence and information technology company, founded on December 30, 1999 in Hefei, Anhui…