CosyVoice
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
About
Fun-CosyVoice 3.0 is an advanced text-to-speech (TTS) system based on large language models (LLM), surpassing its predecessor (CosyVoice 2.0) in content consistency, speaker similarity, and prosody naturalness. It is designed for zero-shot multilingual speech synthesis in the wild.
Open Source Health
- Stars
- 23,597
- Forks
- 2,683
- License
- Apache-2.0
- Last commit
- 5 months ago
Related Categories
Vendor
QwenAudio
Open-source speech and audio language models from the QwenAudio Team
Quick Links
Open Source
More by QwenAudio
Related Products
Repomix
Repomix is a powerful tool that packs your entire repository into a single, AI-friendly file. Perfect for when you need to feed your codebase to Large Language Models (LLMs) or other AI tools like Claude, ChatGPT, DeepSeek, Perplexity, Gemini, Gemma, Llama, Grok, and more.
Top category match
Quivr
Opiniated RAG for integrating GenAI in your apps 🧠 Focus on your product rather than the RAG. Easy integration in existing products with customisation! Any LLM: GPT4, Groq, Llama. Any Vectorstore: PGVector, Faiss. Any Files. Anyway you want.
Top category match
Chatbot
A full-featured, hackable Next.js AI chatbot built by Vercel
Top category match
WikiChat
WikiChat is an improved RAG. It stops the hallucination of large language models by retrieving data from a corpus.
Top category match
Tribe
Low code tool to rapidly build and coordinate multi-agent teams
Top category match