Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
About
Checkout our HuggingFace🤗 collection for the latest open weight checkpoints and demos!
Open Source Health
- Stars
- 18,435
- Forks
- 3,608
- License
- Apache-2.0
- Last commit
- 28 days ago
Related Categories
Vendor
NVIDIA-NeMo
Open-source projects on GitHub: Speech, Guardrails and Switchyard
Quick Links
Open Source
More by NVIDIA-NeMo
Related Products
Whisper Ctranslate2
Whisper command line client compatible with original OpenAI client based on CTranslate2.
Top category match
Voxtype
Voice-to-text with push-to-talk for Wayland compositors
Top category match
pyVideoTrans
Translate the video from one language to another and embed dubbing & subtitles.
Top category match
Kaldi Gstreamer Server
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
Top category match
Vosk Api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Top category match