Speech-Backbones
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
About
Speech-Backbones is a repository focused on speech processing, speech recognition, and speech synthesis technologies. It includes implementations of various models such as Grad-TTS, SPIRAL, and DiffVC, which are designed for tasks like text-to-speech and voice conversion. This resource is intended for researchers and developers interested in advancing speech technology.
Open Source Health
- Stars
- 601
- Forks
- 130
- License
- Not stated
- Last commit
- 3 years ago
Resources & Links
Related Categories
Vendor
HUAWEI Noah's Ark Lab
Working with and contributing to the open source community in data mining, artificial intelligence, and related fields
Quick Links
Open Source
More by HUAWEI Noah's Ark Lab
Related Products
Whisper Ctranslate2
Whisper command line client compatible with original OpenAI client based on CTranslate2.
Top category match
Voxtype
Voice-to-text with push-to-talk for Wayland compositors
Top category match
pyVideoTrans
Translate the video from one language to another and embed dubbing & subtitles.
Top category match
Kaldi Gstreamer Server
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
Top category match
Vosk Api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Top category match