Deepvoice3 Pytorch
PyTorch implementation of convolutional neural networks-based text-to-speech synthesis models
About
DeepVoice3_pytorch provides a framework for text-to-speech synthesis utilizing convolutional sequence-to-sequence models with attention mechanisms. It supports both multi-speaker and single-speaker configurations, making it suitable for various speech synthesis applications. The software includes pre-trained models and audio samples, as well as a language-dependent frontend text processor for English and Japanese, catering to developers and researchers in the field of speech processing.
Open Source Health
- Stars
- 1,975
- Forks
- 481
- License
- Not stated
- Last commit
- 3 years ago
Resources & Links
Related Categories
Vendor
Ryuichi Yamamoto
Speech Synthesis, Voice Conversion, Machine Learning, Singing Voice Synthesis
Quick Links
Open Source
Related Products
Rclip
Semantic photo search for the command line
Top category match
ChineseNER
A neural network model for Chinese named entity recognition
Top category match
Head Pose Estimation
Realtime human head pose estimation with ONNXRuntime and OpenCV.
Top category match
Video Subtitle Extractor
视频硬字幕提取,生成srt文件。无需申请第三方API,本地实现文本识别。基于深度学习的视频字幕提取框架,包含字幕区域检测、字幕内容提取。A GUI tool for extracting hard-coded subtitle (hardsub) from videos and generating srt files.
Top category match
U 2 Net
The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."
Top category match