GLM-ASR
GLM-ASR-Nano: A robust, open-source speech recognition model with 1.5B parameters
About
GLM-ASR-Nano-2512 is a robust, open-source speech recognition model with 1.5B parameters. Designed for real-world complexity, it outperforms OpenAI Whisper V3 on multiple benchmarks while maintaining a compact size.
Open Source Health
- Stars
- 854
- Forks
- 79
- License
- Apache-2.0
- Last commit
- 7 months ago
Related Categories
Vendor
Z.ai
ChatGLM, GLM-4.5, CogVLM, CodeGeeX, CogView, CogVideoX | CogDL, AMiner | Zhipu.ai (Z.ai)
Quick Links
Open Source
More by Z.ai
Related Products
Whisper Ctranslate2
Whisper command line client compatible with original OpenAI client based on CTranslate2.
Top category match
Voxtype
Voice-to-text with push-to-talk for Wayland compositors
Top category match
pyVideoTrans
Translate the video from one language to another and embed dubbing & subtitles.
Top category match
Kaldi Gstreamer Server
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
Top category match
Vosk Api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Top category match