SenseVoice
TranscriptionVoice, Speech & Audio AI

SenseVoice

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

Open Source

About

SenseVoice is a speech foundation model with multiple speech understanding capabilities, including automatic speech recognition (ASR), spoken language identification (LID), speech emotion recognition (SER), and audio event detection (AED).

Open Source Health

Not enough history
Stars
9,297
Forks
823
License
MIT
Last commit
29 days ago
C

Related Categories