Speech To Speech
Machine Learning Model TrainingML Frameworks

Speech To Speech

Low-latency speech-to-speech pipeline

Open Source

About

A low-latency, fully modular voice-agent pipeline: VAD -> STT -> LLM -> TTS, exposed through the core OpenAI Realtime GA event set over WebSocket and WebRTC. Every component is swappable. The LLM slot speaks OpenAI-compatible protocols, so you can point it at a hosted provider, at HF Inference Providers, or at a vLLM or llama.cpp server on your own hardware for a fully local, fully open stack.

Open Source Health

Not enough history
Stars
13,184
Forks
1,664
License
Apache-2.0
Last commit
1 months ago
Python

Related Categories