Marketplace

By Use Case Marketplace

Compare By Use Case products across curated subcategories and trusted providers

Products
15,413
Subcategories
698

Selected subcategory

Transcription

Compare By Use Case products across curated subcategories and trusted providers

Explore By Use Case with structured category paths, practical filters, and independent product data

Showing 1–12 of 76

Active filtersOpen sourceClear all

Whisper Ctranslate2

by Softcatalà

Whisper command line client compatible with original OpenAI client based on CTranslate2.

TranscriptionVoice, Speech & Audio AIOpen Source

Voxtype

by Pete Jackson

Voice-to-text with push-to-talk for Wayland compositors

TranscriptionVoice, Speech & Audio AIOpen Source

pyVideoTrans

by okmyworld

Translate the video from one language to another and embed dubbing & subtitles.

TranscriptionVoice, Speech & Audio AIOpen Source

Kaldi Gstreamer Server

by Tanel Alumäe

Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.

TranscriptionVoice, Speech & Audio AIOpen Source

Vosk Api

by Alpha Cephei

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

TranscriptionVoice, Speech & Audio AIOpen Source

🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

TranscriptionVoice, Speech & Audio AIOpen Source

Speech

by NVIDIA-NeMo

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

TranscriptionVoice, Speech & Audio AIOpen Source

Handy

by CJ Pais

A free, open source, and extensible speech-to-text application that works completely offline.

TranscriptionVoice, Speech & Audio AIOpen Source

VideoChat

by Weiheng Chi

实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.

TranscriptionVoice, Speech & Audio AIOpen Source

Whisper

by Openai

Robust Speech Recognition via Large-Scale Weak Supervision

TranscriptionVoice, Speech & Audio AIOpen Source

FunASR

by ModelScope

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

TranscriptionVoice, Speech & Audio AIOpen Source

Speech-Backbones

by HUAWEI Noah's Ark Lab

This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.

TranscriptionVoice, Speech & Audio AIOpen Source

Frequently asked questions