OpenAI Whisper ASR Webservice API
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
Build local voice agents with open-source models
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.