sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket ser
Category
Speech-to-text
Type
Open source
Platform
C++ project
Pricing
Free / open source
License
Apache-2.0
GitHub stars
14,165
Last updated
2026-08-13
Need the recording as text?
Whatever you record or edit with sherpa-onnx, Whipscribe turns it into an accurate, speaker-labeled transcript — 100+ languages, SRT/VTT/DOCX export, private self-hosted Whisper. About $2 per audio hour, 30 minutes free daily.
Transcribe a file →