SenseVoice
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
Category
Speech-to-text
Type
Open source
Platform
C project
Pricing
Free / open source
License
MIT
GitHub stars
9,068
Last updated
2026-08-12
Need the recording as text?
Whatever you record or edit with SenseVoice, Whipscribe turns it into an accurate, speaker-labeled transcript — 100+ languages, SRT/VTT/DOCX export, private self-hosted Whisper. About $2 per audio hour, 30 minutes free daily.
Transcribe a file →