awesome-speech-recognition-speech-synthesis-papers alternatives (speech-to-text)
awesome-speech-recognition-speech-synthesis-papers: Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC) Below are the closest alternatives in the same category — every entry with a live profile in the hub.
Need the recording as text?
transformers🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.whisper.cppPort of OpenAI's Whisper model in C/C++voiceboxThe open-source AI voice studio. Clone, dictate, create.HandyA free, open source, and extensible speech-to-text application that works completely offline.meetilyPrivacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloudllamafileDistribute and run LLMs with a single file.faster-whisperFaster Whisper transcription with CTranslate2whisperXWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)buzzBuzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.screenpipeYC (S26) | Record your screen 24/7 and plug into your agents. Local, private, secure. Connect to OpenClaw, Hermes agent and 100+ apps
Whatever you record or edit with awesome-speech-recognition-speech-synthesis-papers, Whipscribe turns it into an accurate, speaker-labeled transcript — 100+ languages, SRT/VTT/DOCX export, private self-hosted Whisper. About $2 per audio hour, 30 minutes free daily.
Transcribe a file →← Back to awesome-speech-recognition-speech-synthesis-papers · All speech-to-text →