docker-whisper
Docker image for a self-hosted Whisper speech-to-text server with speaker diarization and OpenAI-compatible transcription and translation APIs. Powered by faster-whisper. Supports all Whisper models, NVIDIA GPU (CUDA) acceleration, JSON/SRT/VTT output, SSE streaming, offline mode, and multi-arch (am
Category
Speech-to-text
Type
Open source
Platform
Python project
Pricing
Free / open source
GitHub stars
92
Last updated
2026-08-08
Need the recording as text?
Whatever you record or edit with docker-whisper, Whipscribe turns it into an accurate, speaker-labeled transcript — 100+ languages, SRT/VTT/DOCX export, private self-hosted Whisper. About $2 per audio hour, 30 minutes free daily.
Transcribe a file →