How to transcribe ESP8266Audio recordings
ESP8266Audio makes the recording — Whipscribe makes it text. Three steps, and the first preview is free with no signup.
The three steps
Get the audio out. Render or bounce the track from ESP8266Audio to WAV or MP3 (usually File → Render/Export/Bounce) — spoken-word stems, podcast edits and voiceovers all transcribe cleanly.
Drop it below. Upload the file right here — no signup for the instant preview. Long, multi-hour files are fine, and video uploads work too (we transcribe the audio track).
Take the text with you. Accurate, speaker-labeled text with timestamps — export TXT, SRT, VTT or DOCX, in 100+ languages, processed on Whipscribe's own private cloud.
About ESP8266Audio
“Arduino library to play MOD, WAV, FLAC, MIDI, RTTTL, OGG/Opus, MP3, and AAC files on I2S DACs or with a software emulated delta-sigma DAC on the ESP8266 and ESP32 and Pico.”
ESP8266Audio facts & alternatives → · GitHub ↗
Frequently asked
Can Whipscribe transcribe sessions and voice tracks from ESP8266Audio?
Yes — any audio or video file ESP8266Audio produces can be uploaded directly. Render or bounce the track from ESP8266Audio to WAV or MP3 (usually File → Render/Export/Bounce) — spoken-word stems, podcast edits and voiceovers all transcribe cleanly.
Do I need to convert the file first?
No. Common audio and video formats upload as-is; video's audio track is transcribed automatically.
How accurate is it?
Clear speech in major languages comes back near-publishable; noisy or heavily accented audio deserves a review pass. The instant preview shows real output before you commit.
What does it cost?
About $2 per audio hour as pay-as-you-go credits that never expire — the first preview is free with no signup.