How to transcribe Modality-toolkit recordings
Modality-toolkit makes the recording — Whipscribe makes it text. Three steps, and the first preview is free with no signup.
The three steps
Get the audio out. Render or bounce the track from Modality-toolkit to WAV or MP3 (usually File → Render/Export/Bounce) — spoken-word stems, podcast edits and voiceovers all transcribe cleanly.
Drop it below. Upload the file right here — no signup for the instant preview. Long, multi-hour files are fine, and video uploads work too (we transcribe the audio track).
Take the text with you. Accurate, speaker-labeled text with timestamps — export TXT, SRT, VTT or DOCX, in 100+ languages, processed on Whipscribe's own private cloud.
About Modality-toolkit
“A SuperCollider toolkit to simplify the creation of personal (electronic) instruments utilising hardware and software controllers of any kind.”
Modality-toolkit facts & alternatives → · GitHub ↗ · Website ↗
Frequently asked
Can Whipscribe transcribe sessions and voice tracks from Modality-toolkit?
Yes — any audio or video file Modality-toolkit produces can be uploaded directly. Render or bounce the track from Modality-toolkit to WAV or MP3 (usually File → Render/Export/Bounce) — spoken-word stems, podcast edits and voiceovers all transcribe cleanly.
Do I need to convert the file first?
No. Common audio and video formats upload as-is; video's audio track is transcribed automatically.
How accurate is it?
Clear speech in major languages comes back near-publishable; noisy or heavily accented audio deserves a review pass. The instant preview shows real output before you commit.
What does it cost?
About $2 per audio hour as pay-as-you-go credits that never expire — the first preview is free with no signup.