How to transcribe › avatars4all

How to transcribe avatars4all recordings

avatars4all makes the recording — Whipscribe makes it text. Three steps, and the first preview is free with no signup.

The three steps

1

Get the audio out. Copy the clip off avatars4all (any common video format) and upload it as-is.

2

Drop it below. Upload the file right here — no signup for the instant preview. Long, multi-hour files are fine, and video uploads work too (we transcribe the audio track).

3

Take the text with you. Accurate, speaker-labeled text with timestamps — export TXT, SRT, VTT or DOCX, in 100+ languages, processed on Whipscribe's own private cloud.

About avatars4all

“Live real-time avatars from your webcam in the browser. No dedicated hardware or software installation needed. A pure Google Colab wrapper for live First-order-motion-model, aka Avatarify in the browser. And other Colabs providing an accessible interface for using FOMM, Wav2Lip and Liquid-warping-GA.”

GitHub stars
373
Built in
Jupyter Notebook
Category
Cameras Capture

avatars4all facts & alternatives →  ·  GitHub ↗

Frequently asked

Can Whipscribe transcribe footage captured with avatars4all?

Yes — any audio or video file avatars4all produces can be uploaded directly. Copy the clip off avatars4all (any common video format) and upload it as-is.

Do I need to convert the file first?

No. Common audio and video formats upload as-is; video's audio track is transcribed automatically.

How accurate is it?

Clear speech in major languages comes back near-publishable; noisy or heavily accented audio deserves a review pass. The instant preview shows real output before you commit.

What does it cost?

About $2 per audio hour as pay-as-you-go credits that never expire — the first preview is free with no signup.