Audio in. Captions out.
Drop in a file, shape it exactly how you want, and export frame-accurate captions ready for Final Cut, Premiere, DaVinci, or CapCut.
Drop your file here
or click to browse
M4A ยท MP3 ยท WAV ยท OGG ยท WEBM ยท MP4
Splits evenly into chunks of N words.
Holds each caption until the next one starts, so there is no blank frame between lines. Longer pauses are left alone, and the last caption stays up for the same amount after its final word.
CaptionsThis transcribes your recording and gives every word its own timestamp, so the subtitles you get back are editable rather than fixed. Change where a line breaks, split a long cue in two, or retype a misheard name, and the timing follows, because it is anchored to the audio and not interpolated from character counts.
Export SRT when the captions need to travel: YouTube, Premiere, DaVinci Resolve, a client. Export FCPXML when you are finishing in Final Cut Pro and want titles that land on the right frame with your styling already applied.
Yes. You can transcribe a clip up to 5 minutes without an account to see the timing for yourself. A free account unlocks SRT and FCPXML export and gives you 15 minutes and 3 transcriptions every month. After that, credit packs are one-time purchases that never expire. There is no subscription.
Audio: MP3, M4A, WAV, OGG, and WEBM. Video: MP4, MOV, and WEBM. Video files are decoded to audio in your browser before upload, so a large video does not need to be uploaded in full.
SRT is a plain list of timecoded text cues that works in almost every player and editor, but carries no styling. FCPXML is Final Cut Pro's interchange format: subtitles arrive as editable title clips with fonts, positioning, and per-word styling intact, at your project's exact frame rate.
Transcription runs on Whisper large-v3, then a separate forced-alignment pass locates each individual word in the audio. That means every word has its own start and end time, so you can split or join subtitle lines without the second half drifting out of sync.
Yes. Pick a target language and the transcript is translated segment by segment, keeping the original timing boundaries. You can export the translated track as SRT or FCPXML like any other.
Uploads are capped at 25 MB. Because video is converted to compressed audio in your browser first, that limit covers roughly two hours of speech from a typical recording.
Yes. Transcribing a short clip works without one, but exporting the SRT or FCPXML file requires a free account. Signing in also gives you 15 minutes and 3 transcriptions each month.
No. Audio is sent to the transcription service, processed, and returned as text. It is not stored. Your edits are saved only in your own browser's local storage.