Video to SRT
Turn any video into a ready-to-use .srt subtitle file, generated by AI.
More accurate, slower — recommended
The first run downloads the speech model to your browser (cached after that). Best on desktop — mobile devices may be slower.
Drop a video or audio file to transcribe
or click to browse
How it works
Drop a video or audio file in. WavyVid decodes the audio track right in your browser.
An AI speech-recognition model (Whisper, running fully on-device via WebAssembly/WebGPU) transcribes the audio into timestamped text.
Edit the transcript if needed, then export as .srt/.vtt, or burn the captions directly into your video.
SRT is the most widely supported subtitle format — YouTube, Vimeo, and virtually every video editor accepts it. This tool generates a properly timestamped SRT straight from your video's audio track, so you can drop it into your editor or upload it alongside your video without manually timing a single line.
Frequently asked questions
What is an SRT file?
A plain-text subtitle format listing cue numbers, start/end timestamps, and text — supported by nearly every video platform and editor.
Can I edit the timestamps or text before exporting?
Yes — the transcript is editable in the browser before you download the SRT file.
Does it work on audio-only files too?
Yes, though for an audio file the SRT is more useful as a transcript-with-timing than as video captions.
Which languages are supported?
The underlying Whisper model supports many languages, though accuracy is strongest for English in this browser-optimized configuration.