← All tools
Video to VTT
One video → a WebVTT caption file ready for HTML5 video, web players, and streaming platforms.
50 credits free every month
No card needed
Private files
Add your file
Upload after your free sign-upMP4, MOV, MKV, WEBM, AVI, or M4V · up to 500 MB, one file
Sign up to continue
Secure upload
EXAMPLE
What you get
check_circleOne .vtt file named after your video, with the WEBVTT header in place
check_circleStandard two-line cues, at most 42 characters per line
check_circleTiming that tracks the dialog, ready for the HTML5 track element
check_circlePlain LF line endings that strict parsers accept
check_circleA clear message if the video has no audio or no speech
How it works
1Upload one videoDrop in an MP4, MOV, M4V, WEBM, MKV, or AVI file up to 500 MB.
2We transcribe and time the cuesThe audio is extracted and transcribed on our own servers in silence-aligned chunks, then shaped into standard two-line cues of up to 42 characters per line.
3Download the VTTOne .vtt file named after your video, ready for the HTML5 track element, your web player, or a platform caption upload.
Who uses this tool
WEB DEVELOPMENT
Captions for the track element
E-LEARNING
Course videos with captions
SAAS & PRODUCT
Captioned demos in your docs
MEDIA & STREAMING
A caption file for your player
ACCESSIBILITY
Meet caption requirements on the web
MARKETING
Landing-page videos that autoplay muted
SECURITY
Built for sensitive documents
Bank statements, medical files, case records. Security isn't a feature we added - it's the foundation.
lock
Encrypted in transit
Encrypted in transit over TLS 1.2+, and stored in access-controlled, encrypted object storage. Files are protected the moment they leave your browser.
admin_panel_settings
Your files stay yours
Workspaces are isolated per account. Role-based access shows teammates only what they need.
auto_delete
Deleted, not stored
Files are deleted after processing. Everything runs on our own hardware and is never sent to an outside AI service, so your data is never used to train models.
dnsProcessed on our own hardwareauto_deleteDeleted after processinglockEncrypted in transit
Questions
MP4, MOV, M4V, WEBM, MKV, or AVI, one video per job, up to 500 MB and about 60 minutes of audio - longer recordings are rejected up front with a clear message. We extract the audio track and transcribe the speech; if the file has no audio or no recognizable speech, the job stops and tells you.
Host the .vtt next to your video and reference it from a track element inside your HTML5 video tag, or upload it to your player or platform as the caption track. The file starts with the required WEBVTT header, so it works as-is.
Cue timing comes from silence-aligned speech chunks, which keeps captions in sync with the dialog for normal viewing. It is not word-level karaoke timing, so individual words are not highlighted as they are spoken.
No. Cues hold the spoken words shaped into standard two-line blocks of up to 42 characters per line, with no speaker names.
Subtitling is 5 credits per minute of audio. Because the length is not known until the file is processed, the upfront hold is estimated from file size at a typical video bitrate (about 8 MB per minute), so a typical minute works out to about 5 credits. The minimum is 5 credits.
Yes. Videos run on our own servers, are never used to train any model, and are auto-deleted 7 days after the job finishes.
Related tools
Audio · video → SRT · VTT
Subtitle Generator
Any recording with speech → ready-to-load SRT and VTT subtitle files, plus a plain-text transcript.
Video → SRT
Video to SRT
One video → an SRT subtitle file that drops straight into editors, players, and upload dashboards.
SRT ↔ VTT ↔ TXT
Subtitle Converter
SRT and VTT caption files → the format your player wants: SRT for editors and TVs, VTT for web players, TXT to read the dialog.
More in Media & video
Video → MP3 · M4A · WAV
Video to Audio
Any common video → a clean MP3, M4A, or WAV of just the audio track, with the picture dropped.
Audio → MP3 · WAV
Audio Converter
Any mix of MP3, WAV, M4A, AAC, FLAC, OGG, OPUS, or WMA → one format you pick, a whole batch in a single upload.
Audio → smaller MP3
Compress Audio
Heavy recordings, voice memos, and podcast files → small MP3s that fit email limits and archives, a whole batch at once.
Audio → trimmed clip
Trim Audio
A long recording → just the section you need: enter a start and end time and keep only that clip, in the same format you uploaded.
Many clips → one
Merge Audio
Voice memos, episode parts, or interview sides → one continuous audio file, joined in the order you upload them.
Audio → -16 LUFS
Normalize Loudness
Quiet voice memos and loud music bounces → every file at -16 LUFS, the loudness podcast and streaming platforms target.
Try it free. 50 credits every month.
No card needed. Files are deleted after processing.