← All tools
Subtitle Generator
Any recording with speech → ready-to-load SRT and VTT subtitle files, plus a plain-text transcript.
50 credits free every month
No card needed
Private files
Add your file
Upload after your free sign-upMP3, M4A, WAV, MP4, MOV, and more · up to 500 MB, one file
Sign up to continue
Secure upload
EXAMPLE
What you get
check_circleAn SRT file that loads in every major editor and desktop player
check_circleA WebVTT file for HTML5 video and web players
check_circleA plain-text transcript of everything spoken
check_circleStandard two-line cues, at most 42 characters per line
check_circleA clear message if the file has no audio or no speech
How it works
1Upload one recordingDrop in an audio or video file up to 500 MB: a talking-head video, a screen recording, a podcast episode, a lecture.
2We transcribe and time the cuesThe speech is transcribed on our own servers in silence-aligned chunks, then shaped into standard two-line cues of up to 42 characters per line.
3Download your subtitle filesAn SRT, a WebVTT, and a plain-text transcript, all named after your upload and ready to load into your editor or player.
Who uses this tool
VIDEO PRODUCTION
Subtitles for the edit
SOCIAL MEDIA
Captions before you post
EDUCATION
Course videos students can follow
PODCASTING
Captions and show notes in one run
MARKETING
Demo videos that work on mute
ACCESSIBILITY
Captions for viewers who need them
SECURITY
Built for sensitive documents
Bank statements, medical files, case records. Security isn't a feature we added - it's the foundation.
lock
Encrypted in transit
Encrypted in transit over TLS 1.2+, and stored in access-controlled, encrypted object storage. Files are protected the moment they leave your browser.
admin_panel_settings
Your files stay yours
Workspaces are isolated per account. Role-based access shows teammates only what they need.
auto_delete
Deleted, not stored
Files are deleted after processing. Everything runs on our own hardware and is never sent to an outside AI service, so your data is never used to train models.
dnsProcessed on our own hardwareauto_deleteDeleted after processinglockEncrypted in transit
Questions
Common audio and video: MP3, M4A, WAV, AAC, FLAC, OGG, OPUS, WMA, and video like MP4, MOV, M4V, WEBM, MKV, and AVI. For video we extract the audio track first, then transcribe the speech. One file per job, up to 500 MB and about 60 minutes of audio - longer recordings are rejected up front with a clear message, so split them first.
Cue timing comes from silence-aligned speech chunks, which keeps subtitles in sync with the dialog for normal viewing. It is not word-level karaoke timing, so individual words are not highlighted as they are spoken.
No. Cues hold the spoken words shaped into standard two-line blocks of up to 42 characters per line, with no speaker names. For a talking-head video or narration that is exactly what you want; for a multi-person panel you would need to add names yourself.
Three files: SRT (the format editors and desktop players expect), WebVTT (the format HTML5 video and web players expect), and a plain-text transcript with the timing stripped. If you only need one format, the Video to SRT and Video to VTT tools produce a single file.
Subtitling is 5 credits per minute of audio. Because the length is not known until the file is processed, the upfront hold is estimated from file size at a typical bitrate for the format (about 1 MB per minute for MP3 audio, more for WAV, FLAC, or video), so a typical minute works out to about 5 credits. The minimum is 5 credits.
Yes. Files run on our own servers, are never used to train any model, and are auto-deleted 7 days after the job finishes.
Related tools
Video → SRT
Video to SRT
One video → an SRT subtitle file that drops straight into editors, players, and upload dashboards.
Video → VTT
Video to VTT
One video → a WebVTT caption file ready for HTML5 video, web players, and streaming platforms.
SRT ↔ VTT ↔ TXT
Subtitle Converter
SRT and VTT caption files → the format your player wants: SRT for editors and TVs, VTT for web players, TXT to read the dialog.
Audio → text · DOCX
Audio to Text
Any audio recording → a typed transcript you can search, quote, and edit, delivered as TXT and DOCX.
More in Media & video
Video → MP3 · M4A · WAV
Video to Audio
Any common video → a clean MP3, M4A, or WAV of just the audio track, with the picture dropped.
Audio → MP3 · WAV
Audio Converter
Any mix of MP3, WAV, M4A, AAC, FLAC, OGG, OPUS, or WMA → one format you pick, a whole batch in a single upload.
Audio → smaller MP3
Compress Audio
Heavy recordings, voice memos, and podcast files → small MP3s that fit email limits and archives, a whole batch at once.
Audio → trimmed clip
Trim Audio
A long recording → just the section you need: enter a start and end time and keep only that clip, in the same format you uploaded.
Many clips → one
Merge Audio
Voice memos, episode parts, or interview sides → one continuous audio file, joined in the order you upload them.
Audio → -16 LUFS
Normalize Loudness
Quiet voice memos and loud music bounces → every file at -16 LUFS, the loudness podcast and streaming platforms target.
Try it free. 50 credits every month.
No card needed. Files are deleted after processing.