Generate SRT and VTT subtitles for your podcast
Quick answer: upload an episode and this produces SRT and VTT subtitle files - the two formats most video editors and web players expect - built from the same timestamped transcript the studio's Transcript mode already generates. It is free, works right now, and the captions are AI-generated, so review them before publishing.
Drop your episode here
or click to choose a file - up to 500 MB
MP3 · WAV · M4A · FLAC · OGG · OPUS · MP4 · MOV · MKV · WEBM
Paste a direct link to an audio or video file, or a page a supported downloader can read.
Not every site is supported - private, paywalled or DRM-protected episodes can never be imported this way. If a link fails, download the episode yourself and use Upload File instead.
RSS import is planned so a whole show's feed can be browsed by episode, but it is not built yet. Use Upload File to try Transcript today.
You can keep this tab open and carry on - longer episodes take a few minutes. Refreshing the page reconnects to this job.
Upload an episode above to see its transcript here.
Smart Clips is a preview
Finds moments that could work as standalone clips - hooks, quotes, insights - with a start/end you can adjust. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Reel Maker is a preview
Turns a clip into a captioned 9:16 vertical video with a waveform and safe-zone framing. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Audiograms is a preview
A waveform video over your cover art, sized for the feed - no video footage required. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Summary is a preview
A short and a detailed summary plus key takeaways, generated from the transcript. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Show Notes is a preview
Chapters with timestamps, plus a starting point for links and guest details. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Blog Article is a preview
Restructures the transcript into a readable article outline, not a raw paste of the text. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
Silence Remover is a preview
Finds and previews silent stretches so you can shorten them without cutting natural pauses. This mode has no working backend yet, so nothing is generated or uploaded when you select it - it will switch on here the moment it ships.
No generated assets yet
Once clips, Reels or audiograms can be rendered, finished files will appear here with format, size and expiry, downloadable one at a time or as a ZIP.
Key facts
Transcript vs. SRT vs. VTT
The transcript is plain timed text. SRT is a subtitle file format most video editors accept. VTT is the web-video equivalent. All three come from the same segments.
Format support
Audio: MP3, WAV, M4A, FLAC, OGG, OPUS, AAC. Video: MP4, MOV, MKV, WEBM - only the audio track is read.
File size
Up to 500 MB per upload, the same limit as audio-to-text since both use the same pipeline.
Not supported yet
Speaker diarization and burned-in video captions. Captions are timestamped text you export, not rendered into a video frame.
What this tool does
Transcribes an uploaded podcast episode and exports the result as SRT and VTT subtitle files, plus the transcript itself.
Accepted input
An audio or video file up to 500 MB, in the formats listed above.
Expected output
Downloadable SRT and VTT files with numbered, timestamped caption blocks, plus the same TXT export and on-page transcript Transcript mode provides.
How to use it
Upload
Drag your episode's audio or video file onto the upload area, or click to browse.
Analyze
Confirm you have the rights to the content, then press Analyze My Podcast.
Review and export
Check the captions against the audio, then download SRT or VTT.
A realistic example
A 25-minute solo episode exports as an SRT file with roughly 140-170 numbered caption blocks, each a few seconds long with start and end timestamps - ready to load into a video editor to burn captions into a video export, or attach directly to a web player as a VTT track.
Common use cases
Video editors
Import SRT directly for caption timing on a video cut of the episode.
Web publishers
Attach the VTT file as a caption track on a hosted video player.
Accessibility
Give viewers accurate, timestamped captions instead of none at all.
Repurposing
The same transcript and captions Reel Maker and Subtitles both draw from.
Limitations
- AI-generated captions require review - names, technical terms and overlapping speech are the most common mistakes.
- No speaker diarization yet - captions are not labelled by speaker.
- No burned-in video captions yet - you load the SRT/VTT file into your own player or editor.
Privacy and retention
- Uploaded files are held only long enough to transcribe them, then removed by routine cleanup.
- Nothing is published or shared - the captions are shown only to you.
- Only process content you own or have permission to use.
Related podcast tools
Frequently asked questions
Is the Podcast Subtitle Generator free?
Yes, and it works today - upload an episode and download SRT and VTT subtitle files with no sign-up.
What is the difference between the transcript, SRT and VTT?
The transcript is plain text split into timed segments, shown on the page. SRT (SubRip) is a subtitle file format most video editors and platforms accept, with numbered, timestamped caption blocks. WebVTT (VTT) is the equivalent format built for web video players. All three come from the same underlying segments - SRT and VTT just package them for playback software.
Do AI-generated subtitles need review?
Yes. Automatic speech recognition makes mistakes, particularly with names, technical terms, overlapping speech and background noise. Read through the captions before publishing them, the same as you would for any auto-generated subtitle.
Does it label who is speaking?
No, not currently. Captions are timestamped but not attributed to a speaker - speaker diarization is not available yet.
Will the captions be burned into a video?
No, not currently. This produces SRT and VTT files you load into a video player or editor yourself - burned-in (rendered directly into the video frame) captions are not available yet.
How is this different from audio-to-text or the Transcript mode?
All three share the same transcription engine and the same job. This page frames it around exporting captions - SRT and VTT are the primary outputs here - while Podcast Transcript Generator frames the identical result as a readable transcript and audio-to-text frames it as general-purpose transcription for any recording.
Last reviewed: 2026-09-08.
Get your subtitles now
Upload your episode above and download SRT or VTT once transcription finishes.
Back to the workbenchRate this tool
Was this tool useful? Your feedback helps us improve it.