StylesHow it worksFeaturesUse casesPricing
Home/Use cases/Add Captions to Video Podcast Clips

Add captions to a video podcast clip.

Start from the clip you already cut from a recorded episode. Generate the captions, correct the names and the moments where both people talk at once, keep the text clear of faces, then export the MP4.

Drop a video here

MP4, MOV, WebM or MKV · up to 500 MB · 10 minutes

No signup. No watermark. Guest projects stay editable for 24 hours, or 7 days once you sign in.

What matters

Two people talking is harder to caption than one

Hosts interrupt, laugh over each other and say guests' names a speech model has never heard. The review step, not the generation, decides whether a podcast clip reads well.

Input
A video file
Audio-only upload
Not supported
Speaker labels
Not added
Output
Captioned MP4, SRT, VTT, TXT

How it works

  1. 01

    Cut the clip first

    Export the moment you want to share from your editing or recording software as MP4, MOV, WebM or MKV. This page does not find highlights or trim an episode.

  2. 02

    Generate, then read every line

    Correct names, show titles and jargon in the script panel. Listen again wherever both speakers overlap.

  3. 03

    Place captions around both faces

    Drag and resize the caption block on the frame so it sits below or between the speakers rather than across a mouth.

  4. 04

    Export for where it is going

    Download the burned-in MP4 for social feeds, or SRT or VTT when a player draws its own captions.

Built for the correction after generation.

Fix crosstalk by hand

When two people speak at once, words from both can land in one caption or be misheard. Correct the words, then drag a caption's start or end on the timeline so it lines up with who is audible.

Correct guest names where they appear

The script panel lists the whole transcript in order. Read it through, fix each spelling of a name or brand, and the timing of the surrounding words is kept.

Frame-aware placement

Two-shot and split-screen podcast layouts leave different space free. Move the caption block on your actual footage rather than trusting a fixed bottom position.

Readable without sound

Word-level timing lets styles follow the words being spoken. Every style can be previewed; Clean Minimal and Documentary export on Free.

The same clip for several feeds

Burn captions into one MP4 and post it where you publish clips. The platform guides cover each app's interface and canvas.

What this does, and what podcast suites also do

Some podcast tools start from a full episode, find clips and label speakers. This workflow starts later, from a clip you already chose.

You choose the moment

No automatic highlight detection or trimming. That keeps the edit in your hands and means the clip has to exist before you upload.

Speakers are not identified

Captions are not split or labelled by voice. If you need names on screen for each speaker, that is not available in this editor.

Video in, video out

An audio-only MP3 or WAV cannot be uploaded, and there is no RSS or episode-link import. Export a video clip from your recording first.

Correction is the core

The time goes into reading the transcript on the real footage, which is where overlapping speech and names are caught.

Limits that matter for podcast footage.

Accepted files
MP4, MOV, WebM, MKV
Audio-only files
Not accepted
Free length and size
10 minutes, 500 MB
Starter length and size
120 minutes, 5 GB
Longest video on any plan
120 minutes
Free video projects
3 per month
Free styles to export
Clean Minimal, Documentary
Account to download
Google or email sign-in
Speaker labels or diarization
Not available
Watermark
None

Not the right tool for these.

Worth knowing before you upload an episode.

  • Captioning a whole long episode on Free

    Free accepts videos up to 10 minutes. A full episode needs Starter, which accepts up to 120 minutes and 5 GB; anything longer has to be split first.

  • Finding clips or posting them for you

    There is no highlight detection, automatic trimming, social account connection or scheduled posting. The output is a file for you to publish.

  • Transcripts with named speakers or translations

    Captions are in the spoken language and are not labelled by speaker. Use a transcription service built for interviews when a named, speaker-separated transcript is the deliverable.

Questions before upload

Can I upload an MP3 or WAV podcast episode?+

No. The uploader accepts MP4, MOV, WebM and MKV video files. Export a video clip from your recording or editor first.

Will it label who is speaking?+

No. Captions are generated from the speech without identifying speakers. Correct the words and timing yourself where two voices overlap.

Can it pick the best moments from my episode?+

No. It captions a clip you have already cut. Choose and trim the moment in your editing software, then upload that clip.

How long can the clip be?+

Up to 10 minutes and 500 MB on Free, and up to 120 minutes and 5 GB on Starter.

Does a revised export use another free slot?+

No. Free counts 3 distinct video projects a month. Downloading a corrected version of the same clip that month uses the slot it already has.

How do I keep captions off the speakers' faces?+

Drag the caption block on the video and resize it until it sits in free space, usually below both speakers or between them in a split-screen layout. Check the result in the preview before exporting.

Add captions to Instagram Reels →Add captions to TikTok →Caption YouTube Shorts →Burned-in vs closed captions →Extract a transcript from a VTT file →

Generate, edit, style, and export subtitles without turning somebody else's logo into part of your video.

Product

  • AI subtitle generator
  • Video caption generator
  • Add subtitles to video
  • YouTube subtitles

Resources

  • Free subtitle tools
  • Help center
  • Blog
  • Caption guides
  • Alternatives and comparisons
  • Features
  • Use cases
  • Pricing
  • Getting started
  • Supported formats
  • Español (preview)

Legal

  • Privacy
  • Terms
  • Refunds