Skip to content
Auto captions

Burned-in captions, coloured per voice, for a whole batch

Upload one video or ten. Clipfer transcribes, recognises who is speaking, places styled captions where you choose, and lets you fix every word before export. The picture, frame rate and colours of the original stay intact.

Sign-ups opening soonSee how it works

Sign-ups not open yet · opening soon

clipfer.com/sous-titres-auto
Choose the caption style and place it on the picture.
01How it works

Five steps, one setting for the whole batch

Upload, configure, transcribe, check, export. The chosen style and placement apply to every video in the batch.

  1. 01

    1. Upload the batch

    Several videos, a whole folder or a ZIP. You can also upload existing SRT or VTT files: matched by name, they skip transcription.

    The step bar of Auto captions: upload and configuration approved.
  2. 02

    2. Choose the style and placement

    Three styles — Neon conversation, Plain, Impact — or your own settings: font, size, outline, shadow, glow, case, one to four words at a time. Then Top, Center, Bottom, or free drag on the reference frame.

    The vertical preview with a sample caption and the placement panel.
  3. 03

    3. Transcribe, check, export

    Deepgram transcribes (language detected automatically, or French / English). You fix the text, the timing, the speaker; filter uncertain lines; search and replace across the batch. Then the service renders the videos.

    The batch summary: videos, length, language, speakers, estimated time.
02What you get

What you get

01

Three ready styles, all adjustable

Neon conversation (three words, colours per voice, glow), Plain (white, crisp outline), Impact (one word at a time, uppercase, yellow and pink). Touch a setting and the style becomes yours.

02

One colour per voice

Deepgram separates the voices, timbre classifies them, and lip movement corrects: the on-screen voice in pink, the off-screen voice in blue. You check each speaker with an audio excerpt.

03

Word-by-word correction

Every caption can be fixed, shifted, hidden; an “Uncertain” filter shows first what deserves a look. The preview is identical to the export.

04

The original intact

The full video is rendered by the service with the frame rate, resolution and colours of the source; captions are drawn by the same engine as the preview.

03Example result

The captioned batch, with its SRT and VTT files

Every video comes out with burned-in captions at the original resolution; the captions also download as SRT and VTT.

Format
MP4 H.264 (10-bit HEVC for an HDR source), AAC audio, resolution and frame rate of the original.
Files
Neutral names (vid_…), ZIP archive with SRT and VTT folders and a mapping report.
Fonts
Nunito, Poppins, Montserrat, Inter, or an uploaded font (TTF, OTF, WOFF, WOFF2).
The vertical preview, a cyan caption placed at the bottom of the picture.
The placement approved for the whole batch, as it will be rendered.
04FAQ

Questions about Auto captions

How many videos at once?

As many as you want: a whole folder or a ZIP. One style and one placement for the whole batch; a batch mixing vertical and horizontal gets a proportional position for each format.

Which languages?

Automatic detection by default (every sentence in its own language), or French / English forced. Speaker detection can be turned off for a single voice.

Can I put a background behind the text?

No: the styles rely on outline, shadow and glow, never on a box. That is what keeps captions readable without hiding the picture.

What if the transcript is wrong?

You fix the text at the check step; the fix is carried over to the timed words. Search and replace fixes a name across the whole batch at once.

What about my existing captions?

Upload your SRT or VTT files with the videos: matched by name, they are burned in without a new transcription.

Your videos captioned, today.

Sign-ups are opening soon.

Sign-ups opening soonI already have an account