Fonti Studio

HomeTools › SRT to text

SRT to text: extract a plain transcript from subtitles

Turn an SRT or WebVTT file into plain text: timing, cue numbers and formatting tags removed, one line per cue. For scripts, translation memories, dialogue lists and search. Free, in your browser.

Runs in your browser · your file never leaves your machine

Convert format

Timing is preserved to the millisecond and text passes through untouched. Both formats store real clock time, so no frame rate is involved.

Frame-rate repair

For subtitles that drift steadily: in sync at the start, seconds off by the end. Timestamps are rescaled by the exact ratio between the two rates.

Constant shift

For subtitles that are early or late by the same amount all the way through. Positive moves cues later, negative moves them earlier.

Applied after the frame-rate repair when both are set.

How findings are graded
BlockingBreaks the spec with no legitimate exception. A platform QC bounces the file.
ReviewExceeds a limit for a reason a person has to judge: a two-speaker cue, a card over dialogue, or dialogue the original cuts just as fast.
RepairableFixed mechanically when the file is written, not by an editor.

Checked against the public Netflix Timed Text limits: 42 characters per line, 2 lines, reading speed, 5/6 s to 7 s durations, 2-frame gaps, overlaps. Formatting tags are excluded from the counts, and cues of 12 characters or fewer are exempt from reading speed because they are read at a glance.

What you get when you convert SRT to text

A subtitle file is a transcript wrapped in timing. Strip the cue numbers, the timecodes and the formatting tags and what is left is the dialogue, in order, one cue per line. That is what this tool writes: plain UTF-8 text you can paste into a script, a translation memory, a search index, a dialogue list for a sales deck, or a document for a reader who only needs the words. Nothing is summarised or reflowed; if a sentence was split across two cues it stays on two lines, which is usually what you want when the text has to be traced back to the picture.

What is kept and what is dropped

Kept: every line of dialogue, speaker dashes, punctuation, and the cue order. Dropped: cue numbers, start and end times, italics and other tags, positioning codes, and empty cues. Lines inside one cue are joined with a space, so a wrapped two-line subtitle becomes one line of text; a two-speaker cue keeps one speaker per line. Both SRT and WebVTT are accepted as input; the WebVTT header and cue settings are discarded with the timing.

When you need the opposite

Going from a plain transcript back to subtitles is not a conversion, it is spotting: every line needs an in and out time against the picture, and the result has to respect line length, reading speed and shot changes. The formats guide explains why, and what a buyer expects a real subtitle file to contain. If what you have is a finished subtitle file that needs a different container, use the SRT to VTT converter; if it needs to be checked before delivery, the validator.

Need it fixed, not just flagged?

We deliver broadcast-grade subtitles that pass these checks and the ones no tool can run: glossary lock, semantic alignment, and a native-speaker read. €400 per language for a feature, flat, with a free preview first.

Send a brief

Prefer email? hello@fonti.studio