Skip to content
Novus Examples
srt335 B

SRT — Speaker Labels and Dialogue Dashes

The conventions captioners actually use — speaker names in caps, leading dashes for alternating dialogue, square brackets and parentheses for sound effects, and music notes around lyrics. None is part of the SubRip grammar, so tools that extract speaker names or strip non-speech cues have to pattern-match them. Useful for testing transcript extraction and caption-cleaning passes.

Preview — first 21 linessrt
1
00:00:01,000 --> 00:00:03,000
ALEX: Speaker label in caps with a colon

2
00:00:03,500 --> 00:00:06,000
- First speaker on a dash line
- Second speaker replies

3
00:00:06,500 --> 00:00:09,000
[door closes]

4
00:00:09,500 --> 00:00:12,000
(SOFT MUSIC PLAYING)

5
00:00:12,500 --> 00:00:15,000
♪ Lyrics marked with music notes ♪

Specifications

Cues
5
Conventions
NAME:, leading dashes, [sound], (SOUND), music notes
Standard
convention only — none of this is in the SubRip grammar

What is a .srt file?

SRT (SubRip) is a plain-text subtitle format listing numbered cues, each with a start and end timecode and one or more lines of text. It is simple, human-readable, and extremely widely supported by players. It carries no styling metadata beyond basic inline tags.

How to use this file

Use an example SRT to test subtitle parsing, timecode handling, and converters that translate between SRT and WebVTT or other caption formats.

How to use this file for testing

“SRT — Speaker Labels and Dialogue Dashes” is a deterministic Novus Examples fixture for Subtitle parsing, Video QA. The same captions across SubRip, WebVTT, ASS/SSA, SBV, and TTML, plus a synced LRC lyric file — for testing subtitle parsers, players, burn-in tools, and format converters against known timings.

Documented properties for this file: 5 cues. Compare results against paired or grouped companions on this page when present (clean↔damaged, searchable↔scanned, or format twins) so scores stay reproducible across runs.

Download the file once, keep the path stable in CI or local scripts, and treat the spec table as the contract: dimensions, seeds, field lists, and roles are intentional. Corrupt or invalid samples are labelled as such — expect parsers to fail loudly rather than silently accept them.

Media fixtures are short and synthetic by design. Prefer waveform or transcript ground truth in the same group when measuring ASR, trim, upscale, or sync tools; do not assume broadcast-quality masters.

Parse the cues and check timings against the documented count; the same captions ship across subtitle formats so you can diff a converter against a known target.

Code examples

<video controls src="clip.mp4">
  <track kind="captions" srclang="en" label="English" src="speakers.srt" default>
</video>

Generated by generation/video_timedtext.py. Free for any use, no attribution required — license.