The VTT Cue Combiner

The VTT Cue Combiner merges consecutive WebVTT cues (also called segments) that share the same speaker label into unified paragraphs, updating each combined section’s timecode to reflect the full duration. The resulting WebVTT file is optimized for use with OHMS (the Oral History Metadata Synchronizer), which prefers transcript timecodes aligned to speaker changes. In addition to VTT files, this tool can also process SRT transcripts — provided the SRT includes identifiable speaker labels.

Your data stays private: This tool processes transcript files entirely within your web browser — nothing is uploaded to any server and nothing is saved.

No files selected.
Configure
Cue Combining & Export
Unlabeled VTT cues inherit the last labeled speaker until a new label appears.
Unlabeled SRT cues inherit the last labeled speaker until a new label appears.
Recommended to avoid overwriting originals.
Blather Buster
Enabled by default. Splits long single-speaker sections into more manageable continued sections.
Default: split speaker sections longer than 3 minutes near the 2-minute mark.
seconds
seconds
seconds
seconds
Batch ResultsHide Batch Results
No files loaded.

Preview Processed File

Choose one processed file to preview both the original input and grouped output.

Original / Input Preview

No file loaded.

Grouped Speaker Output VTT Preview

The VTT Cue Combiner was created by Douglas A. Boyd. This tool is provided for non-commercial and archival purposes and is used at your own risk. No warranty is expressed or implied; results should be reviewed for accuracy before publication or deposit. For questions or feedback, please use the contact form.|Help / About