Add VTT subtitles to your video.
Drop your video and your VTT. Style the captions. Export with the subtitles burned in. Runs entirely in your browser.
Drop your video
or click to browse
MP4 · WEBM · MOV
Drop your VTT
or click to browse
VTT
Free. No account. No watermark.
How it works
Drop, style, export.
Drop your video and VTT
The video and the subtitles stay on your device. Both are read locally, into the browser's private storage. Nothing uploads.
Style the captions
Pick a template, tune sizes, colours, and animations. The preview updates in place. Word timings come straight from the VTT cues.
Export a subtitled video
The captions are burned into the pixels. The exported MP4 plays with subtitles baked in on any player, no separate track to lose.
Beyond VTT
Word-level precision. AI on top.
VTT files carry timings per line. That works for readable captions, but the word-by-word styles that go viral on TikTok, Reels, and Shorts drift when the timing source is per-line. tscaps cloud transcribes with AI at the word level and layers editing intelligence on top of it.
Word-level timings
Every word timed precisely, so word-by-word caption templates fire on the exact syllable instead of drifting inside the cue.
Auto-picked emojis
Emojis chosen per phrase from the transcript context — no manual pass, no template that reads flat.
Bad takes removed
Filler words, stumbles, and repeated takes flagged and cut so the final video reads clean without editing frame by frame.
Hooks + tagged words
The opening hook is detected and styled distinctly. Semantic tags mark the important words so templates emphasise them naturally.
FAQ
Frequently asked questions.
What is a WebVTT file?
WebVTT (`.vtt`) is a plain-text subtitle format designed for the web. Each cue has a start time, an end time, and a line of text — same shape as SRT but with millisecond precision and native support in HTML `<track>` elements.
How do I add a VTT file to a video?
Drop the video in the first slot and the VTT in the second, then press Start. The editor opens with the cues loaded as captions you can restyle or retime. Export when it looks right — the result is a video with the subtitles permanently added.
Does the tool upload my video or my VTT?
No. Both files stay on your device. The tool opens them locally, the browser renders the captions on top of the video, and the export composites frames in-page. Nothing is sent to a server.
What video formats work?
MP4, WebM, and MOV. Support depends on your browser's built-in codecs — most modern browsers handle MP4 (H.264) out of the box. Chrome, Edge, and Firefox on desktop are the smoothest experience today.
What does it mean to burn the subtitles?
The captions are baked into the pixels of the exported video. There is no separate text track that can be toggled off or lost. The output plays the same way on TikTok, Reels, Shorts, or any other player.
What VTT features are supported?
The parser reads the cue timings and the visible text. Cue identifiers are skipped. NOTE, STYLE, and REGION blocks are ignored — they only affect how VTT is rendered inside a native `<track>` and do not change the burned pixels. Inline tags like `<v Speaker>` or `<c.class>` are stripped; the tool renders captions using its own template styles.
Can I edit the captions before burning?
Yes. The editor loads the VTT cues as segments you can rewrite, retime, and style. You can also change wording, adjust timings per cue, and pick a caption template for each part of the video.
Does the exported video have a watermark?
No. Everything runs on your device, so there is no watermark on the output.
Can I use an SRT file instead of VTT?
Yes — there is a companion tool that accepts SRT files directly. If you have an SRT file, open the SRT variant of this tool; if you have both formats, either one works.
How long can my video be?
There is no fixed cap. Longer videos take longer to export because every frame is rendered in the browser; for practical results, most visitors work with clips under fifteen minutes.