Social video with silent autoplay
Keep essential speech and context visible when viewers start with audio muted.
Import subtitle files or write captions manually, then burn styled text into the video for reliable playback anywhere.
Drop a video here
or click to choose a file
Supported formats: MP4, WebM, MOV, MKV
Maximum file size: 2 GB
Supported subtitle formats: SRT and VTT
Choose the original character encoding only when older subtitle files display garbled text.
Text becomes part of the picture and cannot be switched off in the exported file.
File name
-
Original size
-
Output size
-
Resolution
-
Duration
-
Add subtitles
-
Add the video, import an SRT or VTT file or create timed cues, adjust the subtitle style, preview key moments and encode the result. Burned-in subtitles become part of each frame, so viewers cannot turn them off or edit them later.
Burned-in subtitles are drawn directly onto each video frame. They display consistently without a separate subtitle track, but cannot be hidden, restyled or corrected after export without encoding the video again.
Keep essential speech and context visible when viewers start with audio muted.
Embed text when the destination cannot reliably load or select an external subtitle track.
Preserve terminology, steps and warnings as part of the visual record.
Create a separate encoded video for each language when a platform cannot switch subtitle tracks.
Choose the source from your device or load a direct public URL that the browser can access.
Upload SRT or VTT, or create timed cues manually. If old files show garbled characters, choose the correct text encoding and recheck every line.
Make every cue start and end at the intended moment. Keep captions concise, avoid accidental overlaps and split long text into readable lines.
Choose a font with all required characters, add outline or background where needed, and keep text away from edges and player controls.
Check the beginning, cue changes, fast speech, scene cuts and the final seconds, then encode and download the video before refreshing the page.
Bring a caption in when the phrase begins and remove it after the thought is complete. Avoid changing captions in the middle of a word or leaving stale text over a new scene.
Two balanced lines are generally easier to scan than one very long line. Break at natural phrase boundaries and avoid covering important visual detail.
White text alone can disappear over bright footage. An outline, shadow or controlled background improves readability across changing scenes; preview both light and dark moments.
A font may look good but omit Cyrillic, CJK, accents or symbols. Test representative names, punctuation and special characters before a long encode.
Export or retain the SRT/VTT source before burning captions. If wording or timing changes later, editing the subtitle file is faster than rebuilding it from the rendered video.
A local video and subtitle file are read and rendered in the current browser. Remote video URLs require network access to the source host. Subtitle text may contain sensitive names or dialogue, so download the result and clear the session on shared devices.
Use SRT or VTT files supported by the current parser. Advanced styling or nonstandard extensions may be simplified or rejected, so inspect every cue after import.
Burned-in text is part of the picture and always visible. Selectable subtitles remain a separate track that a compatible player can enable, disable or restyle. This tool's final render uses burned-in text.
The file may use an older text encoding instead of UTF-8. Select the correct encoding, reload the file and confirm names, punctuation and accented or CJK characters.
Not inside the finished video. Keep or export the SRT/VTT source, make corrections there and encode a new version.
Cue times can be entered precisely, but actual display is constrained by decoded frame timing, frame rate and the encoder. Preview around fast changes and scene cuts.
The selected font may not include those glyphs. Choose a font with suitable language coverage and test representative text before the full render.
The editor may distribute one line per cue across a chosen interval, but it does not understand speech timing. Treat the result as a draft and manually align every cue to the video.
Burning text requires re-encoding the video, so output quality depends on the selected codec and quality settings. Use an appropriate setting and inspect fine detail before delivery.