Timing taken from the video
Cues are derived from the recording itself, so lines appear when the words are spoken. Text transcribed elsewhere and pasted into a subtitle file drifts, and viewers notice drift long before they notice a wrong word.
Subtitles and captions
KolWrite transcribes your video into timed, speaker-separated text you can edit, then exports SRT subtitles. Because the timing comes from the video itself, translated tracks for other markets reuse the same timeline instead of starting again.
Anyone can generate captions. The gap between generated and shippable is short, specific and almost entirely about the pass a human makes in the middle.
Cues are derived from the recording itself, so lines appear when the words are spoken. Text transcribed elsewhere and pasted into a subtitle file drifts, and viewers notice drift long before they notice a wrong word.
Proper nouns, product names, jargon and acronyms are where automatic transcription fails most visibly, and where captions embarrass a brand fastest. Fix them once in the transcript, then export.
Speaker separation keeps an interview, panel or two-hander readable, so a viewer following captions can still tell who is speaking when the shot does not make it obvious.
Translate between supported languages from the same timed transcript rather than transcribing the video again for each market — the cues stay aligned to the picture across every track.
Four different pressures land on the same file, and each of them ends with somebody asking for an SRT.
Deaf and hard-of-hearing viewers need captions, and accessibility standards for prerecorded video expect captions that are accurate, complete and synchronised. Automatic output is a starting point for that; the review pass is what gets you there.
A large share of social and feed video is watched muted. Captions are not an accessibility extra there — they are whether the video communicates at all in the first three seconds.
A transcript makes a video archive findable. Training libraries, conference recordings and product footage stop being opaque files once the words inside them are text you can search.
Translated subtitle tracks open a finished video to audiences it was never made for, at a fraction of the cost of re-recording or dubbing it.
Bring in the video file from desktop, mobile or the web app — the same account works across macOS, Windows, Linux, iOS, iPadOS, Android and the browser.
Get a timed transcript with speaker separation, aligned to what happens on screen.
Play back any line you are unsure about, fix names and terminology, and make the text say what the video says.
Translate to other supported languages if you need more tracks, then export SRT and load it into your editor, player or publishing platform.
SRT — the most widely accepted subtitle format, which video editors, players and publishing platforms will take directly. You can also export the same content as plain text, Word or PDF when someone needs the script rather than the captions.
Yes. Transcribe the video in the spoken language, then translate between supported languages to produce additional subtitle tracks from the same timed transcript. KolWrite supports 100+ languages for transcription.
Not on their own. Accessibility standards for prerecorded video expect captions that are accurate, complete and properly synchronised, and no automatic system guarantees that unreviewed. KolWrite gives you an editable, timed transcript so a person can review and correct it before export — which is the step that makes captions compliant, not the generation.
Yes. Speaker separation marks distinct voices with timestamps, which helps when captioning interviews, panels and discussions. Heavy crosstalk is harder for any system, so review those passages before exporting.
Edit the transcript in the app, using playback to check anything you are unsure of, and export the SRT afterwards. Do the corrections in the transcript rather than in the subtitle file — the timing stays intact and every translated track you generate afterwards inherits the fix.
Upload a video with real speech in it, export the SRT, and drop it onto the picture. Timing and names are where captions are won or lost.
Subtitle a video