Three different “times” people try to delete
An SRT opened in a text editor shows three layers that all feel like clutter. The cue index
(1, 2, 3) is only there so players can count blocks. The arrow line
(00:01:04,200 --> 00:01:07,800) is the real clock. The line breaks inside the cue are where
a captioner wrapped text for the screen, not where a sentence ended.
Converting always drops the index and the arrow syntax. That is not optional, and it is not the timestamp
toggle. Show timestamps controls a fourth thing: a short start–end range printed for
humans, such as 01:04–01:07, above the paragraph. Hours show up in that label only once the
cue passes an hour. Uncheck the box and those labels are absent from the preview and from the
.docx. The sentences remain. Speaker names remain if the file had them.
So “SRT without timestamps” on this site means a reading document. It does not mean a new caption file
with the clocks stripped, and it does not mean Word has become a subtitle editor. The original
.srt or .vtt is unchanged on disk.
When the clocks should be off
Turn them off when the next reader is not going back to the video. Meeting minutes, a client-facing
recap, a newsletter draft, a quote you will paste into a treatment, and a script pass that is about
wording all read better without a time code on every paragraph. The codes shout “this is still a caption
export” even after the cue numbers are gone.
Turn them on when anyone in the chain might ask where a line sits. Translators, lawyers, producers, and
you — six weeks later — use the range as an index. A clean document without the original caption file is
how that question becomes a full re-watch. The toggle is cheap. Download both versions from the same paste
and name them. notes-clean.docx and notes-timed.docx take less time than
reconstructing clocks from memory.
Interviews are the case where people hide timestamps too early. The names are the structure; the times are
the index. Get the names right first
(interview SRT to Word), keep times on through the factual read, and
hide them on the copy that leaves the building.
What merge does once the times are hidden
Merge same speaker is independent. It joins consecutive cues from the same speaker when
the gap is about 2.5 seconds or less, and it joins unlabeled cues the same way. With timestamps hidden,
that join is the difference between a readable paragraph and a column of caption-sized lines that still
look like subtitles.
Leave merge on for a single narrator, a lecture, or a YouTube auto-caption track. Turn it off when every
cue is its own beat: a bilingual pair, a list of lower-thirds, or dialogue you have not labeled yet and do
not want fused into one voice. Hiding timestamps does not freeze the cue boundaries. If the preview shows
one wall of text, merge is why. Split it by turning merge off, or by adding speaker prefixes so only one
person’s lines join.
Whisper files are where this surprise is largest. One-word or one-phrase cues, all unlabeled, become long
paragraphs as soon as merge is on. That is what you want for a reading script, after you have deleted
repeated junk lines. The deletion step is
Whisper SRT to Word. Hiding timestamps before that deletion just hides
the clocks on the junk, too.
Do not do this with Find and Replace
The manual version is: open the SRT in Word, replace every line that contains -->, replace
the cue numbers, then fix the mid-sentence breaks. It works on a 20-line sample and it fails in the same
predictable ways on a real file. A sentence that mentions a time (“see you at 10:30”) can match a sloppy
pattern. A cue number that sits on its own line is easy to delete until a numbered list inside someone’s
speech looks identical. Line wraps that were only visual become paragraph marks you then have to join.
The toggle avoids that class of mistake because the clocks are never written into the document. There is
nothing to search for. If a range is present, you left the box checked. Uncheck it and download again.
Stripping timestamps inside the SRT itself — deleting the arrow lines so the file is “just text” — is worse.
You no longer have subtitles, and you may no longer have a parseable cue list. Keep the caption file whole.
Hide clocks only in the reading export.
A small delivery habit
- Save the source
.srt or .vtt with a name that includes the language and the date. - Download the timed
.docx if anyone will QC against picture. - Download the untimed
.docx for readers. Put one sentence at the top: which recording, which caption file. - If the untimed file is going to Google Docs, upload the
.docx, not the SRT. See SRT to Google Docs.
Zoom meeting transcripts are the same toggle with a different source file. If you are starting from a cloud
recording rather than from SubRip, the download steps are on
Zoom transcript to Word. Hide the clocks there when the output is
minutes. Keep them when the output is a record of who said what, and when.
Related workflows
Download the .docx
The link unchecks Show timestamps and leaves Merge same speaker on. Drop the SRT. If you still see 01:04–01:07 in the preview, the box is checked. Download reading.docx and keep the .srt.