Do this
All guides- Show timestamps On for a record of the conversation
- Merge same speaker On, so one person’s short lines join
- Prefer the SRT if Otter offers one An SRT has --> lines and parses as cues. A conversation export that is already .docx is finished — open it. Do not upload that .docx here; this tool reads caption text, not Word files.
- Separate the transcript from the recap Outline, action items, and the AI summary are not timed cues. Pasting them above the transcript makes the preview fail or pulls junk into the first paragraph. Copy the spoken part only.
- Match a shape the parser knows Works: an SRT or VTT; a speaker line, then a clock such as 00:12:03, then the sentence; or a line like 00:12:03 Alex: sentence. A free-form paragraph with the time buried in the middle often produces no cues.
- Turn merge on and read the bold names Otter splits one answer across several short lines. Merge joins the same speaker when the gap is about 2.5 seconds. If two people share a spelling, they stay one speaker.
- Download otter.docx and keep Otter’s export The .docx is for notes and comments. Otter is still the place to correct a misheard name if you will share the conversation from there again.
Drop an Otter .srt, or paste blocks shaped as a speaker line, a timecode line, then the sentence. If the preview says no cues were found, you pasted the summary or a shape this parser does not read. Download otter.docx only after speakers show up.