Your recording
Paste your recording. Nothing to upload.
Make a transcript
Every segment timed, so any sentence can be found and played back.
Free to try, no card. Your recording opens in Ancher.
Paste your recording. Nothing to upload.
Every one of these stays separately addressable instead of collapsing into one block of text.
The spoken words, cleaned up, with a timestamp on every segment.
A transcript is only useful if it is anchored. Ancher transcribes the audio and puts a start and end time on every segment, cleans filler without paraphrasing, samples frames from anything shared on screen so a number that was pointed at survives, and states plainly that there are no speaker labels rather than inventing them.
The honest version
Without timings, a forty-page transcript of a ninety-minute call is unusable in a dispute — finding the sentence means rereading or rescrubbing. With a timing on every segment, a contested line becomes a click. That is the difference between a transcript that gets referenced and one that gets generated and never opened again.
One second of one recording, at 18:47, read two ways.Illustrative. The figure is invented; the gap it falls into is not.
if you look at the number on the right there, that's the one that changed —
On the shared screen — not readThe specific never reaches the draft
Lands in the draft, with its timecode
What comes out
The spoken words, cleaned up, with a timestamp on every segment.
The spoken words, cleaned of filler, with screen content noted where something was shared.
A start and end time on every segment, not a marker every few minutes.
What was removed, so nothing reads as a paraphrase of what someone said.
Plain text you can paste into a document or a ticket.
Indexed with your other sources, so a search reaches this call by something said in it.
The link on its own gets a summary. This is the instruction that produces the 5 sections above, in that order, with the rules that keep them honest. Paste it with your recording — in Ancher, or in whatever assistant you already use.
Transcribe this recording. (1) Give the spoken words cleaned of filler, without paraphrasing. (2) Put a start and end time on every segment. (3) Note what you removed in cleanup. (4) Where something was shared on screen, describe what was shown and when. (5) State at the top that there are no speaker labels and do not attribute any line to a person.
The detail, if you want it
| From the recording | Into | Why |
|---|---|---|
| Transcript with timings | Timestamps | Every segment carries start and end seconds, which is what turns a claim into a record. |
| Transcript with timings | Full text | The spoken words are the document, and cleaning filler without paraphrasing is what keeps it a transcript rather than a summary. |
| Screen shares, if it is video | Full text | A shared screen is sampled as frames, so a figure that was pointed at rather than read out still appears. |
| Duration and timeline | Searchable in your workspace | Knowing the length lets a search result say where in a ninety-minute call the match is. |
| Transcript with timings | Filler removed | Removing filler is a decision, and recording what was removed is what keeps the transcript defensible. |
Every segment carries start and end seconds, which is what turns a claim into a record.
I'll take that, by Friday, let's park it — these phrases are what separate a task list from a topic list.
What a conversation failed to close is usually the reason for the next one.
A shared screen is sampled as frames, so the number that drove a decision stays attached to it.
Where the time actually went is a finding in itself for any recurring meeting.
No. It gives you what was said and when. For anything that turns on attribution — a disciplinary record, a legal matter — this is the wrong tool and the transcript says so at the top rather than in a footnote.
Good on clear single-speaker audio and worse with crosstalk, room mics or accents the model handles poorly. The failure mode is worst precisely where people interrupt each other, which is where decisions tend to happen.
Because in a screen-shared call the number is on the screen and the speech is that one right there. A transcript without the frames records a sentence that points at nothing.
Everything you save lives in one workspace, so the transcript is built from your sources — not from a model's memory of the internet.
Open Ancher →Not a recording? The same transcript also comes from YouTube video, Instagram post, TikTok video.