Your YouTube video
Paste a YouTube link. Nothing to upload.
Make a podcast
A listenable version of something you would otherwise have to sit and watch.
Free to try, no card. Your YouTube video opens in Ancher.
Paste a YouTube link. Nothing to upload.
Every one of these stays separately addressable instead of collapsing into one block of text.
A narrated version you can listen to instead of read.
Some videos are worth an hour of your attention and some are worth twenty minutes of your commute. Ancher writes a narration script from the argument rather than the transcript, describes the on-screen material that a listener cannot see, marks the segment breaks where the subject actually turns, and puts the creator's credit in the script itself so the source travels with the audio.
The honest version
The reason a read-aloud transcript is unlistenable is not the voice. It is that every reference to something visual becomes a hole — this graph here, look at what happens, the thing on the right. A script that works as audio has to convert what was shown into what can be said, and that is only possible if the frames were read in the first place.
One second of one video, at 14:03, read two ways.Illustrative. The figure is invented; the gap it falls into is not.
...and you can see right here what happens to the second line — — which is the whole point.
On the graph — not readThe specific never reaches the draft
Lands in the draft, with its timecode
What comes out
A narrated version you can listen to instead of read.
Written to be spoken rather than read, with the visual references converted into description.
Breaks placed where the argument turns, so the piece survives being listened to in two sittings.
A consistent register chosen for the material — not an imitation of the original speaker.
A target runtime, with the cuts made to reach it listed rather than hidden.
The creator, the channel and the link, spoken in the audio as well as written down.
The link on its own gets a summary. This is the instruction that produces the 5 sections above, in that order, with the rules that keep them honest. Paste it with your YouTube video — in Ancher, or in whatever assistant you already use.
Write a narration script from this video. (1) Rewrite it to be spoken to a listener, not read from a transcript. (2) Every reference to something on screen must be converted into a description of what was actually shown. (3) Mark segment breaks where the argument turns. (4) Aim for roughly half the original runtime and list what you cut. (5) Write the creator credit and the source link into the script itself. Do not imitate the original speaker's voice or manner.
The detail, if you want it
| From the YouTube video | Into | Why |
|---|---|---|
| Caption track with timings | Script | The transcript is the raw material, but it is rewritten rather than read, because speech written for a viewer is not speech that works for a listener. |
| Key frames from the video | Script | Every on-screen reference is a hole in audio, and the frame is the only way to fill it with something true. |
| Caption track with timings | Segments | Subject changes in the timings mark the breaks, which is more reliable than cutting on a fixed interval. |
| Title, channel, duration, date | Source credit | Title, channel and date make the credit specific enough to be worth saying aloud. |
| Title, channel, duration, date | Length | The source duration sets what a realistic target runtime is before any cutting decisions get made. |
Published and auto-generated captions both come through, and every line keeps the second it was spoken.
Frames are sampled and scored, so a number shown on a slide survives even though it was never said out loud.
The creator's own cited links come along, so a claim can be traced past the video.
Reception is a weak signal about correctness and a strong one about which claims got attention.
Basic provenance, kept so the finished work can attribute the video properly.
No, and deliberately so. The script is written for a neutral reading voice rather than as an imitation of the speaker, because cloning someone's voice to retell their own work is a problem no amount of credit fixes.
Usually about half, because speech to camera carries a lot of repetition that a listener does not need repeated. The cuts are listed with the script, so you can put anything back that you disagree with.
Only with the creator's permission. What it is genuinely good for is a listenable version of something you need to absorb — a talk, a briefing, a long explainer — where the audience is you rather than a feed.
Everything you save lives in one workspace, so the podcast is built from your sources — not from a model's memory of the internet.
Open Ancher →Not a YouTube video? The same podcast also comes from X thread, LinkedIn post, Instagram post.