Podcast transcription
Drop in the audio and get a clean, speaker-labeled, timestamped transcript back in minutes, plus an AI summary you can paste as your episode description. 95%+ accuracy on clear audio, honest about the rest. The first 10 minutes of any file are free.
This is a real transcript synced to its audio. Press play, or click any word to jump straight to that moment. Same thing happens to your episode.
Every line is timestamped and speaker-labeled, so the transcript is ready for your show page and it's easy to read chapter markers off the topic changes. The AI summary gives you a clean prose recap for the episode description. A full transcript is also the biggest SEO win a podcast can get: an hour of audio Google can't read becomes an indexable page.
Every line carries its moment, so chapter markers are just the topic changes. Speaker labels rename in one edit.
A conversation on transcribing podcasts from recorded files rather than live tools. The hosts cover the preview-before-you-pay model, where you see the full transcript quality before paying, and discuss realistic accuracy: strong on clear audio, lower on crosstalk and compressed remote calls. They close on pricing and which plan fits a weekly show.
A clean prose recap, ready to paste as your episode description.
How accurate will it be on my setup, how do I get my file in, and what does it cost. Straight answers, right here on the page.
95%+ on clear audio. Where your setup lands depends on how it was recorded. Pick yours, honestly. See the live benchmark.
Two hosts, local-recorded (Riverside, Zencastr): Distinct voices on their own tracks stay close to the top. You review the flagged words in the editor, so the version you ship is right either way.
Drag to your real episode length. The price is the price.
Pay per episode
$3
pay once, $12/mo at this cadence
Basic plan
$10/mo
covers ~10 hrs of audio a month
At 4 episodes a month, the $10 Basic plan is cheaper than paying per episode for the same product (every plan has every feature).
Hours round up: a 1 hour 2 minute episode bills as 2 hours.
Where did you record it?
Tip: Riverside records each person locally, so its audio is about as clean as remote gets.
The stuff that actually decides whether your transcript is good, none of which is about the transcription tool.
If you record on Riverside or Zencastr, each host lands on their own track. That is close to the best case for telling voices apart. A single mixed-down MP3 still works, but when two similar voices are glued into one channel, expect to fix a few speaker labels by hand. Either way, renaming "Speaker 1" to a real name updates it everywhere in one edit.
A music-only cold open just comes back as silence in the text, which is correct. Speech over a light music bed transcribes cleanly. It is only dense music mixed loud under talking that shades a few words, and the per-word confidence view flags exactly those so you are not re-reading the whole thing.
A guest dialing in over a compressed call will transcribe a little worse than your studio mic. That is the audio, not the model. If it matters, ask guests to record their side locally (a voice memo works) and send it. Our interview transcription guide goes deeper on multi-person audio.
For a show page, use filler-word cleanup so the transcript reads like writing instead of a court record. The original stays available if you ever need exactly what was said.