Ten free transcription tools compared by what free really covers (open source, free tiers, and previews), plus when free is genuinely enough.
"Free transcription software" is one of those searches where every result says free and almost none of them mean the same thing by it. Some tools are free forever with no limits. Some hand you a monthly allowance. Some give you a taste and then a paywall. None of that is a scandal, since speech models and servers cost real money, but you can lose an afternoon discovering the fine print one signup at a time. This guide does the fine-print reading for you across nine tools, and finishes with an honest answer to the question that actually matters: when is free genuinely enough?
Want to try it on your own file first? Our audio to text tool transcribes the first 10 minutes of any file free, no card needed.
Each tool below is labelled by which kind of free it offers.

The most capable fully free option. Whisper runs on your own computer, transcribes unlimited audio in dozens of languages, and never sends a byte to anyone. The catch is the on-ramp: installation means the command line, and long files want a decent GPU or a lot of patience. Community front-ends soften this, but it remains a tool for the technically comfortable.
A lighter open-source toolkit that runs offline on modest hardware, including phones and single-board computers. Accuracy trails Whisper, but for embedded projects and ageing laptops it is the pragmatic pick.
Not automatic at all: oTranscribe is a free web editor for typing transcripts yourself, with playback shortcuts and one-key timestamp insertion. If your plan was always to hand-type an important interview, it removes the misery of juggling a media player and a document side by side.
Docs will type what it hears through your microphone at no cost and with no hard cap. The limitation is structural: it transcribes live speech, not files. People work around this by playing a recording out loud into the mic; results degrade accordingly, and it needs you sitting there for the full duration.
Our own tool, so treat this as a label rather than a review. You drag an audio or video file onto filetotext.ai and the first ten minutes come back transcribed free, in seconds, with no card needed. That preview is the honest kind of free (enough to judge accuracy on your own recording), and transcribing the rest of the file is paid. One practical consequence: a file under ten minutes fits entirely inside the preview, so a short recording lets you judge the full paid-grade quality before spending anything.
Otter's free plan gives a monthly allowance of live meeting minutes with a per-conversation cap, plus a small number of file imports over the lifetime of the account. Used lightly as a meeting notetaker, it holds up. Used as a free file transcriber, the import limit ends the experiment quickly.

Notta's free plan follows the familiar pattern: a modest monthly minute allowance combined with short per-file limits. Workable for quick voice notes; check the current caps before planning anything larger around it.
On Pixel phones, the Recorder app transcribes as it records: on the device, free, and searchable afterwards. Language coverage is narrower than cloud services and export options are basic, but for meetings and lectures you attend in person with a Pixel in your pocket, it is quietly excellent.
Recent iPhone and Mac releases generate transcripts of Voice Memos recordings automatically, on-device, at no cost. Same shape as Google's option: superb when the recording happens on your own phone, no help for files that arrive from elsewhere.
Word's dictation is free in the browser and types live speech reliably. The companion feature that transcribes uploaded recordings is tied to a paid Microsoft 365 subscription with its own monthly ceiling, so as free transcription software it only half qualifies.
Before signing up for anything above, run the same four checks. They expose the shape of a free plan faster than any feature grid:
The walls appear in predictable places: a two-hour recording meets a thirty-minute cap, a backlog of interviews meets a monthly allowance, a client's confidential call meets a free cloud service with vague data terms. At that point, compare paid options on price per hour of audio rather than on headline price. Our comparison of audio to text converters does exactly that across nine tools.
Yes: the open-source route. Whisper and Vosk have no usage limits because you supply the computing power. Every hosted service limits free usage somewhere, since processing audio costs the provider money on every file.
On-device options (Whisper, Vosk, Pixel Recorder, Apple's transcripts) never upload your audio, which is as safe as it gets. With free cloud tiers, read the data policy: where audio is stored, for how long, and whether it trains models. If the answers are unclear, keep sensitive recordings away.
Because minutes are the cost driver. Speech recognition consumes compute for every second of audio, so vendors meter exactly that. Feature gates come second; minute gates are nearly universal.
A ninety-minute lecture blows through most free tiers in a single file. The realistic free paths are recording it on a Pixel or iPhone and using the on-device transcript, or running the file through Whisper if you have the setup. Otherwise a small one-off payment is usually cheaper than the workarounds.
Curious what paid-grade output looks like before spending anything? Drop a file on FileToText at filetotext.ai. The first ten minutes are transcribed free in seconds, and you only pay if the quality convinces you.