Convert audio to text — free and unlimited

Drop in almost any audio file — MP3, WAV, M4A, FLAC, OGG — and get an editable transcript with timestamps. There is no minute cap, no account, and no queue. Your files are never uploaded — everything runs on your device.

Most transcription sites limit the free tier because every minute costs them server money. This one runs the speech model inside your own browser, on your own hardware, so there is nothing to meter. The engine is open source, and after the first model download it works offline too.

Checking what this browser supports…

How to transcribe audio

  1. Pick your audio file

    Drag it into the box above or click Choose file. Common audio formats are supported directly — no conversion needed first.

  2. Wait for the one-time model download

    On first use the speech model (about 80 MB) downloads into your browser cache. It stays cached, so later transcriptions start immediately.

  3. Watch the transcript stream in

    Long recordings are processed in chunks, so text starts appearing before the file is finished. Progress and a time estimate are shown throughout.

  4. Edit and export

    Click any segment to fix a word, then download the result as TXT, SRT, VTT, or JSON — or copy the plain text.

Why this converter is different

Actually unlimited

No minute caps, no monthly quota, no premium tier. Because inference runs on your device, a three-hour file costs us the same as a three-second one: nothing.

Private by design

The audio is decoded and transcribed locally. Privacy here is not a policy promise — uploading your file is simply not part of the design.

No signup, no nagging

Open the page, drop a file, leave with your transcript. No account, no email, no trial countdown.

Frequently asked questions

Is it really free without limits?
Yes — YourDevice transcription is free with no account, no watermark, no per-file cap and no limit on how long a recording can be. Whisper runs in your browser on your own hardware, so there is no per-minute server cost for us to pass on; the site is paid for by ads rather than by metering the tool.
Which audio formats work?
Anything your browser can decode: MP3, WAV, M4A/AAC, FLAC and OGG cover almost everything. Video files work too — YourDevice extracts the audio track in the browser, so an MP4 needs no converting first. The finished transcript downloads as plain text, SRT or VTT subtitles, or a Word document.
How accurate is the transcription?
YourDevice transcribes with Whisper, OpenAI’s open-source speech-recognition model and one of the most widely used available. Accuracy depends on audio quality, accents and background noise, so we won’t quote a single percentage — a clean interview and a recording made across a noisy room are not the same job. Run a real file through it and judge the result; nothing is uploaded either way.
Which languages are supported?
Forty, listed by name — every European language plus the major world languages, curated from the roughly 99 Whisper knows. You pick the spoken language yourself before transcribing. There is no auto-detect, deliberately: the underlying library has not implemented language detection, so an “Auto” option would quietly transcribe German audio as English rather than admit it did not know. There is also a translate option that outputs English regardless of the source language.
Does it work offline?
After the first visit, yes. Your browser caches the Whisper model and the app itself, so once the page has loaded, transcription needs no internet connection at all — the work happens on your own processor or graphics card, on a plane or offline.
Where does my file go?
Nowhere. YourDevice reads, decodes and transcribes the recording inside your own browser tab. Your files are never uploaded to any server — ours or anyone else’s — and nothing is kept once you close the tab.