Speech to text, on your device
Transcribe anything, no minute limits
Transcription services meter you by the minute and want your recording on their servers. This one runs OpenAI's Whisper model inside your browser tab, so interviews, lectures and client calls are transcribed without ever leaving your device — and without a meter.
Nothing uploaded
Confidential interviews and client calls stay on your machine.
Subtitles included
Export .srt with timestamps, ready for any video editor.
90+ languages
The multilingual model transcribes far beyond English.
How local transcription works
Why is there a download the first time?
The speech recognition model itself has to reach your device before it can run. It is fetched once, cached by your browser, and reused instantly on every later visit.
How long does a recording take?
Roughly real-time to a few times faster on a modern laptop with the fast model. Long recordings are processed in 30-second chunks, so progress is steady rather than all-at-once.
Which files can I transcribe?
Any audio or video your browser can decode — MP3, WAV, M4A, OGG, MP4, WebM and more. The audio track is extracted and resampled automatically.
How accurate is it?
Whisper is strong on clear speech and holds up well with accents. Heavy background noise, crosstalk and very quiet recordings are where any model, local or cloud, starts to struggle.