Free Speech to Text (Runs in Your Browser) — Uttir
Uttir

Speech to Text

Transcribe audio to text in your browser. Upload a file or record from your microphone — a small local model (~40 MB) does the work. Nothing is uploaded.

Speech to Text — This tool transcribes audio to text using a small AI model that runs in your browser. The audio and the model never leave your device.

Runs entirely on your device. The audio and the model stay on your device — nothing is uploaded. The first run downloads a small local model (~40 MB); after that, transcription works offline.

Was Speech to Text useful?

What is speech to text?

This tool transcribes audio to text using a small AI model that runs in your browser. The audio and the model never leave your device.

The first run downloads the model (~40 MB). After that, transcription works offline. You can upload an audio file or record from your microphone.

Output is shown as text with timestamps. You can copy the text, download as .txt, or download as .srt subtitles (compatible with YouTube, VLC, and most video editors).

Everything — the audio, the model, and the result — stays on your device. There is no upload, no account, and no analytics on the audio itself.

How to use it

  1. Pick an audio source

    Choose an audio file (mp3, m4a, wav, webm, ogg) or click Record to capture from your microphone.

  2. Download the model on first use

    The first run downloads a small local model (~40 MB). After that, transcription works offline.

  3. Click Transcribe

    Inference is local. A one-minute clip typically finishes in a few seconds on a modern laptop. The result appears below with timestamps.

  4. Copy or download

    Use Copy for plain text, or download as .txt or .srt. The .srt format works with most video editors and YouTube for subtitles.

Frequently asked questions

Related tools

Last reviewed: 2026-08-28