GoodWebTools

Voice to Text

beta

This free voice-to-text tool transcribes speech from a recording or an audio/video file into editable text, timestamps and subtitles using on-device Whisper AI. You can record right in the page or drop a file, and everything is processed in your browser, so your audio never leaves your device.

or
Model
Record or drop a file to enable this.

How to use Voice to Text

  1. 1 Click Record to capture from your mic, or drop an audio or video file (mp3, wav, m4a, mp4…).
  2. 2 Pick a model — fast English, accurate English, or multilingual — and, for multilingual, choose the spoken language.
  3. 3 Click Transcribe; the model downloads once on first use, then runs on your device.
  4. 4 Switch between Text, Timestamped and Subtitles tabs, edit the text, and download .txt, .srt or .vtt.

Frequently asked questions

Is my audio uploaded to a server?

No. Transcription runs entirely in your browser with an on-device Whisper model. Your recording or file never leaves your device, so private conversations stay private.

Can it transcribe languages other than English?

Yes. Choose one of the multilingual models and select the spoken language — Indonesian, Malay, Japanese, Spanish and many more — for the best accuracy. Auto-detect is also available.

Can I export subtitles?

Yes. Open the Subtitles tab and download SRT or VTT files with timestamps, or grab plain text (.txt) or a timestamped version instead.

Why does the first run take a while?

The AI model is downloaded and cached on first use, and larger or 'Better' models are slower — especially on phones. After the initial download it loads from cache, and transcription runs locally on your device.