Speech to Text
Transcribes an audio file, or types what you say as you speak.
Runs in your browser. The main work runs in your browser. Only the step named below is sent to our server, and it is not stored. Free, no account.
How it works
You have a recording and you need the words out of it. Upload an audio file up to 12 MB and this tool sends it to a speech model that returns the text with a timestamp for each segment. Or pick a language and let your browser type as you talk.
- 1
Upload an audio file up to 12 MB (about ten minutes of MP3). It is sent to a hosted speech model, which returns the text with a timestamp for each segment.
- 2
Or use live dictation: your browser's own speech recognition types as you talk, in the language you choose.
- 3
Edit the transcript in place, then copy it or download it as text.
What it does not do
Live dictation is provided by the browser and works in Chrome, Edge and Safari. Audio may be processed by the browser maker's speech service; availability and processing depend on the browser and settings.
File transcription has a shared daily capacity and resets at midnight UTC.
Background noise and overlapping speakers lower accuracy. Check names and numbers.
When a source is busy or a site does not answer, this tool says so. It never fills the gap with a guess.
Questions
Why does my transcript get names and numbers wrong?
The model works from the sound, so a noisy room, music or two people talking at once blur the words it hears. Names and numbers suffer most because they are rarely predictable from context. Read the transcript before you use it and correct those by hand. The editor lets you fix the text in place.
What happens to my audio file when I upload it?
When you upload a file, it goes to a hosted speech model and comes back as text with a timestamp for each segment. Live dictation is different: the browser handles it, and audio may be processed by the browser maker's speech service, depending on the browser and settings.
Why does live dictation not work in my browser?
Live dictation uses the speech recognition built into your browser, and that feature is available in Chrome, Edge and Safari. Other browsers do not provide it, so the microphone option will not start. Use one of those three, or upload an audio file instead and let the server transcribe it.
Tools that answer the next question
- Online Text EditorA plain-text editor with live word and character counts, find and replace, sorting and a save button.
- Text SummarizerPicks the sentences that carry the most of a text and returns them in order, at the length you choose.
- Grammar CheckerFinds spelling, grammar and punctuation mistakes and offers a fix for each, without your text leaving the browser.