Speech to Text Online Free
Convert speech to text in real time using your microphone. Supports 15 languages with live transcription. No signup, no upload — uses your browser's built-in voice recognition.
⏱ 6 min read · Complete guide below
How Speech to Text Works
- 1Select your language from the dropdown. The default is English (US).
- 2Click Start Recording. Your browser will ask for microphone permission — allow it.
- 3Speak clearly. Words appear in real time — grey text is interim (being processed), black text is finalised.
- 4Click Stop when done, then Copy text to paste your transcript anywhere.
Powered by the Web Speech API
This tool uses the browser's native Web Speech API — no external library or API key required. Chrome and Edge route the audio through Google's speech recognition infrastructure; Safari uses Apple's on-device recognition. No data passes through PublicSoftTools' servers. Accuracy is equivalent to the dictation feature built into your operating system.
Why Dictation Can Be Faster Than Typing
Most people speak far faster than they type. A comfortable speaking pace is around 130 words per minute, while average typing sits closer to 40, so dictating a first draft can be several times quicker than writing it by hand. Speech to text is especially useful for capturing ideas while they are fresh, drafting emails and messages on the move, taking notes hands-free during a call, or simply resting your wrists. The workflow that works best is to speak first, edit later: let the words flow without stopping to fix every small error, then tidy the transcript afterwards. Fighting for perfect accuracy as you go breaks your train of thought and cancels out the speed advantage.
Getting the Best Accuracy
Recognition quality depends heavily on your setup. The biggest factors are a quiet environment and a decent microphone positioned reasonably close to your mouth — background music, television, or a noisy room will confuse the engine. Speak at a natural, steady pace rather than unusually slowly; the recognition is trained on normal speech patterns. Crucially, select the correct language before you start, since dictating English while the tool is set to Spanish produces gibberish. For punctuation, many recognition engines accept spoken commands such as saying “comma” or “new line,” though support varies by browser and language.
Privacy and How Recognition Happens
It is worth understanding where your voice is processed, because it differs by browser. This tool uses the standard Web Speech API built into your browser rather than any service of ours, so your audio never touches PublicSoftTools' servers. On Chrome and Edge, the browser sends the audio to Google's speech recognition infrastructure to turn it into text — the same mechanism the browser uses everywhere. On Safari, Apple processes it on-device. If on-device privacy is a priority, Safari is the better choice; if you want the widest language support, Chrome or Edge tend to lead. Either way, the transcript stays in your browser until you copy it out.
Use Cases
Meeting Notes
Dictate meeting notes hands-free and copy the transcript into your note-taking app or email in seconds.
Accessibility
Useful for anyone who types slowly or has difficulty using a keyboard — just speak and get text.
Content Drafting
Dictate first drafts of articles, emails, or social media posts faster than typing, then edit the transcript.
Language Practice
Check your pronunciation by speaking in a foreign language — if the transcription is correct, the recognition engine understood you.
Frequently Asked Questions
Which browsers support speech to text?
The tool uses the Web Speech API, which is supported in Chrome, Edge, and Safari. Firefox does not currently support the Web Speech API. For the best experience, use the latest version of Chrome or Edge on desktop or Android, or Safari on iOS.
Which languages are supported?
The tool supports 15 languages: English (US and UK), Spanish, French, German, Italian, Portuguese (Brazilian), Arabic, Chinese (Simplified), Japanese, Korean, Hindi, Russian, Turkish, and Dutch. Select your language from the dropdown before starting.
Is my voice recorded or sent to a server?
The transcription uses your browser's built-in Web Speech API. On Chrome and Edge, the audio is processed by Google's speech recognition servers, which is how the browser's native API works. On Safari, it is processed on-device by Apple. No data is sent to PublicSoftTools' servers.
Can I transcribe a long recording?
Yes. The tool uses continuous recognition mode, so it keeps transcribing as long as you are speaking and have not pressed Stop. For very long sessions, pause and resume as needed. There is no enforced time limit, though browser tab focus and microphone permissions must remain active.
Why does it stop after a few seconds of silence?
The Web Speech API automatically pauses recognition after a period of silence as a browser-level behaviour. Click Resume to continue from where you left off — your transcript is preserved.
How do I improve transcription accuracy?
Speak clearly at a moderate pace in a quiet environment. Use a good-quality microphone and position it close to your mouth. Select the correct language before starting — using the wrong language will produce garbled output. Avoid background music or TV during transcription.
Is speech to text really free with no limits?
Yes. Because the tool uses your browser's built-in Web Speech API rather than a paid third-party service, there is no cost, no account, and no artificial usage cap. You can transcribe as much as you like. The only practical constraints are your browser's own behaviour — such as pausing after long silences — and keeping the tab active with microphone permission granted.
Can I add punctuation while dictating?
Often, yes. Many recognition engines respond to spoken punctuation commands such as saying "comma", "period", "question mark", or "new line", though support varies by browser and language. If your setup does not insert punctuation automatically or via commands, the fastest approach is to dictate the words freely and add punctuation in a quick editing pass afterwards.
Can I transcribe an existing audio or video file?
This tool transcribes live speech from your microphone rather than uploaded files. If you want to transcribe a recording, you can play it aloud near your microphone, though quality will be lower than dictating directly. For accurate transcription of existing audio and video files, a dedicated file-based transcription tool designed for that purpose will give better results.
Does speech to text work on my phone?
Yes, on supported mobile browsers. It works in Chrome on Android and Safari on iOS, using the same Web Speech API as on desktop. Grant microphone permission when prompted, choose your language, and dictate. On a phone the built-in microphone is usually close enough to work well, making it a handy way to capture notes or messages hands-free while away from a keyboard.