Speech to Text

Convert your voice to text instantly in 16 languages — free, browser-based.

Developer ToolsFreeNo Signup
Speech to Text
Free Tool

How to use Speech to Text

Our Speech to Text tool converts your spoken words into typed text directly in your browser using the Web Speech API — no audio ever leaves your device, no account needed, completely free. Before starting, ensure your browser has microphone permission. Chrome and Edge offer the best Web Speech API support. Firefox has limited support. The first time you use it, your browser will ask for microphone access — click Allow. Select your language from the 16 supported options: English (US), English (UK), Spanish, French, German, Italian, Portuguese, Hindi, Japanese, Korean, Chinese (Simplified), Chinese (Traditional), Arabic, Russian, Dutch, and Polish. The language selection determines which speech recognition model is used. Choose your listening mode: • Single Utterance — listens for one sentence or phrase, then stops automatically. Best for short, deliberate inputs like filling forms or searching. • Continuous Mode — keeps listening and appending text until you manually stop. Best for dictation, note-taking, and transcribing long speech. Click "Start Listening" — the status indicator shows when the microphone is active. Speak clearly at a normal pace. The confidence score shows how certain the recognition engine is about the transcribed words — higher is better. Your transcription appears in real time. The tool auto-saves to browser localStorage so your text persists if you accidentally close the tab. Use the "Copy" button to copy the full transcript to your clipboard, or "Download" to save a .txt file. Tips for better accuracy: use a quality microphone, minimize background noise, speak at conversational pace (not too fast or slow), and use punctuation by saying "period," "comma," "question mark," or "new paragraph." This tool is ideal for hands-free note-taking, accessibility needs, quick voice-to-document workflows, and anyone who types faster by speaking.

Frequently Asked Questions

Which browsers support speech to text?

Google Chrome and Microsoft Edge have the best Web Speech API support. Safari on Mac and iOS also supports it. Firefox has limited or no support depending on the version. For the most reliable experience, use Chrome — it powers the speech recognition behind most browser-based voice tools.

Is my voice data private?

The Web Speech API sends audio to Google's speech recognition servers for processing (in Chrome and Edge). The audio is processed and discarded — it is not stored permanently. If privacy is critical, consider an offline speech recognition solution. This tool never stores your audio or transcript on our servers.

What languages are supported?

This tool supports 16 languages: English (US and UK), Spanish, French, German, Italian, Portuguese (Brazil), Hindi, Japanese, Korean, Chinese Simplified and Traditional, Arabic, Russian, Dutch, and Polish. Accuracy varies by language — English, Spanish, and French have the most training data and typically perform best.

Can I add punctuation by speaking?

Yes. In most languages you can say punctuation marks aloud. Say "period" or "full stop" for a period (.), "comma" for a comma (,), "question mark" for (?), "exclamation mark" for (!), and "new paragraph" or "new line" to start a new line. Results vary by language and browser.

What is continuous mode and when should I use it?

Continuous mode keeps the microphone active and keeps appending text until you stop it. Use it for dictating long documents, taking notes in meetings, or transcribing recorded speech. Single utterance mode is better for one-at-a-time inputs — it stops after detecting the end of a sentence, reducing accidental transcription of background noise.

Recommended

Related Tools