Voice to Text

Transcript stays in your browser
T
ALL TOOLS · THE CATALOG
113 single-purpose utilities · runs local
Browse all →
All tools process files entirely in your browser · Your data never leaves your device

Voice to Text — Free Live Transcription

Speak and watch your words appear in real time. Uses the Web Speech API built into Chrome, Edge, and Safari — pick from 77 languages, see in-progress phrases dimmed until they finalize, then copy the transcript or download it as a TXT file, with optional timestamps.

Your transcript never touches our servers — it lives only on this page until you copy or download it. Speech recognition runs on your browser's built-in engine (Chrome routes audio through its vendor's speech service; recent Safari versions can recognize on-device).

How to transcribe speech to text in your browser

  1. Pick your language. Choose from 16+ locales — English, Spanish, French, German, Chinese, Japanese, Hindi, Arabic, and more. You can switch mid-session.
  2. Start and allow the mic. Click Start transcribing and grant microphone permission when the browser asks.
  3. Speak naturally. In-progress phrases appear dimmed and italic; once the engine is confident, they turn into final transcript paragraphs with a timestamp.
  4. Copy or download. Stop anytime — the transcript stays. Toggle timestamps, then Copy to clipboard or download as a .txt file.

Use Cases

Dictate notes, emails, or a first draft faster than typing
Capture meeting or lecture takeaways hands-free
Transcribe your side of an interview as you speak
Practice pronunciation in a foreign language and check what the engine hears
Journal by voice and export the day's entry as a TXT file

Frequently Asked Questions

Does this work in every browser?

It needs the Web Speech API: Chrome, Edge, and Safari support it; Firefox currently does not. On unsupported browsers the page explains this and recommends Chrome or Safari.

Is my voice sent anywhere?

Your transcript stays on this page — orangebot.ai never receives your audio or text. Recognition itself is performed by your browser's built-in speech engine; in Chrome that engine processes audio on Google's speech servers, while newer Safari versions can recognize on-device.

Does it add punctuation?

Sometimes. Chrome adds automatic punctuation for English and a few other languages; most locales come back unpunctuated, so expect to add periods and commas while editing.

Why does transcription pause after silence?

Browsers end a recognition session after a stretch of silence. This tool restarts it automatically so dictation feels continuous — your previous text is always kept.

Can I transcribe an audio or video file?

No — the Web Speech API only listens to the live microphone. Play the file out loud near your mic as a workaround, or use a dedicated file-transcription service for accuracy.

How accurate is it?

Very usable for clear speech in a quiet room — typically 90%+ for major languages. Accents, background noise, and technical vocabulary lower accuracy; a good microphone helps more than anything.

Which browser transcribes where — on-device vs vendor servers

Live speech-to-text on the web runs through the Web Speech API, and the single most misunderstood thing about it is that "in the browser" does not always mean "on your device". Where the audio is processed depends entirely on which browser you opened the page in.

BrowserRecognition engineAudio leaves device?Works offline?
Chrome / Edge (desktop)Google speech serviceYes — audio is streamed to the vendorNo
Safari 16+ (macOS/iOS)Apple on-device recognitionNo, for supported languagesOften yes
FirefoxNot implementedN/A — the API is unavailableN/A
Chrome on AndroidGoogle speech service, may use on-device modelUsually yesSometimes

Related tools

More free, in-browser tools from the AI Tools set — every one runs locally, with no sign-up and no upload.

Browse all free online tools →