Just speak.
It's already typed.

Turn your voice into finished text. Works in Slack, Gmail, Word, your terminal and any other app.

Microphone Array (Default) Stop Ctrl Space Close Esc

Ready to see how it works?

0:00 / 0:00

Speech in, text out

No commas spoken.
Every comma typed.

Nobody talks in capitals and full stops. Murmur hears the run-on and types the sentence.

Free · Windows 10 1809 or later, 64-bit Intel or AMD · Nothing leaves the machine

Works with what
you already use

There is no integration to install and no plugin to approve.
If the window takes a cursor, it takes your voice.

Everywhere

Your words don't care
which window is open

Murmur pastes into whatever had focus when you pressed the key. Pick a destination and watch the same four seconds of speech land in it.

0:04

What's inside

Everything it does.

Wi-Fi

✕ Works offline

The model sits on your disk and runs on your processor or your graphics card. No Wi-Fi, no problem.

Models library
Tiny75 MB
SmallIn use
Medium1.5 GB
Large v33.1 GB

◧ Six models, one click

From instant-but-rough to slow-and-precise, downloaded on demand and kept afterwards. All of them free, forever.

🎙
💬
✉️

✎ Punctuation included

Commas, full stops and capitals arrive without being dictated, because the decoder is a language model.Say it, don't spell it.

helloこんにちは merhabaolá привет안녕하세요 bonjourمرحبا

◍ 99 languages

Turkish included. Detection is automatic, but naming your language in the menu is 2.6x faster and cannot guess wrong. It will translate them all to English too.

+

Reminder to send the quarterly numbers to Dilek before Friday.

➤

⧉ Pastes itself

No copying, no switching window. The transcript goes straight into the field the cursor was already in.Your hands never leave the work.

Graphics card0.85s
Processor2.6s
Speaking it4.0s

⚡ Faster than you spoke

4.7× realtime on an RTX 5070 Ti with the recommended model; 1.5× on the processor alone. The text is waiting before you are.

17:42 · 1.2s
Ship the installer before the demo.
17:39 · 0.9s
Ask about the Blackwell kernels.
17:31 · 1.4s
Move the pack download to the Speed page.

↺ Everything you said

The History tab keeps the session's transcripts so you can grab one again. Gone when you quit, unless you ask it to keep them.

🗎

File transcription

Point it at a recording with --file and get the text, no microphone needed.

☝

Press to talk Default

Press once to start, once again to stop and paste. Esc throws the take away.

⌨

Your own shortcut

Ctrl+Space to begin with, changed in the setup guide and kept in config.json.

◐

Silence gate

A voice-activity filter runs first, so an empty room never produces an invented sentence.

Integrations

Works anywhere
you can type

Slack, Cursor, Notion, Word, your terminal.
You say it, Murmur puts it there.

Murmur
Models library
Bigger is more accurate, smaller is faster.
SmallIn use
Distil Large v3Download
Large v3Download
Speed
Where transcription happens
Graphics card (CUDA)Ready
ProcessorFallback
GPU pack · 1.6 GBInstalled
Shortcut
Press the keys you want
Ctrl Space
DiscardEsc
QuitF10

Set it up once

Make it yours

Five screens on first run, then a window you rarely need to open again.

⚙

Pick the trade-off yourself

Six models between instant and immaculate. Switch whenever the job changes; the download is kept.

⚡

Turn the graphics card on

The installer ships the processor build. One button on the Speed page adds the CUDA libraries, and it needs no administrator rights.

⌨

Choose your own keys

Ctrl+Space is only the default. Per-application paste overrides are there for consoles that want Ctrl+Shift+V.

Beep boop

Faster workflows

Use Murmur with Claude Code, Cursor, Codex or any other agent that takes a prompt, without touching your keyboard.

⚙

Works where your code is

Any editor or terminal that accepts a paste, which is all of them.

⌁

You build more and type less

People speak roughly three times faster than they type. Give the model the whole picture instead of the short version.

Windows Terminal

Models

Pick your trade-off

Bigger is more accurate, smaller is faster. They download on demand and are kept afterwards.

ModelSizeLanguagesBest for
Tiny75 MB99Short commands and clear speech in a quiet room.
Base145 MB99A sensible floor on machines with no graphics card.
SmallRecommended480 MB99Everyday dictation. Text comes back before you've thought of the next sentence.
Distil Large v31.5 GBEnglishThe best quality-per-second here, if you only dictate in English.
Medium1.5 GB99The multilingual step up. Better with names, jargon and noisy rooms.
Large v33.1 GB99Strong accents and difficult audio. The slowest of the set.

Get it

Free, and it stays free

Open source, no subscription, no account. The speech models are MIT licensed and cost nothing to run.

Windows will say it protected your PC - the installer is not signed. Choose More info → Run anyway.

Windows 10 version 1809 or later 64-bit Intel or AMD - not ARM64 8 GB RAM A microphone A graphics card makes it quicker, but is not required

Questions

Does it work in every application?

Anywhere you can paste. Two exceptions worth knowing: windows running as Administrator will not accept the paste, because Windows blocks input across privilege levels, and a few older consoles want Ctrl+Shift+V instead of Ctrl+V, which you can set per application.

How fast is it really?

On a recent NVIDIA laptop GPU the recommended model transcribes four to five times faster than you speak, so a four second sentence returns in under a second. Without a graphics card, expect roughly real time on the smaller models.

Do I pay for the speech models?

No. Every model is MIT licensed and publicly hosted. You download once and run locally forever, with no per-minute charge and no account.

Windows says it protected my PC. Is something wrong?

No, and it will say that. The installer is not code-signed, so SmartScreen warns about it and puts the run button behind More info → Run anyway. The warning means the file has no certificate and few people have downloaded it yet, not that anything was found in it, a certificate costs a few hundred pounds a year, which a free tool does not have. If you would rather not take that on trust, the whole program is on GitHub and you can run it from source instead.

Does it speak Turkish?

Yes. Every model except Distil Large v3 covers 99 languages and will detect yours automatically, but name it if you can. Detection reads the opening seconds and commits, and a wrong guess does not produce a translation, it produces confident nonsense, because the decoder picks the wrong vocabulary and then writes fluently in it. Naming the language also skips the detection pass: on the same eight-second English clip here, 0.73s named against 1.88s detected, 2.6x faster. It is in the tray and pill menus as well as on the Language page, so switching between two languages is a right-click, not a trip into settings.

Is any audio stored?

No. Audio exists in memory only for the seconds between pressing the shortcut and the text appearing, then it is discarded. There is no code path in Murmur that writes a recording to a file.

So what does it write down?

Four files, all in %LOCALAPPDATA%\Murmur and none of them sent anywhere. config.json is your settings. stats.json is totals, words, seconds, and which applications you dictated into; no transcript and no text of any kind reaches it. murmur.log is startup and error output, kept because the installed build is a windowed program with no console to print to; it holds no transcripts, but the file paths in it contain your Windows username, so read it before pasting it into a bug report. history.json only exists if you tick keep these after Murmur closes on the History page, which is off by default; it holds the last 300 transcripts as plain text, and the Clear button on that page empties it. Leave the tick off and your transcripts are gone the moment Murmur quits.