murmur

A whole radio station, for an audience of one.

murmur is a local-first companion radio with an agent for a brain. It finds something to talk about, puts on a song, and comes back to keep going. Type back whenever you like, and it answers in a voice that sounds human.

$ git clone https://github.com/wine-fall/murmur
$ cd murmur && make dev   # the radio comes on the air

local-first open-source, MIT keyboard in, voice out

Most AI right now is selling something: marketing, productivity, content. murmur isn’t trying to make you faster. It was built for one thing: to be AI that sits closer to the person. A voice in the room while you work. A host who stays. A rapport that grows.

— why this exists

It speaks first.

Nobody presses play. It picks a topic and starts talking, slides into a song, comes back and keeps going. Morning and night get their own greetings; when you step away, it notices and goes quiet on its own.

It repairs by talking.

Missing ffmpeg? No voice endpoint yet? The radio still launches anyway, degraded, then names its own gaps on air and offers to fix them, asking before every change. Setup is a conversation, not a README.

You cut in, it decides.

Your reply doesn’t land in a form field; it lands in an agent’s hands. Ask for something quieter and the music switches mid-song; say you’re done for tonight and it signs off properly, last words and all.

The rest of the station

It remembers you

Three tiers of memory persist across sessions: who you are, what you’ve talked about, which songs already played. The host’s character is a plain text file, yours to edit.

No dead air

While one segment plays, the next is already being written and voiced. Talk hands over to music and back without a gap.

A voice that sounds human

Sentence-paced speech with a pinned timbre, riding over the music with proper ducking: it lowers the song to talk, and never cuts it.

Terminal-native

A TUI with a live music visualizer and a pixel pet; without bun it falls back to plain text and keeps broadcasting.

Three network calls, total

Model inference for the brain, a stream for the songs, a hosted voice for now, until a local TTS lands. Logic, memory, and mixing stay on your machine.

Accepted by ear

Open-source under MIT, with a hard bar: if it doesn’t sound human, it isn’t done. The remaining work is listening.

On the air today

Every spec on the roadmap is built: the loop, the voice, the memory, the pacing, the steering. Most of what’s open is listening: by-ear passes on how it feels through a real day. Clone it and it talks.

Read the source

Node ≥ 24 · pnpm · your agent login (Claude Code today), or run fully offline with --brain stub.