Skip to main content
Glama

noisy-coding

Talk to Claude Code while it works — Jarvis-style voice coding. It's your voice that's noisy, not your code.

Docker Pulls Release CI Last commit License: MIT

Claude speaks short summaries aloud. An always-on listener turns your speech into messages Claude receives mid-task, without stopping it — no push-to-send, no copy-pasting transcripts. Step away from the keyboard and keep steering your agent.

Why you'll like it

  • Interrupt-free flow — speak while Claude is working; your words land in the running session, not a text box.

  • Hands-free reviews — Claude reads its findings aloud; you answer from across the room.

  • Live "tactical HUD" dashboard — conversation log with replay/recall, real-time oscilloscope, mute buttons, costs and latencies at a glance.

  • Per-agent character — voice, speed, and personality dials for every agent, all from the dashboard.

  • Nothing to configure in files — API key, devices, language, push-to-talk: everything lives in the UI and persists.

  • Never talks over you — one voice at a time; speech you missed parks as UNHEARD and a CATCH UP button replays it.

Speech-to-text and text-to-speech run on the Grok (xAI) Voice API — extremely cheap in practice (a small one-time budget lasts months of daily use).

Related MCP server: Elba MCP Server

Install in 2 minutes

The backend ships as a hardware-free Docker image (noisy/noisy-coding): the dashboard browser tab is the microphone and the speaker. You need Docker and a browser — no Python, no git, no environment variables.

# terminal: marketplace + plugin in one line
claude plugin marketplace add noisy/noisy-coding && claude plugin install noisy-coding@noisy
# inside Claude Code (new session):
/noisy-coding:setup

The setup command starts the published image and walks you through first contact. Then finish in the browser at http://127.0.0.1:8765: paste your xAI API key (console.x.ai) and click the amber ENABLE TAB AUDIO banner — that one click makes the tab your microphone and speaker. Keep the tab open and just talk.

Prefer staying inside Claude Code? Same thing, four commands: /plugin marketplace add noisy/noisy-coding/plugin install noisy-coding@noisy/reload-plugins/noisy-coding:setup.

Other setups — plain Docker without the plugin, native install with hardware mic/speakers, remote hosts, all configuration knobs — live in docs/INSTALL.md.

How it works

All speech logic lives in one listener daemon — the single owner of the microphone, the playback queue and the speakers. The MCP server is a thin messenger that forwards speak requests; Claude Code hooks deliver your transcribed speech back into the session (see docs/hooks.md).

mic (hardware or browser tab via WS :8766)
  -> VAD -> Grok STT -> transcript queue -> HTTP :8765
                              ^ polled by Claude Code hooks
speak (MCP, stdio or HTTP :8767) -> POST /speak -> daemon queue
  -> Grok TTS -> speakers (hardware or browser tab)

Tools

Tool

What it does

speak(text, interrupt?)

Queues text for speech and waits until it has played. Voice/speed/language come from the daemon (dashboard character), not the call.

announce(text)

Fire-and-forget variant: returns immediately, plays in the background.

change_voice(voice_id)

Deliberately switches this agent's voice (persists, shows on the dashboard).

list_voices()

Lists the built-in Grok voices (ara, eve, leo, rex, …).

Docs

License

MIT © Krzysztof Szumny

A
license - permissive license
-
quality - not tested
A
maintenance

Maintenance

Maintainers
<1hResponse time
1dRelease cycle
17Releases (12mo)
Commit activity
Issues opened vs closed

Related MCP Servers

  • A
    license
    -
    quality
    D
    maintenance
    Enables bidirectional voice interaction for Claude Code using local speech-to-text and text-to-speech models optimized for Apple Silicon. It provides tools to listen to user speech via microphone and speak responses aloud through system speakers.
    16
    Apache 2.0
  • A
    license
    A
    quality
    D
    maintenance
    Local speech-to-text transcription using Microsoft's VibeVoice-ASR model with speaker diarization, enabling audio transcription directly in AI tools like Claude Code, Cursor, and OpenCode.
    3
    2
    MIT
  • -
    license
    -
    quality
    -
    maintenance
    A voice-enabled interface for Claude Desktop that supports speech-to-text input and text-to-speech output via ElevenLabs, turning Claude into a voice assistant.
    1

View all related MCP servers

Related MCP Connectors

  • Real-time chat hub for AI agents — Claude Code, Cursor, Cline, Codex over MCP or REST.

  • Coding agents from Claude Code, Cursor and Codex claim jobs and lock files on one shared board.

  • Cross-agent artifact workspace with provenance across Claude Code, Codex, Cursor, LangGraph.

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/noisy/noisy-coding'

If you have feedback or need assistance with the MCP directory API, please join our Discord server