Voice Mcp

Voice Mcp is an MCP server that enables bi-directional voice conversations with Claude Code and other Model Context Protocol agents. It interfaces with local speech processing tools such as Whisper.cpp and Kokoro, or cloud endpoints using an OpenAI API key for speech-to-text and text-to-speech services. Software engineers, terminal users, and multitasking developers use this tool when typing is inconvenient or hands-free operation is required, such as stepping away from the desk or debugging during manual tasks. By capturing audio input through the microphone and returning generated audio directly through system speakers, Voice Mcp allows users to speak prompts and listen to model responses. It features smart silence detection that automatically halts recording when the speaker pauses, maintaining conversational flow. Users can run the service completely offline with local machine learning models for privacy, or configure it with remote OpenAI APIs as a fallback, all managed through standard MCP server configurations.

Category: Design, Media & Creative

Tags: audio, speech, tts, voice

Visit Voice Mcp

How to install and configure Voice Mcp

Follow these steps to set up Voice Mcp using the uv tool or Claude Code plugin: 1. Install platform audio dependencies, such as ffmpeg, portaudio, and audio libraries for your operating system (e.g., via brew on macOS or apt on Ubuntu). 2. If using Claude Code plugin mode, run: bash claude plugin marketplace add mbailey/voicemode claude plugin install voicemode@voicemode /voicemode:install 3. Alternatively, install via the Python package installer: bash curl -LsSf https://astral.sh/uv/install.sh | sh uvx voice-mode-install claude mcp add --scope user voicemode -- uvx --refresh --from voice-mode voicemode-mcp-launcher 4. If using cloud speech services instead of local Whisper or Kokoro, set your API key by running export OPENAI_API_KEY=your-openai-key. 5. Start conversational voice mode by executing claude converse or /voicemode:converse.

What you can do with Voice Mcp

  • Conversing with Claude Code hands-free while away from the keyboard, walking, or cooking - Hearing spoken responses from coding assistants to reduce continuous visual strain from screen time - Running speech-to-text and text-to-speech completely offline using local Whisper.cpp and Kokoro tools - Interacting verbally with MCP-capable terminal agents when holding devices, pets, or physical items - Automatically capturing dictated prompts and halting recordings using smart silence detection features

Key facts

  • https://github.com/mbailey/voice-mcp
  • Design, Media & Creative
  • audio, speech, tts, voice

Part of MCP Servers

Related MCP servers

  • MCP MD2PDF Server — MCP MD2PDF Server is an MCP server that enables automated conversion of Markdown documents into formatted PDF files with full…
  • MCP Media Processing Server — MCP Media Processing Server is an MCP server that connects AI assistants like Claude Desktop to local media manipulation utilities,…
  • MCP Music Analysis — MCP Music Analysis is an MCP server that enables AI clients like Claude to inspect and process sound recordings using…
  • MCP OCR Server — MCP OCR Server is an MCP server that provides optical character recognition capabilities by connecting language models directly to the…
  • MCP MiniMax Music Server — MCP MiniMax Music Server is an MCP server that provides AI-powered audio and music generation using the MiniMax Music API…
  • MCP Mermaid Server — MCP Mermaid Server is an MCP server that provides tools for generating and analyzing Mermaid diagrams directly within Model Context…

How do I install Voice Mcp?

You can install Voice Mcp as a Claude Code plugin by adding the mbailey/voicemode marketplace and running the install command. Alternatively, install the uv package manager, execute uvx voice-mode-install to set up dependencies, and register the server into Claude Code using the claude mcp add command targeting voicemode-mcp-launcher.

What can Voice Mcp do?

Voice Mcp connects MCP-capable agents like Claude Code to speech-to-text and text-to-speech engines. It records voice input through your microphone, detects silence to stop recording automatically, sends the transcribed text to the language model, and vocalizes the model's generated response back through your speakers.

Can Voice Mcp work without an internet connection?

Yes, Voice Mcp supports fully local, offline voice processing. You can set up Whisper.cpp for local speech-to-text transcription and Kokoro for local text-to-speech generation. When configured locally, no external API calls or internet connections are required, keeping your audio data private on your local machine.

Which operating systems are compatible with Voice Mcp?

Voice Mcp is compatible with Linux, macOS, Windows natively or through WSL2, and NixOS. It requires Python versions 3.10 through 3.14 along with operating system audio dependencies like ffmpeg and portaudio. WSL2 users specifically require pulseaudio packages for microphone access.

Is Voice Mcp open source?

Voice Mcp is open-source software licensed under the MIT license. You can view, inspect, contribute to, or fork its source code directly on GitHub at https://github.com/mbailey/voice-mcp.

  • AI Tools
  • Categories
  • Industries
  • CLI Coding Agents
  • MCP Servers
  • MCP Categories