Media · Local MCP server
Voicemode
Natural voice conversations for AI assistants - STT/TTS via MCP.
What the MCP Registry states
The entry as published to the official MCP Registry (read 2026-10-04), latest version.
- Registry name
dev.voicemode/voicemode- Version
- 8.12.0
- Status
- Active
- Category
- media
- Transport
- stdio (local process)
- Package
- PyPI
- Published
- 2026-07-21
- Updated
- 2026-07-21
- Publisher
- dev.voicemode
- Repository
- github.com/mbailey/voicemode
- Source
- Registry API entry
Packages
| Registry | Package | Version | Transport |
|---|---|---|---|
| PyPI | voice-mode | 8.12.0 | stdio |
How to connect Voicemode
Voicemode runs locally from a Python package published to PyPI: voice-mode version 8.12.0. It speaks MCP over stdio, so the client starts it as a program and talks to it through standard input and output. It needs Python; clients usually start it with uvx (from uv) or after pip install — the usual command is uvx voice-mode. It reads these environment variables: OPENAI_API_KEY (secret), VOICEMODE_DEBUG, VOICEMODE_SKIP_TTS, VOICEMODE_PREFER_LOCAL, VOICEMODE_AUDIO_FORMAT, VOICEMODE_WHISPER_MODEL and VOICEMODE_DISABLE_SILENCE_DETECTION; set them in the client's configuration for this server.
In the mcpServers JSON format that many desktop and editor MCP clients read, the entry looks like this (placeholders in angle brackets):
{
"mcpServers": {
"voicemode": {
"command": "uvx",
"args": [
"voice-mode"
],
"env": {
"OPENAI_API_KEY": "<secret>",
"VOICEMODE_DEBUG": "<value>",
"VOICEMODE_SKIP_TTS": "<value>",
"VOICEMODE_PREFER_LOCAL": "<value>",
"VOICEMODE_AUDIO_FORMAT": "<value>",
"VOICEMODE_WHISPER_MODEL": "<value>",
"VOICEMODE_DISABLE_SILENCE_DETECTION": "<value>"
}
}
}
}Derived from the registry entry, not tested here. What the server does, and on what terms, is set by its publisher; check its repository or website before giving it access to your accounts or files. How to add an MCP server to an assistant · Before you connect
More media servers
| Server | Runs |
|---|---|
| Voice BridgeMulti-engine TTS for AI coding assistants. 5 engines, free engine included. | Local · stdio |
| VoiceLabsAI voice generation: text-to-speech and voice cloning from any MCP client. | Remote · HTTP |
| VoicelyOffline speech-to-text on macOS: transcribe audio/video files, read dictations and calls. | Local · stdio |
| VoiceMoatThe personal brand OS for Twitter/X and LinkedIn: score, improve, schedule and publish posts. | Remote · HTTP |
| VoxNative macOS MCP server for voice I/O — Swift binary, SFSpeechRecognizer + ElevenLabs TTS. | Local · stdio |
| VR.org VR / AR / XR ReaderRead-only VR / AR / XR news, full-text editorial, events, deals, and buyer guides from VR.org. | Local · stdio |
| Vreddie MCP ServerVR Eddie's headset ratings, game reviews, deals, and compatibility checks via MCP. | Local · stdio |
| VuelaAI video, image, text and audio tools from your vuela.ai account. | Remote · HTTP |