Multi-model AI gateway β 8 providers plus local models, no system prompt injection.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent β or use 1-click editor setup below.
π‘ Paste the JSON block into your client's configuration file under mcpServers, then restart the application.
Multi-model AI gateway for MCP clients.
MCP clients like Claude Code, Claude Desktop, and Cursor are locked to their host model. Vox gives them access to every other model β Gemini, GPT, Grok, DeepSeek, Kimi, or your local Ollama β through a single chat tool.
The design is deliberately minimal: prompts go to providers unmodified, responses come back unmodified. No system prompt injection. No response formatting. No behavioral directives. The only value Vox adds is routing and conversation memory β everything else is pure passthrough.
Send a prompt, optionally attach files or images, pick a model (or let the agent pick), and get back the model's raw response. Conversation threads persist in memory via continuation_id for multi-turn exchanges across any provider β start a thread with Gemini, continue it with GPT. Threads are shadow-persisted to disk as JSONL for durability and can be exported as Markdown.
3 tools:
| Tool | Description |
|---|---|
chat | Send prompts to any configured AI model with optional file/image context |
listmodels | Show available models, aliases, and capabilities |
dump_threads | Export conversation threads as JSON or Markdown |
8 providers:
| Provider | Env Variable | Example Models |
|---|---|---|
| Google Gemini | GEMINI_API_KEY | gemini-2.5-pro |
| OpenAI | OPENAI_API_KEY | gpt-5.1, gpt-5, o3, o4-mini |
| Anthropic | ANTHROPIC_API_KEY | claude-opus-4-8, claude-sonnet-5, claude-haiku-4-5 |
| xAI | XAI_API_KEY | grok-4.5, grok-4.3 |
| DeepSeek | DEEPSEEK_API_KEY | deepseek-v4-pro |
| Moonshot (Kimi) | MOONSHOT_API_KEY | kimi-k2.6 |
| OpenRouter | OPENROUTER_API_KEY | Any OpenRouter model |
| Custom | CUSTOM_API_URL | Ollama, vLLM, LM Studio, etc. |
Vox runs as a stdio MCP server. Each client needs to know how to launch it.
Replace /path/to/vox-mcp with the absolute path to your cloned repo.
Or add to .mcp.json in your project root:
Add to claude_desktop_config.json:
macOS: ~/Library/Application Support/Claude/claude_desktop_config.json
Windows: %APPDATA%\Claude\claude_desktop_config.json
Add to .cursor/mcp.json (project) or ~/.cursor/mcp.json (global):
Add to ~/.codeium/windsurf/mcp_config.json:
The canonical stdio configuration:
Tips:
.env.env file in the vox-mcp directory is loaded automatically, so API keys can go there instead of in the client configVOX_FORCE_ENV_OVERRIDE=true in .env if client-passed env vars conflict with your .env valuesCopy .env.example to .env and configure:
DEFAULT_MODEL β auto (default, agent picks) or a specific model nameGOOGLE_ALLOWED_MODELS, OPENAI_ALLOWED_MODELS, etc.CONVERSATION_TIMEOUT_HOURS β thread TTL (default: 24h)MAX_CONVERSATION_TURNS β thread length limit (default: 100)See .env.example for the full reference.
See CONTRIBUTING.md for code style, project structure, and how to add providers.
Apache 2.0 β see LICENSE and NOTICE.
Derived from pal-mcp-server by Beehive Innovations.
No reviews yet β be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/vox-mcp)<a href="https://allmcps.com/mcp/vox-mcp"><img src="https://allmcps.com/api/badge/vox-mcp?style=directory" alt="Vox MCP on AllMCPs" /></a>