The full upstream README, mirrored here for reference. Install config, tool schemas, adoption signals, and an original overview live on the MCP Fish listing page.
A Model Context Protocol (MCP) server for Fish Audio TTS (Text-to-Speech) via the AceDataCloud platform. Generate natural-sounding speech and explore the Fish voice model library.
Pass one public HTTPS reference audio URL plus its exact transcript. This conditions only the current TTS request and does not create a reusable voice model:
Use reference_id for saved or public voices, and the one-shot reference fields for a temporary voice. Do not combine them. Reference audio supports MP3/WAV and should be 10–270 seconds. Billing remains based on the target text's UTF-8 byte count.
Set your AceDataCloud API token:
Get your token from https://platform.acedata.cloud.
| Tool | Description |
|---|---|
fish_generate_audio | Generate speech from text via a Fish voice model |
fish_list_models | List available Fish voice models |
fish_get_model | Fetch metadata for a specific Fish voice model |
fish_get_task | Get the status / result of a generation task |
fish_get_tasks_batch | Batch-fetch the status / result of multiple tasks |
fish_get_usage_guide | Get the API usage guide |
MIT — see LICENSE at the repository root.