Voice Processing · AI Chatbot

MCP Fish Audio Server

da-okazaki/mcp-fish-audio-server

An MCP (Model Context Protocol) server that provides seamless integration between Fish Audio's Text-to-Speech API and LLMs like Claude, enabling natural language-driven speech synthesis.

Install

npx -y @alanse/fish-audio-mcp-server

Client configuration

{
  "mcpServers": {
    "fish-audio": {
      "command": "npx",
      "args": [
        "-y",
        "@alanse/fish-audio-mcp-server"
      ],
      "env": {
        "FISH_API_KEY": "<FISH_API_KEY>",
        "FISH_MODEL_ID": "<FISH_MODEL_ID>",
        "FISH_REFERENCE_ID": "<FISH_REFERENCE_ID>",
        "FISH_OUTPUT_FORMAT": "<FISH_OUTPUT_FORMAT>",
        "FISH_STREAMING": "<FISH_STREAMING>",
        "FISH_LATENCY": "<FISH_LATENCY>",
        "FISH_MP3_BITRATE": "<FISH_MP3_BITRATE>",
        "FISH_AUTO_PLAY": "<FISH_AUTO_PLAY>",
        "AUDIO_OUTPUT_DIR": "<AUDIO_OUTPUT_DIR>"
      }
    }
  }
}

Environment variables

FISH_API_KEYFISH_MODEL_IDFISH_REFERENCE_IDFISH_OUTPUT_FORMATFISH_STREAMINGFISH_LATENCYFISH_MP3_BITRATEFISH_AUTO_PLAYAUDIO_OUTPUT_DIRFISH_REFERENCESFISH_DEFAULT_REFERENCEFISH_REFERENCE_1_ID
Category
Voice Processing, AI Chatbot
License
MIT
Updated
Oct 6, 2026

Features

  • High-Quality TTS: Leverage Fish Audio's state-of-the-art TTS models

  • Streaming Support: Real-time audio streaming for low-latency applications

  • Multiple Voices: Support for custom voice models via reference IDs

  • Smart Voice Selection: Select voices by ID, name, or tags

  • Voice Library Management: Configure and manage multiple voice references

  • Flexible Configuration: Environment variable-based configuration

  • Multiple Audio Formats: Support for MP3, WAV, PCM, and Opus

  • Easy Integration: Simple setup with any MCP-compatible client

Details on this page are taken from the project's README. Open README

Related MCP servers

More servers
Elevenlabs MCP logo

Elevenlabs MCP

elevenlabs/elevenlabs-mcp

620

ElevenLabs official MCP server providing text-to-speech and audio processing API interaction.

Voice Processing
MiniMax MCP Server logo

MiniMax MCP Server

MiniMax-AI/MiniMax-MCP

302

Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech and video/image generation APIs. This server allows MCP clients like Claude Desktop, Cursor, Windsurf, OpenAI Agents and others to generate speech, clone voices, generate video, generate image and more.

Voice Processing
Elevenlabs MCP Server logo

Elevenlabs MCP Server

mamertofabian/elevenlabs-mcp-server

76

A Model Context Protocol (MCP) server that integrates with ElevenLabs text-to-speech API, featuring both a server component and a sample web-based MCP Client (SvelteKit) for managing voice generation tasks.

Voice Processing
Smart Pet with MCP logo

Smart Pet with MCP

shijianzhong/smart-pet-with-mcp

46

An intelligent pet companion application based on the MCP protocol, enabling interaction with virtual pets through speech recognition and natural language processing, with support for multiple platforms.

Voice Processing
Minimax MCP Tools logo

Minimax MCP Tools

PsychArch/minimax-mcp-tools

45

A Model Context Protocol (MCP) server for Minimax AI integration, providing async image generation and text-to-speech with advanced rate limiting and error handling.

Voice Processing
Speech Interface (Faster Whisper) logo

Speech Interface (Faster Whisper)

kvadratni/speech-mcp

33

Speech MCP provides a voice interface for Goose, allowing users to interact through speech rather than text. It includes.

Voice Processing