@mindstone/mcp-server-elevenlabs

ElevenLabs MCP server for Model Context Protocol hosts. Generate speech, music, and sound effects, browse voices, and transcribe audio using the ElevenLabs API through a standardised MCP interface.
Status
- Version: 0.5.2 ยท npm
- Auth: API key (
ELEVENLABS_API_KEY)
- Tools: 32 (account, usage, voices, speech, music, transcription, voice conversion, isolation, alignment, cloning, dialogue, voice design, dubbing, history, pronunciation dictionaries)
- Surface: cloud-api
- Machine-readable:
STATUS.json
Requirements
One-click install

After clicking the button, your host will prompt you to fill: ELEVENLABS_API_KEY.
Manual config for Claude Desktop / Claude Code / Goose / Continue.dev (ElevenLabs)
{
"mcpServers": {
"ElevenLabs": {
"command": "npx",
"args": [
"-y",
"@mindstone/mcp-server-elevenlabs"
],
"env": {
"ELEVENLABS_API_KEY": ""
}
}
}
}
Quick Start
Install & build
cd <path-to-repo>/connectors/elevenlabs
npm install
npm run build
npx (once published)
npx -y @mindstone/mcp-server-elevenlabs
Local
Configuration
Environment variables
ELEVENLABS_API_KEY โ ElevenLabs API key (starts with sk_)
MCP_HOST_BRIDGE_STATE โ optional path to a host bridge state file used for credential management
MINDSTONE_REBEL_BRIDGE_STATE โ backwards-compatible alias for MCP_HOST_BRIDGE_STATE
Host configuration examples
Claude Desktop / Cursor
{
"mcpServers": {
"ElevenLabs": {
"command": "npx",
"args": ["-y", "@mindstone/mcp-server-elevenlabs"],
"env": {
"ELEVENLABS_API_KEY": "your-api-key"
}
}
}
}
Local development (no npm publish needed)
{
"mcpServers": {
"ElevenLabs": {
"command": "node",
"args": ["<path-to-repo>/connectors/elevenlabs/dist/index.js"],
"env": {
"ELEVENLABS_API_KEY": "your-api-key"
}
}
}
}
Configuration
configure_elevenlabs_api_key โ Save your ElevenLabs API key
Account & discovery (FREE)
check_subscription โ Check subscription tier and character credit usage
get_usage_stats โ Credit usage over time grouped by product/model/voice (workspace analytics API)
list_models โ List TTS models with languages and capabilities
Voices
list_voices โ Search and browse voices on your account
get_voice โ Get full details for one voice by voice_id
search_shared_voices โ Search the public shared voice library (filters include accent)
clone_voice โ Create an instant voice clone from local audio samples (destructiveHint)
delete_voice โ Permanently delete a voice (destructiveHint)
design_voice โ Generate voice-design previews from a text description (slow; previews saved under the workspace)
create_voice_from_preview โ Save a design preview as a permanent voice (destructiveHint)
Speech & conversion
generate_speech โ Generate spoken audio from text using text-to-speech (supports seed and pronunciation_dictionary_locators)
generate_speech_with_timestamps โ Generate speech with character-level timing; also writes an .srt subtitle file and alignment JSON
generate_sound_effect โ Generate sound effects from a text description
speech_to_speech โ Convert source audio to a different voice
text_to_dialogue โ Multi-voice dialogue from a script (one voice per line)
History (FREE)
list_history โ List previously generated audio items (find that voiceover from last week)
get_history_item_audio โ Re-download the audio of a past generation by history_item_id
Pronunciation dictionaries
list_pronunciation_dictionaries โ List pronunciation dictionaries (brand names, jargon) (FREE)
get_pronunciation_dictionary โ Get one dictionary's metadata and current rules (FREE)
add_pronunciation_dictionary โ Create a dictionary from alias/IPA rules (destructiveHint)
archive_pronunciation_dictionary โ Archive a dictionary so it is no longer applied (destructiveHint; reversible in the dashboard)
Audio processing
isolate_audio โ Remove background noise from an audio file (source must be โฅ ~4.6s; shorter clips fail upstream)
forced_alignment โ Align transcript text to audio with per-word timestamps
Dubbing (v1 API โ async submit โ poll โ download)
create_dubbing โ Submit a dubbing job (local file via sandbox or source_url for ElevenLabs-side fetch)
get_dubbing โ Poll job status until dubbed, failed, or cancelled
download_dubbed_audio โ Download dubbed audio (Content-Type sniffed)
delete_dubbing โ Permanently delete a dubbing job (destructiveHint)
Music
generate_music โ Generate music from a text prompt
create_music_plan โ Create a composition plan for music generation (free)
generate_music_from_plan โ Generate music from a composition plan
Transcription
transcribe_audio โ Transcribe speech from an audio file to text, with optional speaker diarization (diarize, num_speakers, diarization_threshold) and word-level timestamps (include_word_timestamps); scribe_v1 and scribe_v2 models
Local file paths for upload tools must be inside MCP_WORKSPACE_PATH (or os.tmpdir() when unset). See src/tools/file-input.ts. Generated and downloaded files (speech audio, subtitles, history/dubbing downloads, voice-design previews) are written into the same canonical workspace root with exclusive creation โ existing files are never overwritten. See src/tools/path-safety.ts.
Licence
FSL-1.1-MIT โ Functional Source License, Version 1.1, with MIT future licence. The software converts to MIT licence on 2030-04-08.