Text-to-Speech

Generate TTS audio and manage voices from your MCP client.

llms.txt Full docs for LLMs

Browse voices, generate audio, and check job status. Supports both ElevenLabs and Stealth TTS providers.

TOOL list_voices

Browse and search available TTS voices. Returns voice IDs, names, previews, and metadata. Set stealth=true for Stealth provider voices.

Example prompt
"Show me young male American English voices"
ParameterTypeDefaultDescription
searchstringnullSearch by voice name or labels
genderstringnullFilter: male, female, or neutral
agestringnullFilter: young, middle_aged, or old
languagestringnullLanguage code (e.g. en, es, fr)
accentstringnullFilter by accent (e.g. american, british)
sortstringtrendingSort: trending, created_date, usage_character_count_1y
page_sizeinteger30Results per page (max 100)
pageinteger0Page number (0-indexed)
stealthbooleanfalseUse Stealth voices endpoint instead of ElevenLabs
minimaxbooleanfalseUse MiniMax voices endpoint (cloned MiniMax voices only)
TOOL generate_tts

Generate text-to-speech audio. Submits a job, polls until complete, returns the permanent hosted audio URL.

Example prompt
"Generate a voiceover using the voice 'Adam' saying: Welcome to today's video about productivity tips"
ParameterTypeDefaultDescription
scriptstringrequiredText to convert to speech (max 200,000 characters)
voice_idstringrequiredVoice ID from list_voices, or voice name for Stealth provider
providerstringelevenlabsTTS provider: elevenlabs, stealth, or minimax
model_idstringnullElevenLabs model (e.g. eleven_multilingual_v2, eleven_flash_v2_5)
stabilitynumbernullVoice consistency 0.0-1.0 (ElevenLabs only)
similarity_boostnumbernullVoice match accuracy 0.0-1.0 (ElevenLabs only)
stylenumbernullStyle exaggeration 0.0-1.0 (ElevenLabs only)
speednumbernullPlayback speed 0.7-1.2 (ElevenLabs) or 0.5-2.0 (MiniMax)
temperaturenumbernullExpressiveness — higher is more expressive (Stealth only)
speaking_ratenumbernullSpeaking speed multiplier (Stealth only)
stealth_modelstring1.5Stealth model tier (Stealth only). 1.5 = standard model (1× characters, default); 2.0 = Stealth 2.0, our newest, most capable model (2× characters).
pitchintegernullPitch shift -12..+12 (MiniMax only)
volumenumbernullVolume 0.0-10.0 (MiniMax only)
voice_namestringnullHuman-readable voice label for your reference
custom_titlestringnullCustom filename for output MP3 (without extension)
srt_formatstringoffSRT transcript format (ElevenLabs only). 'off' = no SRT, 'default' = original multi-word phrases as returned by forced alignment, '1' / '2' / '3' / '5' = N words per caption (social-video style), 'sentence' = one full sentence per caption. Non-'off' values request `generate_srt=true` upstream and bill 1.2× the script character count (the SRT pass runs an extra forced-alignment step); 'off' bills 1×. 1/2/3/5/sentence additionally fetch the upstream SRT and regroup it. Mirrors the web UI's 'Generate SRT Transcript' preset.
TOOL get_tts_job_status

Check the status of a TTS generation job. Use this to retrieve the audio_url for a previously submitted job.

ParameterTypeDefaultDescription
job_idstringrequiredJob ID returned from generate_tts
TOOL list_tts_jobs

List your recent TTS generation jobs, sorted by creation time (newest first). No parameters required.

Example prompt
"Show me my recent TTS generations"
TOOL clone_voice

Clone a voice using the Stealth voice engine or MiniMax. Upload an audio sample and get a reusable voice ID. Requires an active subscription.

Example prompt
"Clone this voice sample as 'My Narrator' using the audio at https://example.com/sample.mp3"
ParameterTypeDefaultDescription
display_namestringrequiredDisplay name for the cloned voice
audio_urlstringrequiredURL to the audio sample (MP3, WAV, M4A, OGG, WEBM; max 15MB)
lang_codestringEN_USLanguage code (e.g. EN_US, ES_MX, FR_FR, DE_DE, JA_JP)
descriptionstringnullDescription of the voice
providerstringstealthVoice engine: stealth or minimax
voice_namestringnullDisplay name (MiniMax only — use display_name for Stealth)
languagestringEnglishLanguage tag (MiniMax only) — e.g. English, Spanish, French. See full list at /docs.
need_noise_reductionstringtrueStrip background noise (MiniMax only). Values: true or false.
TOOL delete_voice

Delete a cloned Stealth or MiniMax voice by ID.

ParameterTypeDefaultDescription
voice_idstringrequiredID of the cloned voice to delete
providerstringstealthVoice engine: stealth or minimax
esc
×

Buy More Credits

Boost
150
credits
$4.99
Good for:
  • ~12 min caption removal
Studio
1,000
credits
$31.49
Good for:
  • ~80 min caption removal