What this MCP does
Converts text and scripts to spoken audio, supports multi-voice and timed tracks, lists voices, and transcribes audio for eligible accounts.
Tools
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'properties': {}}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['script'], 'properties': {'script': {'type': 'string', 'maxLength': 25000, 'minLength': 1, 'description': "Lines like 'Abd: Tonight we cook shakshuka.' one per line. Names before the colon."}, 'voices': {'type': 'object', 'description': 'Speaker name to FreeTTS voice id, like {"Abd": "en-US-AndrewNeural", "Guest": "en-US-EmmaNeural"}. Unnamed speakers get a voice each.', 'additionalProperties': {'type': 'string'}}, 'language': {'type': 'string', 'description': 'Language of the lines when no voices are given.'}, 'pause_ms': {'type': 'integer', 'maximum': 5000, 'minimum': 0, 'description': 'Silence between lines in milliseconds. Default 600.'}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'properties': {'limit': {'type': 'integer', 'maximum': 200, 'minimum': 1, 'description': 'How many to return. Default 15.'}, 'gender': {'enum': ['female', 'male', 'any'], 'type': 'string', 'description': 'Filter by gender.'}, 'search': {'type': 'string', 'description': "A word to match in the voice name or personality, like 'Andrew', 'calm' or 'news'."}, 'language': {'type': 'string', 'description': "Language or locale, like 'German', 'de-DE', 'Spanish (Mexico)' or 'en'. Empty lists every language."}, 'free_only': {'type': 'boolean', 'description': 'Only voices that work without a FreeTTS key.'}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['script'], 'properties': {'speed': {'type': 'integer', 'maximum': 30, 'minimum': -30, 'description': 'Percent slower or faster for the main lines. Default 0.'}, 'script': {'type': 'string', 'maxLength': 60000, 'minLength': 1, 'description': 'The script text. See freetts.org/text-to-speech-for-actors for the syntax.'}, 'my_part': {'type': 'string', 'description': 'In a scene, the character you play: their lines become timed gaps for you to say them.'}, 'count_in': {'enum': ['none', '3', '5', '10'], 'type': 'string', 'description': 'Spoken count-in before each track, one number a second. Default none.'}, 'cue_voice': {'type': 'string', 'description': 'Voice for CUE: lines and [bracketed] words. Default Aria.'}, 'track_end': {'enum': ['none', 'beep', 'end'], 'type': 'string', 'description': 'After the last line of each track: nothing, a long beep, or the word End. Default none.'}, 'line_voice': {'type': 'string', 'description': 'Voice for the main lines: a steady voice name or id (Guy, Andrew, Ava, Aria, Jenny, Brian, Emma, Sonia, Ryan, Libby, Natasha, William, Clara, Liam...). Default Guy.'}, 'sentence_pause_ms': {'type': 'integer', 'maximum': 3000, 'minimum': 0, 'description': 'Silence after each sentence. Default 500.'}, 'paragraph_pause_ms': {'type': 'integer', 'maximum': 5000, 'minimum': 0, 'description': 'Silence after each paragraph. Default 1000.'}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['language'], 'properties': {'use': {'type': 'string', 'description': 'What the audio is for, in a few words.'}, 'gender': {'enum': ['female', 'male', 'any'], 'type': 'string'}, 'language': {'type': 'string', 'description': "Language or locale, like 'English (US)', 'de-DE' or 'Portuguese (Brazil)'."}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['text'], 'properties': {'text': {'type': 'string', 'maxLength': 30000, 'minLength': 1, 'description': 'The text to read. Plain text; SSML is not needed.'}, 'speed': {'type': 'integer', 'maximum': 30, 'minimum': -30, 'description': 'Percent slower (negative) or faster (positive). Default 0.'}, 'voice': {'type': 'string', 'description': "A FreeTTS voice id from list_voices, like en-US-JennyNeural (a short name like 'Jenny' also works). If empty, a good standard voice for the language is chosen."}, 'format': {'enum': ['mp3', 'wav'], 'type': 'string', 'description': 'mp3 (default) or wav (FreeTTS PRO).'}, 'language': {'type': 'string', 'description': "Language of the text when no voice is given, like 'German' or 'pt-BR'."}, 'include_audio': {'type': 'boolean', 'description': 'Also return the audio bytes in the result (large). Default false; the link is enough for most clients.'}}, 'additionalProperties': False}
Input schema
{'type': 'object', '$schema': 'http://json-schema.org/draft-07/schema#', 'required': ['audio_url'], 'properties': {'language': {'type': 'string', 'description': "Language of the recording, like 'en-US', 'de-DE', 'Arabic' or 'es', or 'auto' to detect it (default)."}, 'audio_url': {'type': 'string', 'format': 'uri', 'description': 'Public https URL of the audio file.'}}, 'additionalProperties': False}
Recent tool changes
Similar MCP servers
Andreax
Offers pay-per-call AI services for inference, agent and workflow design, OCR and transcription, code generation and review, clas…
Social Fetch
Retrieves public content and metadata from social networks, music services, Facebook Marketplace, events, profiles, posts, commen…
Hermoso
Supports ad research, competitor analysis, video and static ad creation, campaign policy checks, content publishing, and performa…
ElevenLabs
Manages ElevenLabs speech, voices, pronunciation dictionaries, podcasts, dubbing, audio projects, workspaces, and related media r…
John's Essentials
Analyzes and converts documents, images, audio, video, archives, data files, and other media formats, with PDF chat and file insp…
MiOffice — AI-Powered Workspace Studio
Provides browser-based tools for processing PDFs, images, video, and audio, including generation, enhancement, conversion, transc…
BlitzReels Video Editor
Creates and edits short-form videos, including timelines, captions, transitions, AI-generated visuals, music, voiceovers, sound e…
Heista
Analyzes video advertising, builds brand intelligence profiles, researches markets, and generates advertising scripts and creativ…