create_text_to_speech_generation
Create Speech Generation.
This record as markdown: /tools/aiwerk-mcp-server-elevenlabs/create-text-to-speech-generation.md
What create_text_to_speech_generation does on Elevenlabs
AI agents use create_text_to_speech_generation to create or update resources in Elevenlabs, usually the action step of a workflow, after the agent has gathered context. Every call changes real data in your Elevenlabs environment.
| Parameter | Type | Required | Description |
|---|---|---|---|
text | string | — | The text to synthesize into speech. |
voice | string | — | The ID of the voice to speak with. |
webhook | object | — | |
model_id | string | — | The model to use for the generation. |
language_code | string | null | — | |
output_format | string | — | The audio encoding of the output, as `codec_sampleRateHz_bitrateKbps`. `mp3_44100_192` requires the Creator tier or above. |
voice_settings | object | — | Overrides for the voice's saved settings, applied to one generation. |
pronunciation_dictionary_locators | array | — | Pronunciation dictionaries to apply to the text, in order of precedence. Up to 3. |
Parameters from the server's own tool schema.
Why create_text_to_speech_generation is rated Medium
An AI agent can call create_text_to_speech_generation faster than any human can review: one bad instruction and it creates or modifies resources in Elevenlabs by the hundred, each call as confident as the last.
Risk signalsAccepts URL/endpoint input (webhook)
Attacks that exploit this kind of access
The rule that runs create_text_to_speech_generation safely
PolicyLayer is an MCP gateway: it sits between your AI agents and Elevenlabs, and checks every tool call against a rule you set before the call runs. Nothing changes on the server itself. For create_text_to_speech_generation, this is the rule to start with:
create_text_to_speech_generation stays usable, but capped: an agent stuck in a loop can't make hundreds of changes a minute. Everything else on the server is denied unless you say otherwise.
The button opens the PolicyLayer dashboard: create your workspace, connect Elevenlabs, apply this rule, and every create_text_to_speech_generation call is checked against it from then on.
Questions about create_text_to_speech_generation
Create Speech Generation. It is categorised as a Write tool in the Elevenlabs MCP Server, which means it can create or modify data. Consider rate limits to prevent runaway writes.
create_text_to_speech_generation accepts 8 parameters: text, voice, webhook, model_id, language_code, output_format, voice_settings, pronunciation_dictionary_locators. The full parameter table on this page comes from the server's own tool schema.
Register the Elevenlabs MCP server in PolicyLayer and add a rule for create_text_to_speech_generation: allow, deny, rate-limit, or require approval. Point your MCP client at the PolicyLayer proxy URL and the rule is enforced on every call, before it reaches Elevenlabs. Nothing to install.
create_text_to_speech_generation is a Write tool with medium risk. Write tools should be rate-limited to prevent accidental bulk modifications.
Yes. Add a rate_limit block to the create_text_to_speech_generation rule in your PolicyLayer policy. For example, setting max: 10 and window: 60 limits the tool to 10 calls per minute. Rate limits are tracked per agent session and reset automatically.
Set action: deny in the PolicyLayer policy for create_text_to_speech_generation. The AI agent will receive a policy violation error and cannot call the tool. You can also include a reason field to explain why the tool is blocked.
create_text_to_speech_generation is provided by the Elevenlabs MCP server (@aiwerk/mcp-server-elevenlabs). PolicyLayer sits as a proxy in front of this server to enforce policies before tool calls reach the server.
More on Elevenlabs, and thousands of servers like it.
This server
Across the catalogue