New Your team’s decisions, in one playbook every coding agent works from. Never answer your agent twice

text_to_speech_stream

Text To Speech Streaming. Spends ElevenLabs credits. Returns audio/mpeg bytes; pass output_path to save them.

SERVERElevenlabs SOURCE@aiwerk/mcp-server-elevenlabs
Medium RISK CLASS
Category Write
Parameters 122 required
Recommended Rate-limitedsee the rule below
Registry record Grade F, identity unverified Pull the record →

This record as markdown: /tools/aiwerk-mcp-server-elevenlabs/text-to-speech-stream.md

What text_to_speech_stream does on Elevenlabs

AI agents use text_to_speech_stream to create or update resources in Elevenlabs, usually the action step of a workflow, after the agent has gathered context. Every call changes real data in your Elevenlabs environment.

ParameterTypeRequiredDescription
seed number | null
text string Yes The text that will get converted into speech.
model_id string Identifier of the model that will be used, you can query them using GET /v1/models. The model needs to have support for text to speech, you can check this using
voice_id string Yes Voice ID to be used, you can use https://api.elevenlabs.io/v1/voices to list all the available voices.
next_text string | null
output_path string Where to write the returned bytes. Relative paths resolve against ELEVENLABS_OUTPUT_DIR. Omit it to get the data inline as base64 (small files only).
language_code string | null
output_format string Output format of the generated audio. Formatted as codec_sample_rate_bitrate. So an mp3 with 22.05kHz sample rate at 32kbs is represented as mp3_22050_32. MP3 w
previous_text string | null
enable_logging boolean When enable_logging is set to false zero retention mode will be used for the request. This will mean history features are unavailable for this request, includin
use_pvc_as_ivc boolean If true, we won't use PVC version of the voice for the generation but the IVC version. This is a temporary workaround for higher latency in PVC versions.
voice_settings object

Parameters from the server's own tool schema.

Why text_to_speech_stream is rated Medium

An AI agent can call text_to_speech_stream faster than any human can review: one bad instruction and it creates or modifies resources in Elevenlabs by the hundred, each call as confident as the last.

Risk signalsHigh parameter count (18 properties)

Questions about text_to_speech_stream

What does the text_to_speech_stream tool do? +

Text To Speech Streaming. Spends ElevenLabs credits. Returns audio/mpeg bytes; pass output_path to save them. It is categorised as a Write tool in the Elevenlabs MCP Server, which means it can create or modify data. Consider rate limits to prevent runaway writes.

What parameters does text_to_speech_stream accept? +

text_to_speech_stream accepts 12 parameters: seed, text, model_id, voice_id, next_text, output_path, language_code, output_format, previous_text, enable_logging, use_pvc_as_ivc, voice_settings. Required: text, voice_id. The full parameter table on this page comes from the server's own tool schema.

How do I enforce a policy on text_to_speech_stream? +

Register the Elevenlabs MCP server in PolicyLayer and add a rule for text_to_speech_stream: allow, deny, rate-limit, or require approval. Point your MCP client at the PolicyLayer proxy URL and the rule is enforced on every call, before it reaches Elevenlabs. Nothing to install.

What risk level is text_to_speech_stream? +

text_to_speech_stream is a Write tool with medium risk. Write tools should be rate-limited to prevent accidental bulk modifications.

Can I rate-limit text_to_speech_stream? +

Yes. Add a rate_limit block to the text_to_speech_stream rule in your PolicyLayer policy. For example, setting max: 10 and window: 60 limits the tool to 10 calls per minute. Rate limits are tracked per agent session and reset automatically.

How do I block text_to_speech_stream completely? +

Set action: deny in the PolicyLayer policy for text_to_speech_stream. The AI agent will receive a policy violation error and cannot call the tool. You can also include a reason field to explain why the tool is blocked.

What MCP server provides text_to_speech_stream? +

text_to_speech_stream is provided by the Elevenlabs MCP server (@aiwerk/mcp-server-elevenlabs). PolicyLayer sits as a proxy in front of this server to enforce policies before tool calls reach the server.

More on Elevenlabs, and thousands of servers like it.

Across the catalogue

// THE MCP REGISTRY

PolicyLayer tracks 44,603 MCP servers and 515,000+ tools.

Every server has a live record: who publishes it, whether it answers without auth, its risk grade, every tool classified, the recommended policy. This page is one line of Elevenlabs's. Pull the full record:

Teams ship this data inside their own products. See what a licence covers →

// GET IN TOUCH

Have a question or want to learn more? Send us a message.

Message sent.

We'll get back to you soon.