ElevenLabs TTS V4 Turbo
Also known as Eleven v4 Turbo
Eleven v4 Turbo is the faster ElevenLabs text-to-speech tier. It generates speech from a chosen voice with inline audio tags, stability and similarity controls, and IPA pronunciation in forward slashes. Supports language selection, text normalization, and an optional seed. Best when you want v4 delivery at a lower character cost.
See more
Eleven v4 Turbo is the faster ElevenLabs text-to-speech tier. It generates speech from a chosen voice with inline audio tags, stability and similarity controls, and IPA pronunciation in forward slashes. Supports language selection, text normalization, and an optional seed. Best when you want v4 delivery at a lower character cost.
Connect Claude, Cursor, VS Code, or Codex to Layer’s MCP server and ElevenLabs TTS V4 Turbo becomes a tool your agent can reach for — it picks the parameters, runs the generation, and shows you the result in the conversation.
https://mcp.app.layer.ai/mcpAsk your agent
“Use ElevenLabs TTS V4 Turbo for a heavy stone door grinding open in a dungeon.”
execute_forge({
"base_model_id": "eleven-labs-tts-v4-turbo",
"prompt": "A heavy stone door grinding open in a dungeon",
"batch_size": 1
})Authenticate with a personal access token, post the model’s own form, and poll the inference until it completes. Every field below is accepted by this endpoint.
# Start the generation — returns an inference id straight away.
curl -X POST https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/base-models/elevenlabs-tts-v4-turbo/inferences \
-H "Authorization: Bearer $LAYER_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"prompt": "A heavy stone door grinding open in a dungeon"
}'
# Poll until status is "complete" — the response then carries the asset URLs.
curl https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/inferences/$INFERENCE_ID \
-H "Authorization: Bearer $LAYER_TOKEN"- prompt stringrequired
- The text the voice speaks.
- seed integer
- The starting point for the model's randomness. The same seed and settings give the same result.
- stability numberdefault 0.50–1
- How steady the voice stays. Lower is more expressive, higher more consistent.
- similarity_boost numberdefault 0.750–1
- How closely the voice keeps to the original speaker.
- text_normalization enumdefault AUTOAUTO, ON, OFF
- Spell out numbers, dates and abbreviations as words before speaking them.
- prompt_language string
- The language the voice speaks in.
Why use ElevenLabs TTS V4 Turbo for audio?
Natural voice generation
ElevenLabs TTS V4 Turbo produces expressive, high-quality audio suitable for character dialogue, narration, and voiceover work.
Flexible content types
Supports a range of use cases from in-game dialogue and cinematics to marketing narration and social media content.
Fast iteration
Generate and refine audio content quickly, enabling rapid prototyping of character voices and sound design.
Character voiceover and dialogue production
Generate expressive character voices for in-game dialogue, cutscenes, and interactive narratives. Iterate on tone and delivery rapidly.
Marketing narration and promotional audio
Create professional voiceovers for trailers, app store videos, and social media content without booking voice talent.
Sound design exploration and prototyping
Quickly prototype sound effects, ambient audio, and musical elements to test creative directions early in production.
FAQ
How much does ElevenLabs TTS V4 Turbo cost? Is it free?+
How does Layer's pricing work?+
There are no seat fees, feature gates, or plans on Layer. Instead, our platform uses a consumption based system with Creative Units (CUs) with a flexible monthly subscription. Every generation on Layer (image, video, 3D, or audio) consumes a Creative Unit, and you only pay for what you create.
Other audio models on Layer
ElevenLabs TTS V4
ElevenLabs
ElevenLabs' expressive v4 speech model with audio tags, stability, and similarity control.
ElevenLabs TTS V3
ElevenLabs
ElevenLabs' most expressive TTS model with inline audio tags for emotion and delivery.
ElevenLabs Multilingual V2
ElevenLabs
Multilingual text-to-speech with natural voice selection and stability controls.
Lyria 3.5
Google DeepMind's latest music model, writing full-length songs with vocals from one prompt.
VEED Clean Audio 1.0
VEED
Remove background noise from speech while keeping quiet and distant words intact.
MiniMax Music 3
MiniMax
High-performance MiniMax music model for complete songs up to five minutes with structure tags and seed control.