P-Video Avatar TTS
P-Video Avatar TTS is Pruna AI's talking-avatar model driven by text. Write what the character says and it generates the speech with a built-in voice and animates a single image of the person speaking it, at 720p or 1080p. Attach a character to set the voice and, when no image is given, the portrait it speaks from.
See more
P-Video Avatar TTS is Pruna AI's talking-avatar model driven by text. Write what the character says and it generates the speech with a built-in voice and animates a single image of the person speaking it, at 720p or 1080p. Attach a character to set the voice and, when no image is given, the portrait it speaks from.
Connect Claude, Cursor, VS Code, or Codex to Layer’s MCP server and P-Video Avatar TTS becomes a tool your agent can reach for — it picks the parameters, runs the generation, and shows you the result in the conversation.
https://mcp.app.layer.ai/mcpAsk your agent
“Use P-Video Avatar TTS to animate a torch-lit dungeon corridor with a slow dolly forward.”
execute_forge({
"base_model_id": "pruna-p-video-avatar-tts",
"prompt": "A torch-lit dungeon corridor, slow dolly forward",
"batch_size": 1
})Authenticate with a personal access token, post the model’s own form, and poll the inference until it completes. Every field below is accepted by this endpoint.
# Start the generation — returns an inference id straight away.
curl -X POST https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/base-models/pruna-p-video-avatar-tts/inferences \
-H "Authorization: Bearer $LAYER_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"prompt": "A torch-lit dungeon corridor, slow dolly forward"
}'
# Poll until status is "complete" — the response then carries the asset URLs.
curl https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/inferences/$INFERENCE_ID \
-H "Authorization: Bearer $LAYER_TOKEN"- prompt stringrequired
- Describe what to generate, or the change to make.
- aspect_ratio enumdefault LANDSCAPE_16_9_720p6 options
- The shape and size of the output.
- seed integer
- The starting point for the model's randomness. The same seed and settings give the same result.
- guidance_file_first_frame file[]required
- The frame the video starts on
- prompt_language string
- The language the prompt is written in, so it reaches the model as you meant it.
Why use P-Video Avatar TTS for video?
Cinematic video generation
P-Video Avatar TTS creates high-fidelity video content suitable for trailers, gameplay previews, and marketing materials with temporal consistency.
Flexible input modes
Generate from text prompts, reference images, or existing video. Supports various aspect ratios and durations for different content needs.
Production-ready output
Clean motion, stable framing, and coherent transitions make outputs ready for direct use in UA campaigns and social content.
Game trailers and cinematic previews
Create attention-grabbing trailers and gameplay previews that showcase your game at its best. High temporal consistency ensures smooth, professional results.
UA video creative at scale
Generate dozens of video ad variants for A/B testing across platforms. Rapid iteration lets performance teams optimize creative angles quickly.
Character animations and in-game cinematics
Bring characters to life with animated sequences for cutscenes, promotional materials, and social content.
FAQ
How much does P-Video Avatar TTS cost? Is it free?+
How does Layer's pricing work?+
There are no seat fees, feature gates, or plans on Layer. Instead, our platform uses a consumption based system with Creative Units (CUs) with a flexible monthly subscription. Every generation on Layer (image, video, 3D, or audio) consumes a Creative Unit, and you only pay for what you create.
Other video models on Layer
P-Video Replace
Pruna AI
Pruna's character swap: replaces the person in a video with one from reference images.
P-Video Animate
Pruna AI
Pruna's motion transfer: a reference image performs the motion and audio of a source video.
P-Video Edit
Pruna AI
Pruna's video editor: change a clip of up to 15 seconds with a prompt and optional reference images.
P-Video Avatar
Pruna AI
Pruna's fast, low-cost talking avatar: a photo lip-synced to an audio track.
P-Video 2 Pro
Pruna AI
Pruna's highest-quality video model, with speed and quality modes at 480p or 768p.
P-Video 2
Pruna AI
Pruna's quality-focused successor to P-Video, at 720p or 1080p.