MiniMax H3 Max Lip Sync
MiniMax H3 Max Lip Sync from MiniMax generates a talking video from a still image and supplied audio, synchronizing mouth movements to the soundtrack. Audio from 5 to 15 seconds drives clip length, with optional transcription guidance and output at 480p, 768p, 1080p, or 2K using the aspect ratio nearest the image. Ideal for virtual presenters, dubbed character stills, and narrated talking-head content.
See more
MiniMax H3 Max Lip Sync from MiniMax generates a talking video from a still image and supplied audio, synchronizing mouth movements to the soundtrack. Audio from 5 to 15 seconds drives clip length, with optional transcription guidance and output at 480p, 768p, 1080p, or 2K using the aspect ratio nearest the image. Ideal for virtual presenters, dubbed character stills, and narrated talking-head content.
Connect Claude, Cursor, VS Code, or Codex to Layer’s MCP server and MiniMax H3 Max Lip Sync becomes a tool your agent can reach for — it picks the parameters, runs the generation, and shows you the result in the conversation.
https://mcp.app.layer.ai/mcpAsk your agent
“Use MiniMax H3 Max Lip Sync to animate a torch-lit dungeon corridor with a slow dolly forward.”
execute_forge({
"base_model_id": "minimax-h3-max-lip-sync",
"prompt": "A torch-lit dungeon corridor, slow dolly forward",
"batch_size": 1
})Authenticate with a personal access token, post the model’s own form, and poll the inference until it completes. Every field below is accepted by this endpoint.
# Start the generation — returns an inference id straight away.
curl -X POST https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/base-models/minimax-h3-max-lip-sync/inferences \
-H "Authorization: Bearer $LAYER_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"guidance_file_first_frame": [
{
"url": "https://example.com/input.png"
}
]
}'
# Poll until status is "complete" — the response then carries the asset URLs.
curl https://api.app.layer.ai/api/v2/workspaces/$WORKSPACE_ID/inferences/$INFERENCE_ID \
-H "Authorization: Bearer $LAYER_TOKEN"- seed integer
- The starting point for the model's randomness. The same seed and settings give the same result.
- guidance_file_first_frame file[]required
- The frame the video starts on
- guidance_file_reference_audio file[]required
- Guides style and voice without being modified
Why use MiniMax H3 Max Lip Sync for video?
Cinematic video generation
MiniMax H3 Max Lip Sync creates high-fidelity video content suitable for trailers, gameplay previews, and marketing materials with temporal consistency.
Flexible input modes
Generate from text prompts, reference images, or existing video. Supports various aspect ratios and durations for different content needs.
Production-ready output
Clean motion, stable framing, and coherent transitions make outputs ready for direct use in UA campaigns and social content.
Game trailers and cinematic previews
Create attention-grabbing trailers and gameplay previews that showcase your game at its best. High temporal consistency ensures smooth, professional results.
UA video creative at scale
Generate dozens of video ad variants for A/B testing across platforms. Rapid iteration lets performance teams optimize creative angles quickly.
Character animations and in-game cinematics
Bring characters to life with animated sequences for cutscenes, promotional materials, and social content.
FAQ
How much does MiniMax H3 Max Lip Sync cost? Is it free?+
How does Layer's pricing work?+
There are no seat fees, feature gates, or plans on Layer. Instead, our platform uses a consumption based system with Creative Units (CUs) with a flexible monthly subscription. Every generation on Layer (image, video, 3D, or audio) consumes a Creative Unit, and you only pay for what you create.
Other video models on Layer
MiniMax H3 Reference
MiniMax
Reference-guided MiniMax H3 video at 2K that follows a reference clip shot for shot while the prompt changes the subject or action.
MiniMax H3
MiniMax
MiniMax H3 (Hailuo-03) next-gen open-weight video with native stereo audio at 2K and 24 FPS.
Minimax Hailuo-2.3 Standard
MiniMax
Latest standard video model. Improved prompt understanding and visual consistency for daily creation.
FLUX.3 Video Upscale Creative
Black Forest Labs
FLUX.3-powered video upscaler that adds detail, up to 4K.
LTX Video 2.3 Quality Outpaint
Lightricks
Spatially outpaint a source video to a new aspect ratio with LTX 2.3 Quality.
FLUX.3 Video Upscale
Black Forest Labs
FLUX.3-powered source-faithful video upscaler, up to 4K.