Overview
This model generates video from text prompts or input images via an image_url parameter. It delivers sharper detail and smoother motion than its predecessor, with a focus on photorealistic human subjects and cinematic framing. On Kyma, requests route through an OpenAI-compatible endpoint with automatic failover to maintain reliability. Prompt caching is supported, billing repeated prompt prefixes at this model’s cached input rate. Every response returns the exact compute cost in usage.cost and identifies the active model in the X-Kyma-Model header. The model operates on a premium cost tier and runs at a slower inference speed compared to fast-tier alternatives. It does not natively generate audio; for synchronized soundtracks, you must route to the dedicated audio variant.Specs
Pricing
Use this when
- Cinematic Character Shots — Generate high-fidelity video clips focused on realistic human faces and smooth motion.
- Brand Hero Videos — Produce premium marketing assets from text prompts or reference images.
- Image-to-Video Conversion — Animate static reference frames into continuous, photorealistic video sequences.
- Premium Visual Prototyping — Test high-quality video concepts before committing to full production pipelines.
Not ideal for
Do not use this model for real-time applications, low-budget batch generation, or workflows that require native audio synthesis.Pick something else when
- You need faster inference speeds: use
seedance-2-fastorveo-3-fast. - You require native audio generation: use
kling-3-pro-audio. - You are generating lower-resolution drafts: use
hailuo-02-512porhailuo-02-768p. - You need a cheaper alternative for bulk testing: use
kling-2.5-pro.