Text to Video
Built for Professional Video Generation
Kandinsky 5.0 Video Pro is a large-scale video diffusion model built for creators and teams who demand visual fidelity, motion stability, and precise semantic control. Write a prompt and generate a short cinematic clip.
Describe the scene → Set aspect ratio → Get your video
Related: Image to Video · Kandinsky 5.0 Video Pro · Examples · FAQ
Video settings
FPS fixed at 24. Duration 4–10s · default 5s.
Preview
Generated video appears here
What is text to video?
Text to video generates a clip entirely from language. You describe the scene, action, camera, and mood; Kandinsky 5.0 Video Pro synthesizes the frames with stable motion and strong prompt alignment.
How to write a strong prompt
- Scene — Place, time of day, weather, and materials.
- Action — What moves, who performs, and how fast.
- Camera — Dolly, pan, zoom, or locked-off framing.
- Mood — Lighting, color grade, and cinematic tone.
When to use image to video instead
Start from a still when identity or composition must stay locked. Use Image to Video for product frames, portraits, and storyboard keys.