Veo 3.1 is Google’s current video generation model, replacing the now-discontinued Veo 2. Per Google DeepMind’s own model page, it generates 8-second clips by default with native audio built directly into generation — dialogue, sound effects, and ambient sound synced to the video, not added afterward — and supports resolutions up to 4K.
Beyond straightforward text-to-video, Veo 3.1 supports image-to-video, character consistency across shots, scene extension, object insertion and removal, camera and motion controls, and style-reference matching — more directorial control than a single-shot prompt-to-clip model.

Promptus AI is the easiest way to generate realistic photos, videos, 3D and ComfyUI workflows with artificial intelligence.
Our AI photo generator produces lifelike portraits, product images, and creative concepts in seconds, making it the perfect tool for creators and brands.
Join our distributed GPU compute network. Help us make AI accessible, scalable
and secure for designer, developers and start-ups.