Models/Veo 3.1
MODEL · VIDEONEWPOPULAR

Veo 3.1

Google DeepMind Veo 3.1 — the highest realism in the suite, and the only model that generates synced audio with the visuals. Picked by the router when the prompt asks for "cinematic", "live action", names a real-world camera move, or needs sound.

Selected works.

8 OF MANY · GENERATED BY THE COMMUNITY
Veo 3.1 sample 10:08
Veo 3.1 sample 20:08
Veo 3.1 sample 30:08
Veo 3.1 sample 40:08
Veo 3.1 sample 50:08
Veo 3.1 sample 60:08
Veo 3.1 sample 70:08
Veo 3.1 sample 80:08

CAPABILITIES

What it can
do well.

  • Text-to-videosupported
  • Image-to-videosupported
  • Clip lengthup to 8 seconds
  • Output resolution1080p
  • Audio generationnative, synced to the visuals
  • ×Lip-syncuse Lip Sync tool

Prompts that hit.

1 PATTERNS · COMMUNITY-SOURCED
PATTERN · CAMERA

"Slow dolly forward on [subject]. [Light condition]. [Mood]. [Ambient sound]."

Naming the camera move and the sound is the biggest unlock for Veo.

ABOUT THIS MODEL

FAQ.

Q.01

When does the auto-router pick Veo 3.1?

When the prompt requests video output and the requested duration / resolution falls within Veo 3.1's capabilities.

Q.02

Can I pin this model as default?

Yes. The model picker in any chat lets you pin per-turn or per-session. Set as default in /settings.