Crosshair
Leaderboard
World ModelsProprietary

Veo 3

Google DeepMind's text/image-to-video model with native audio and real-world physics; among the top performers on the PAI-Bench physical-AI generation benchmark.

videoaudioOfficial site
Crosshair Index
#6 of 10 · World Models
Provider
Google DeepMind
Released
2025-05-20
Parameters
Undisclosed

Capability web

Strengths across world-modeling capabilities — each axis is one benchmark domain. Coverage is sparse while these evaluations mature; empty axes read “no data yet.”

UnderstandingAnticipationGen. PhysicsPhysical AI

Physical-AI Generation

82
skill

Generation quality plus physical plausibility for embodied / physical-AI scenes — driving, robotics, ego-centric — judged on PAI-Bench.

Scorecard

BenchmarkScoreSourceStatus
Something-Something v2
Motion Understanding
not evaluated
EPIC-Kitchens-100 Anticipation
Action Anticipation
not evaluated
Perception Test
Video Understanding
not evaluated
Physics-IQ
Generative Physics
not evaluated
PAI-Bench-G
Physical-AI Generation
82.2PAI-Bench (arXiv:2512.01989, Table 3)paperunverified