Skip to benchmark contentMotionBenchHosted by Baz Studio Human assessment
Pilot datasetOne prompt · one trial · 33 routes · not the official Core rankingRead the limits
openai · C
o4-mini
o4-miniPilot visual tierB
Post-fix rerun produces a clean real scene and valid 90-frame export.
One diagnostic trial. This is not a durable model capability claim.