Skip to benchmark content
MotionBenchHosted by Baz Studio
Pilot datasetOne prompt · one trial · 33 routes · not the official Core rankingRead the limits
openai · O+C

GPT 5.4

gpt-5.4
Pilot visual tierC
Human assessment
Valid but small and visually conservative in sampled frames.

One diagnostic trial. This is not a durable model capability claim.

GPT 5.4 — MotionBench