HumanEval.org

Benchmarks

LMArena Text-to-Video Arena

last retrieved 2026-09-08 21:40 UTC

Third-party scores shown as context with provenance — our own arena ratings from blind human votes stay alongside every row and remain the primary signal.

ModelLMArena Text-to-Video ArenaObservedSource
Gemini Omni FlashGoogle15111501 – 15212026-09-04source ↗
Dreamina Seedance 2.0 (720p)ByteDance14791471 – 14872026-09-04source ↗
MiniMax H3MiniMax14621452 – 14722026-09-04source ↗
HappyHorse 1.0Alibaba-ATH14271414 – 14402026-09-04source ↗
Sora 2 ProOpenAI13671360 – 13742026-09-04source ↗

— = not yet evaluated by HumanEval (no eligible arena votes). External values are the newest observation per model; hover a value for source notes and a rating for its 95% CI. Unit: Elo, higher is better.

Not enough overlap with our arena ratings for a correlation scatter yet — fewer than two scored models are rated.