SlopTV

Benchmark

Animals leaderboard

Every model runs the same 40 prompts. A vision model scores each clip blind against the original prompt in a single pass, across visual fidelity, physics and motion, subject consistency, prompt adherence and audio sync. Scores are out of 5, and the 95% CI column shows how much a score could plausibly move on a re-run with this few samples. Full methodology.

Board: Overall · Hands · Faces · Liquids · On-screen text · Multi-shot · Lip-sync · Camera · Animals · Style · Image to video

Clip shown: every model here on the same prompt, “animal-horse-gallop”, so the preview is a fair comparison, not a cherry-picked best clip.

#ModelClipOverall95% CIFidelity VisualPhysicsConsistency AudioAttemptsGen (s) $/runRuns
1 Seedance 2.0
ByteDance
5.00 ±0 5.00 5.00 5.00 327 4
2 Seedance 2.5
ByteDance
5.00 ±0 5.00 5.00 5.00 233 4
3 Kling 3.0 Turbo
Kuaishou
5.00 ±0 5.00 5.00 5.00 5.00 5.00 4
4 Grok Video
xAI
4.54 ±0.5 4.63 4.63 4.38 1.00 85 4
5 MiniMax H3
MiniMax
4.54 ±1.46 4.63 4.63 4.38 448 4
6 OpenAI Sora 2 discontinued
OpenAI
4.46 ±1 4.50 4.75 4.13 1.00 217 4
7 Google Veo 3.1
Google
4.33 ±1.5 4.38 4.50 4.13 131 4
8 Runway Gen-4.5
Runway
4.19 ±0.88 4.40 4.25 3.70 4.40 4
9 PixVerse v5.5
PixVerse
4.00 ±1.03 3.90 4.25 3.53 4.33 4
10 Kling 3.0
Kuaishou
3.75 ±0.98 3.88 4.25 3.13 1.00 190 4
11 MiniMax Hailuo 02
MiniMax
3.71 ±2.56 3.75 4.00 3.38 300 4

Attempts is the average number of generations needed before one was usable, the number that actually decides what a model costs you. 95% CI is computed from that model's own runs on this board, not assumed. On a category board with only 4 runs per model, a wide CI is the honest reflection of a small sample, not a bug. Two models whose CIs overlap are not meaningfully different at this sample size.

Frequently asked

Which AI video model is best at Animals right now?

As of 23 September 2026, Seedance 2.0 by ByteDance ranks #1 on Animals on SlopTV's benchmark, scoring 5.00/5 across 4 scored runs (±0 at 95% confidence).