Text-to-Video leaderboard
Models ranked by blind head-to-head votes. Scores are Elo ratings and update as new matchups complete.
Generation highlights
How the top models compare
Best models, by
Best Text-to-Video Models by Price
Best Text-to-Video Models by Speed
3 models waiting for enough speed data.
Best AI Models for Text-to-Video
| # | Model | Elo |
|---|---|---|
| 1 | 1236 | |
| 2 | 1194 | |
| 3 | 1174 | |
| 4 | 1171 | |
| 5 | 1164 | |
| 6 | 1148 | |
| 7 | 1125 |
As of July 2026, Seedance 2.0 leads the leaderboard with a 1242 Elo and an 80.6% win rate, maintaining a 47-point lead over its closest competitor, Kling V3 (1195 Elo). While high-tier models dominate the top three, PrunaAI’s P-Video (1170 Elo) sits in fourth place despite costing 85% less and generating videos nearly eight times faster than the top-ranked models. This creates a narrow 25-point gap between the third-ranked Sora 2 Pro and the sixth-ranked Grok Imagine Video, highlighting intense competition among current-generation cinematic models.
Highlighted challenges
The Rubik's Gauntlet
This prompt is one of the hardest single tests for 2026 SOTA video models because it simultaneously demands extreme fine-motor precision at high speed, long-term physical consistency (the cube must genuinely solve without morphing), and complex multi-element rendering (hyper-detailed skin, sweat, glossy reflections, and dynamic camera movement). Areas where even top models still frequently break down.
Neon Rain Reverie
This prompt is exceptionally difficult because it combines complex fluid dynamics (rain, splashing, clinging wet fabric), advanced material simulation (flowing silk + hair in wind), and atmospheric lighting; three areas where even top 2026 models still frequently produce artifacts or unrealistic behavior.
Recent SOTA shifts in Text-to-Video
Full historyFAQ
What is the best AI text-to-video model?
Based on blind community voting, Seedance 2.0 is currently the #1 ranked AI text-to-video model with an Elo rating of 1242. Rankings update in real time as new votes come in.
How are AI text-to-video models ranked on Lumenfall?
Lumenfall Arena ranks AI models through blind community voting. In each matchup, two models generate from the same prompt and voters pick the better result without seeing model names. Votes are processed using TrueSkill, a Bayesian rating algorithm developed by Microsoft Research, that produces a single Elo score reflecting each model's relative quality.
What is an Elo rating for AI models?
An Elo rating is a numerical score representing a model's skill relative to other models. Under the hood, Lumenfall uses TrueSkill, which tracks two values per model: mu (estimated skill) and sigma (uncertainty). The displayed Elo is calculated as 1000 + 10 x (mu - 3*sigma), a conservative lower bound. A model must prove itself consistently across many matchups to earn a high rating.
Keep the arena honest
Cast your vote
Pick winners in blind matchups. Every vote nudges the Elo and shapes these rankings.
Cast Your VoteSuggest a prompt
Got an idea worth testing? Submit a prompt and watch the models battle it out.