Text-to-Video
Step-Video-T2V is ranked in the Text-to-Video benchmark from Artificial Analysis.
925
Feb 2025
n/a
n/a
n/a
Category rows come directly from Artificial Analysis when the endpoint exposes category-level Elo scores.
| Category | Elo | 95% CI | Appearances |
|---|---|---|---|
| Sci Fi | 990 | -22/22 | 683 |
| Sports | 984 | -31/31 | 350 |
| Action | 983 | -29/29 | 309 |
| Buildings | 979 | -15/15 | 2,168 |
| Transport | 978 | -17/17 | 1,579 |
| Long prompt | 972 | -24/24 | 1,042 |
| Multi-scene | 968 | -34/34 | 289 |
| People | 963 | -13/13 | 2,570 |
| Text | 958 |
| -30/30 |
| 522 |
| Technology | 957 | -14/14 | 1,716 |
| Specific location or era | 952 | -23/23 | 973 |
| Physics | 942 | -15/15 | 1,746 |