Text-to-Video
Veo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.
1,207
Jan 2026
video
$0.10/s
Catalog
Category rows come directly from Artificial Analysis when the endpoint exposes category-level Elo scores.
| Category | Elo | 95% CI | Appearances |
|---|---|---|---|
| 3D animation | 1,312 | -44/44 | 195 |
| Sports | 1,269 | -26/26 | 601 |
| Fashion | 1,257 | -28/28 | 478 |
| Sci Fi | 1,249 | -21/21 | 886 |
| Cartoon and anime | 1,246 | -25/25 | 575 |
| People | 1,245 | -12/12 | 2,819 |
| Text | 1,244 | -29/29 | 492 |
| Specific location or era | 1,244 | -22/22 | 857 |
| Transport |
| 1,243 |
| -16/16 |
| 1,578 |
| Fantasy | 1,237 | -31/31 | 437 |
| Long prompt | 1,235 | -24/24 | 729 |
| Indoor | 1,233 | -18/18 | 1,255 |