Veo 3.1 Fast
Veo 3.1 is the latest text-to-video model from Google that generates high-fidelity, cinematic videos with synchronized audio from a simple text prompt.
text + image + video → audio + video · made by Google
$0.15per second
DeepInfra
Sold by 3 ways
| Seller | Lane | Rate |
|---|---|---|
| Google 0 ways | $0.1 | |
| Google apinot on the seller's list today | standard | $6per minuteai.google.dev · read 2026-08-24 |
| Google Vertex AI 0 ways | $0.1 | |
| Google Vertex AI cloudnot on the seller's list today | standard | $6per minutecloud.google.com · read 2026-08-24 |
| DeepInfra aggregator | standard | $9per minuteapi.deepinfra.com · read 2026-08-25 |
Measured 1 standing
| Place | Board | Metric | Score |
|---|---|---|---|
| 16thof 32 | Artificial Analysis · Text to Video Arena | Elo | 1086 |
About
Veo 3.1 Fast — an audio and video model from Google, sold by one company from $9 per minute, placed 16th of 32 on Artificial Analysis · Text to Video Arena.
It takes text, images and video and returns audio and video, with a context window of 1,024 tokens. It was published in October 2025. The catalogue files it under video. Only DeepInfra sells it, at $9 per minute. It stands 16th of 32 on Artificial Analysis · Text to Video Arena.
Every current figure
- Maker
- Register
- model
- Takes
- text + image + video
- Returns
- audio + video
- Context
- 1,024 tokens
- Published
- October 2025
- Licence
- not read
- Sellers
- 1
- Price
- $9 per minute — DeepInfra
- Boards
- 1
- Best place
- 16th of 32 — Artificial Analysis · Text to Video Arena
Known as 2 names
google/veo-3.1-fastveo-3.1-fast-generate-preview