Veo 3.1 Standard
Veo 3.1 is the latest text-to-video model from Google that generates high-fidelity, cinematic videos with synchronized audio from a simple text prompt.
text + image + video → audio + video · made by Google
$0.4per second
DeepInfra
Sold by 3 ways
| Seller | Lane | Rate |
|---|---|---|
| Google 0 ways | $0.4 | |
| Google apinot on the seller's list today | standard | $24per minuteai.google.dev · read 2026-08-24 |
| Google Vertex AI 0 ways | $0.4 | |
| Google Vertex AI cloudnot on the seller's list today | standard | $24per minutecloud.google.com · read 2026-08-24 |
| DeepInfra aggregator | standard | $24per minuteapi.deepinfra.com · read 2026-08-25 |
Measured 1 standing
| Place | Board | Metric | Score |
|---|---|---|---|
| 12thof 32 | Artificial Analysis · Text to Video Arena | Elo | 1090 |
About
Veo 3.1 Standard — an audio and video model from Google, sold by one company from $24 per minute, placed 12th of 32 on Artificial Analysis · Text to Video Arena.
It takes text, images and video and returns audio and video, with a context window of 1,024 tokens. It was published in October 2025. The catalogue files it under video. Only DeepInfra sells it, at $24 per minute. It stands 12th of 32 on Artificial Analysis · Text to Video Arena.
Every current figure
- Maker
- Register
- model
- Takes
- text + image + video
- Returns
- audio + video
- Context
- 1,024 tokens
- Published
- October 2025
- Licence
- not read
- Sellers
- 1
- Price
- $24 per minute — DeepInfra
- Boards
- 1
- Best place
- 12th of 32 — Artificial Analysis · Text to Video Arena
Known as 3 names
Veo 3.1google/veo-3.1veo-3.1-generate-preview