Pass IndexThe State of AISign in

Veo 3.1 Standard

Veo 3.1 is the latest text-to-video model from Google that generates high-fidelity, cinematic videos with synchronized audio from a simple text prompt.

text + image + video → audio + video · made by Google

$0.4per second
DeepInfra

Sold by 3 ways

SellerLaneRate
Google 0 ways$0.4
Google apinot on the seller's list todaystandard$24per minuteai.google.dev · read 2026-08-24
Google Vertex AI 0 ways$0.4
Google Vertex AI cloudnot on the seller's list todaystandard$24per minutecloud.google.com · read 2026-08-24
DeepInfra aggregatorstandard$24per minuteapi.deepinfra.com · read 2026-08-25

Measured 1 standing

PlaceBoardMetricScore
12thof 32Artificial Analysis · Text to Video ArenaElo1090

About

Veo 3.1 Standard — an audio and video model from Google, sold by one company from $24 per minute, placed 12th of 32 on Artificial Analysis · Text to Video Arena.

It takes text, images and video and returns audio and video, with a context window of 1,024 tokens. It was published in October 2025. The catalogue files it under video. Only DeepInfra sells it, at $24 per minute. It stands 12th of 32 on Artificial Analysis · Text to Video Arena.

Every current figure

Maker
Google
Register
model
Takes
text + image + video
Returns
audio + video
Context
1,024 tokens
Published
October 2025
Licence
not read
Sellers
1
Price
$24 per minute — DeepInfra
Boards
1
Best place
12th of 32 — Artificial Analysis · Text to Video Arena

Known as 3 names

Veo 3.1google/veo-3.1veo-3.1-generate-preview