Pass IndexThe State of AISign in

Veo 3.1 Fast

Veo 3.1 is the latest text-to-video model from Google that generates high-fidelity, cinematic videos with synchronized audio from a simple text prompt.

text + image + video → audio + video · made by Google

$0.15per second
DeepInfra

Sold by 3 ways

SellerLaneRate
Google 0 ways$0.1
Google apinot on the seller's list todaystandard$6per minuteai.google.dev · read 2026-08-24
Google Vertex AI 0 ways$0.1
Google Vertex AI cloudnot on the seller's list todaystandard$6per minutecloud.google.com · read 2026-08-24
DeepInfra aggregatorstandard$9per minuteapi.deepinfra.com · read 2026-08-25

Measured 1 standing

PlaceBoardMetricScore
16thof 32Artificial Analysis · Text to Video ArenaElo1086

About

Veo 3.1 Fast — an audio and video model from Google, sold by one company from $9 per minute, placed 16th of 32 on Artificial Analysis · Text to Video Arena.

It takes text, images and video and returns audio and video, with a context window of 1,024 tokens. It was published in October 2025. The catalogue files it under video. Only DeepInfra sells it, at $9 per minute. It stands 16th of 32 on Artificial Analysis · Text to Video Arena.

Every current figure

Maker
Google
Register
model
Takes
text + image + video
Returns
audio + video
Context
1,024 tokens
Published
October 2025
Licence
not read
Sellers
1
Price
$9 per minute — DeepInfra
Boards
1
Best place
16th of 32 — Artificial Analysis · Text to Video Arena

Known as 2 names

google/veo-3.1-fastveo-3.1-fast-generate-preview