Pass IndexThe State of AISign in

Hive Vision Language Model

The Hive Vision Language Model turns images, or image and text pairs, into plain-language answers and structured JSON in one call, so teams can handle moderation, tagging or detection of subtle elements without stitching together multiple models.

text + image → text · made by Hive

$0.5$2.5per Mtok in / out
Hive

Sold by 1 way

SellerLaneRate
Hive apistandard$0.5per Mtok in$2.5per Mtok outthehive.ai · read 2026-08-25

About

Hive Vision Language Model — a text model from Hive, sold by one company from $0.5 in and $2.5 out per million tokens.

It takes text and images and returns text. The catalogue files it under guard and extract. Only Hive sells it, at $0.5 in and $2.5 out per million tokens.

Every current figure

Maker
Hive
Register
model
Takes
text + image
Returns
text
Sellers
1
Price
$0.5 in and $2.5 out per million tokens — Hive