Hive Vision Language Model
The Hive Vision Language Model turns images, or image and text pairs, into plain-language answers and structured JSON in one call, so teams can handle moderation, tagging or detection of subtle elements without stitching together multiple models.
text + image → text · made by Hive
$0.5→$2.5per Mtok in / out
Hive
Sold by 1 way
| Seller | Lane | Rate |
|---|---|---|
| Hive api | standard | $0.5per Mtok in$2.5per Mtok outthehive.ai · read 2026-08-25 |
About
Hive Vision Language Model — a text model from Hive, sold by one company from $0.5 in and $2.5 out per million tokens.
It takes text and images and returns text. The catalogue files it under guard and extract. Only Hive sells it, at $0.5 in and $2.5 out per million tokens.
Every current figure
- Maker
- Hive
- Register
- model
- Takes
- text + image
- Returns
- text
- Sellers
- 1
- Price
- $0.5 in and $2.5 out per million tokens — Hive