Qwen-VL Max
Qwen-VL Max is a model by Alibaba for conversation, it stands 7th of 8 on Spatialviz bench — Epoch AI, and 2 sellers in the catalogue publish a price for it.
text + image → text · made by Alibaba
$0.8→$3.2per Mtok in / out
Alibaba · 2 sellers
Sold by 2 ways
| Seller | Lane | Rate |
|---|---|---|
| Alibaba aggregator | standard | $0.8per Mtok in$3.2per Mtok outmodels.dev · read 2026-08-27 |
| ElectronHub aggregator | standard | $1per Mtok in$3.5per Mtok outapi.electronhub.ai · read 2026-09-16 |
Measured 2 standings
| Place | Board | Metric | Score |
|---|---|---|---|
| 7thof 8 | Spatialviz bench — Epoch AI | Overall score (max) | 0.32 |
| 41stof 50 | Video mme — Epoch AI | Overall (no subtitles) (max) | 0.513 |
About
Qwen-VL Max — a text model from Alibaba, sold by 2 companies from $0.8 in and $3.2 out per million tokens, placed 41st of 50 on Video mme — Epoch AI.
It takes text and images and returns text, with a context window of 131,072 tokens. It was published in April 2024, trained on material up to April 2024. Its sellers say it can call a tool. The catalogue files it under chat. Two companies sell it. The cheapest is $0.8 in and $3.2 out per million tokens at Alibaba. It has been measured on 2 boards, and stands best at 41st of 50 on Video mme — Epoch AI.
Every current figure
- Maker
- Alibaba
- Register
- model
- Takes
- text + image
- Returns
- text
- Context
- 131,072 tokens
- Longest answer
- 8,192 tokens
- Published
- April 2024
- Knowledge to
- April 2024
- Licence
- not read
- Sellers
- 2
- Price
- $0.8 in and $3.2 out per million tokens — Alibaba
- Boards
- 2
- Best place
- 41st of 50 — Video mme — Epoch AI