Qwen3-VL-2B-Instruct
Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.
text + image → text · made by Alibaba
Sold by
Nobody in the catalogue publishes a price for this yet.
About
Qwen3-VL-2B-Instruct — a text model from Alibaba.
It takes text and images and returns text. It was published in October 2025. The catalogue files it under chat. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- Alibaba
- Register
- model
- Takes
- text + image
- Returns
- text
- Published
- October 2025
- Parameters
- 2.1 billion · read from its own weights
- Licence
- apache-2.0
Known as 1 name
Qwen/Qwen3-VL-2B-Instruct