Pass IndexThe State of AISign in

Qwen3-VL-2B-Instruct

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

text + image → text · made by Alibaba

Sold by

Nobody in the catalogue publishes a price for this yet.

About

Qwen3-VL-2B-Instruct — a text model from Alibaba.

It takes text and images and returns text. It was published in October 2025. The catalogue files it under chat. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
Alibaba
Register
model
Takes
text + image
Returns
text
Published
October 2025
Parameters
2.1 billion · read from its own weights
Licence
apache-2.0

Known as 1 name

Qwen/Qwen3-VL-2B-Instruct