Pass IndexThe State of AISign in

Kimi-VL-A3B-Instruct

We present Kimi-VL, an efficient open-source Mixture-of-Experts (MoE) vision-language model (VLM) that offers advanced multimodal reasoning, long-context understanding, and strong agent capabilities—all while activating only 2.8B parameters in its language decoder (Kimi-VL-A3B).

text + image → text · made by Moonshot AI

Sold by

Nobody in the catalogue publishes a price for this yet.

About

Kimi-VL-A3B-Instruct — a text model from Moonshot AI.

It takes text and images and returns text. It was published in April 2025. The catalogue files it under chat. Its weights are published under MIT, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
Moonshot AI
Register
model
Takes
text + image
Returns
text
Published
April 2025
Licence
mit

Known as 1 name

moonshotai/Kimi-VL-A3B-Instruct