Kimi-VL-A3B-Instruct
We present Kimi-VL, an efficient open-source Mixture-of-Experts (MoE) vision-language model (VLM) that offers advanced multimodal reasoning, long-context understanding, and strong agent capabilities—all while activating only 2.8B parameters in its language decoder (Kimi-VL-A3B).
text + image → text · made by Moonshot AI
Sold by
Nobody in the catalogue publishes a price for this yet.
About
Kimi-VL-A3B-Instruct — a text model from Moonshot AI.
It takes text and images and returns text. It was published in April 2025. The catalogue files it under chat. Its weights are published under MIT, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- Moonshot AI
- Register
- model
- Takes
- text + image
- Returns
- text
- Published
- April 2025
- Licence
- mit
Known as 1 name
moonshotai/Kimi-VL-A3B-Instruct