LocateAnything-3B
LocateAnything is a vision-language model for fast and high-quality visual grounding, enabling precise object localization, dense detection, and point-based localization across diverse domains in both Enterprise Intelligence and Physical AI.
text + image → text · made by NVIDIA
Sold by
Nobody in the catalogue publishes a price for this yet.
About
LocateAnything-3B — a text model from NVIDIA.
It takes text and images and returns text. It was published in March 2026. The catalogue files it under search. Its weights are published under the other licence, which attaches conditions the plain open licences do not.
Every current figure
- Maker
- NVIDIA
- Register
- model
- Takes
- text + image
- Returns
- text
- Published
- March 2026
- Parameters
- 3.8 billion · read from its own weights
- Licence
- other
Known as 1 name
nvidia/LocateAnything-3B