Pass IndexThe State of AISign in

LocateAnything-3B

LocateAnything is a vision-language model for fast and high-quality visual grounding, enabling precise object localization, dense detection, and point-based localization across diverse domains in both Enterprise Intelligence and Physical AI.

text + image → text · made by NVIDIA

Sold by

Nobody in the catalogue publishes a price for this yet.

About

LocateAnything-3B — a text model from NVIDIA.

It takes text and images and returns text. It was published in March 2026. The catalogue files it under search. Its weights are published under the other licence, which attaches conditions the plain open licences do not.

Every current figure

Maker
NVIDIA
Register
model
Takes
text + image
Returns
text
Published
March 2026
Parameters
3.8 billion · read from its own weights
Licence
other

Known as 1 name

nvidia/LocateAnything-3B