clip-ViT-B-32
The CLIP model maps text and images to a shared vector space, enabling various applications such as image search, zero-shot image classification, and image clustering.
text → embedding · made by sentence-transformers
$0.005per Mtok in
DeepInfra
Sold by 1 way
| Seller | Lane | Rate |
|---|---|---|
| DeepInfra aggregator | standard | $0.005per Mtok inapi.deepinfra.com · read 2026-08-25 |
About
clip-ViT-B-32 — a vectors model from sentence-transformers, sold by one company from $0.005 per Mtok in.
It takes text and returns vectors, with a context window of 77 tokens. The catalogue files it under embedding. Only DeepInfra sells it, at $0.005 per Mtok in.
Every current figure
- Maker
- sentence-transformers
- Register
- model
- Takes
- text
- Returns
- embedding
- Context
- 77 tokens
- Sellers
- 1
- Maker's own price
- not read
- Price
- $0.005 per Mtok in — DeepInfra
Known as 1 name
sentence-transformers/clip-ViT-B-32