Pass IndexThe State of AISign in

clip-ViT-B-32

The CLIP model maps text and images to a shared vector space, enabling various applications such as image search, zero-shot image classification, and image clustering.

text → embedding · made by sentence-transformers

$0.005per Mtok in
DeepInfra

Sold by 1 way

SellerLaneRate
DeepInfra aggregatorstandard$0.005per Mtok inapi.deepinfra.com · read 2026-08-25

About

clip-ViT-B-32 — a vectors model from sentence-transformers, sold by one company from $0.005 per Mtok in.

It takes text and returns vectors, with a context window of 77 tokens. The catalogue files it under embedding. Only DeepInfra sells it, at $0.005 per Mtok in.

Every current figure

Maker
sentence-transformers
Register
model
Takes
text
Returns
embedding
Context
77 tokens
Sellers
1
Maker's own price
not read
Price
$0.005 per Mtok in — DeepInfra

Known as 1 name

sentence-transformers/clip-ViT-B-32