clip-vit-base-patch16
The original implementation had two variants: one using a ResNet image encoder and the other using a Vision Transformer.
image → text · made by OpenAI
Sold by
Nobody in the catalogue publishes a price for this yet.
About
clip-vit-base-patch16 — a text model from OpenAI.
It takes images and returns text. It was published in March 2022. The catalogue files it under chat.
Every current figure
- Maker
- OpenAI
- Register
- model
- Takes
- image
- Returns
- text
- Published
- March 2022
Known as 1 name
openai/clip-vit-base-patch16