Open-weight embeddings models
86 in the catalogue today. Every one with what it costs, who sells it and where it stands.
Models that turn text into a vector so it can be searched by meaning rather than by words. They are cheap — usually cents per million tokens, and often priced for input only, since nothing comes back but numbers. Two things decide the choice: the dimension of the vector, which sets what your database will cost to hold, and whether the model was trained for your language. Changing model later means re-embedding everything you have.
Weights published under a licence that lets you run them where you like and sell what you build — Apache 2.0, MIT and their kin impose little beyond keeping the notice. This is the list to start from if you need the model on your own hardware, in your own region, or simply want no third party between you and it. The trade is that hosting is now your problem, and the prices shown beside each one are what somebody else charges to do it for you.
- BidirLM-1.7B-EmbeddingBidirLM21st of 86
- BidirLM-1B-EmbeddingBidirLM25th of 86
- BidirLM-Omni-2.5B-EmbeddingBidirLM23rd of 86
- F2LLM-0.6Bcodefuse-ai28th of 183
- F2LLM-1.7Bcodefuse-ai14th of 183· 2 boards
- F2LLM-4Bcodefuse-ai9th of 183· 2 boards
- F2LLM-v2-0.6Bcodefuse-ai29th of 183· 2 boards
- F2LLM-v2-1.7Bcodefuse-ai18th of 183· 2 boards
- F2LLM-v2-14Bcodefuse-ai9th of 238· 3 boards
- F2LLM-v2-4Bcodefuse-ai16th of 238· 3 boards
- F2LLM-v2-80Mcodefuse-ai72nd of 86
- F2LLM-v2-8Bcodefuse-ai10th of 238· 3 boards
- FRIDAai-forever
- Giga-Embeddings-instructai-sage23rd of 183
- GritLM-7BGritLM30th of 86
- Jasper-Token-Compression-600Minfgrad2nd of 238· 2 boards
- KaLM-embedding-multilingual-mini-instruct-v2.5KaLM-Embedding20th of 183
- LGAI-Embedding-Previewannamodels6th of 238· 2 boards
- LaBSEsentence-transformers
- Multilingual-e5-large-instructintfloat24th of 86$0.01–$0.08per Mtok in4 selling
- NeoBERTChandar Research Lab
- Octen-Embedding-8BOcten11th of 86
- QZhou-EmbeddingKingsoft-LLM2nd of 238· 2 boards
- UAE-Large-V1WhereIsAI
- Yuan-embedding-2.0-enIEITYuan1st of 238· 2 boards
- all-MiniLM-L12-v2sentence-transformers$0.005per Mtok in2 selling
- all-MiniLM-L6-v2sentence-transformers$0.005–$0.009per Mtok in3 selling
- all-mpnet-base-v2sentence-transformers$0.005per Mtok in2 selling
- bart-basefacebook
- bart-largefacebook
- bge-base-en-v1.5BAAI$0.005–$0.008per Mtok in3 selling
- bge-en-iclBAAI$0.01per Mtok in1 selling
- bge-large-enBAAI$0.1per Mtok in1 selling
- bge-large-en-v1.5BAAI$0.01per Mtok in2 selling
- bge-large-zhBAAI
- bge-large-zh-v1.5BAAI
- bge-m3BAAI$0.01–$0.02per Mtok in3 selling
- bge-small-en-v1.5BAAI
- clip-ViT-B-32-multilingual-v1sentence-transformers$0.005per Mtok in1 selling
- dinov2-basefacebook
- distiluse-base-multilingual-cased-v2sentence-transformers
- e5-base-v2intfloat$0.005per Mtok in2 selling
- e5-large-v2intfloat$0.01–$0.02per Mtok in3 selling
- e5-mistral-7b-instructintfloat20th of 238
- gte-Qwen2-1.5B-instructAlibaba-NLP
- gte-Qwen2-7B-instructAlibaba-NLP8th of 238· 3 boards
- gte-large-en-v1.5Alibaba-NLP
- gte-modernbert-baseAlibaba-NLP
- gte-multilingual-baseAlibaba-NLP
- gte-smallthenlper
- harrier-oss-v1-0.6bMicrosoft7th of 86
- harrier-oss-v1-270mMicrosoft15th of 86
- harrier-oss-v1-27bMicrosoft1st of 86
- instructor-largeNLP Group of The University of Hong Kong
- instructor-xlNLP Group of The University of Hong Kong
- jina-clip-v1Jina AI
- jina-embeddings-v2-base-codeJina AI
- jina-embeddings-v2-base-enJina AI
- jina-embeddings-v2-base-zhJina AI
- mimiKyutai
- modernbert-embed-baseNomic AI
- multilingual-e5-baseintfloat
- multilingual-e5-smallintfloat
- mxbai-embed-large-v1Mixedbread
- nomic-embed-text-v1Nomic AI
- nomic-embed-text-v2-moeNomic AI
- nomic-embed-vision-v1.5Nomic AI
- omnivinciNVIDIA
- paraphrase-MiniLM-L6-v2sentence-transformers$0.005per Mtok in2 selling
- paraphrase-multilingual-MiniLM-L12-v2sentence-transformers
- paraphrase-multilingual-mpnet-base-v2sentence-transformers
- prov-gigapathProv-GigaPath
- pubmedbert-base-embeddingsNeuML
- rubert-tiny2cointegrated
- speed-embedding-7b-instructHaon-Chen10th of 238
- stella_en_1.5B_v5NovaSearch13th of 238
- stella_en_400M_v5NovaSearch
- text2vec-base-chineseshibing624$0.005per Mtok in1 selling
- text2vec-large-chineseGanymedeNil
- vit-base-patch16-224-in21kGoogle
- voyage-code-4voyageai$0.12per Mtok in2 selling
- voyage-context-4Voyage AI$0.12per Mtok in1 selling
- voyage-large-2-instructVoyage AI15th of 238$0.12per Mtok in1 selling
- voyage-multimodal-3Voyage AI$0.12per Mtok in1 selling
- w2v-bert-2.0facebook
- z-image-turbo-flow-dpoF16