Research-only evaluation models
2 in the catalogue today, and the list grows as the market does. Every one with what it costs, who sells it and where it stands.
Tools for judging what a model produced: scoring runs, tracing a chain of calls, catching a regression before a user does. Priced per trace or per evaluation. Most use a model as the judge, which is worth knowing, because it means your evaluation has a bill and an opinion of its own.
Weights published for research and explicitly not for selling. They are here because they exist, are often excellent, and are worth knowing about — and because the cheapest way to learn this is not after you have shipped. If a model on this list is what you want, the maker will usually license it commercially on request; that is a conversation, not a download.
- Isaac 0.2 1BPerceptron$0.15→$1.25per Mtok in / out1 selling
- Isaac 0.2 2B PreviewPerceptron$0.15→$1.25per Mtok in / out1 selling