o3
OpenAI's o3 is their most powerful reasoning model, setting new state-of-the-art benchmarks in coding, math, science, and visual perception.
text + image + file → text · made by OpenAI
Sold by 12 ways
| Seller | Lane | Rate |
|---|---|---|
| Microsoft Azure AI cloud | batch | $1per Mtok in$4per Mtok outprices.azure.com · read 2026-08-24 |
| DigitalOcean cloud | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cachedmodels.dev · read 2026-09-16 |
| Microsoft Azure AI cloud | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cachedmodels.dev · read 2026-09-16 |
| OpenAI api | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cacheddevelopers.openai.com · read 2026-08-24 |
| Requesty aggregatornot on the seller's list today | openai flex | $0.9per Mtok in$3.6per Mtok out$0.23per Mtok cachedrouter.requesty.ai · read 2026-08-25 |
| Requesty aggregator | openai | $2per Mtok in$8per Mtok out$0.5per Mtok cachedrouter.requesty.ai · read 2026-09-10 |
| Nous Research 2 ways | $2per Mtok in$8per Mtok out$0.5per Mtok cached | |
| Nous Research aggregator | batch | $1per Mtok in$4per Mtok out$0.25per Mtok cachedinference-api.nousresearch.com · read 2026-09-12 |
| Nous Research aggregator | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cachedinference-api.nousresearch.com · read 2026-09-12 |
| OpenRouter 2 ways | $2per Mtok in$8per Mtok out$0.5per Mtok cached | |
| OpenRouter aggregator | batch | $1per Mtok in$4per Mtok out$0.25per Mtok cachedopenrouter.ai · read 2026-08-24 |
| OpenRouter aggregator | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cachedopenrouter.ai · read 2026-08-24 |
| ElectronHub aggregator | standard | $1.5per Mtok in$6per Mtok outapi.electronhub.ai · read 2026-09-16 |
| Vercel AI Gateway aggregator | standard | $2per Mtok in$8per Mtok out$0.5per Mtok cachedai-gateway.vercel.sh · read 2026-08-25 |
Measured 52 standings
| Place | Board | Metric | Score |
|---|---|---|---|
| 1stof 15 | Cad eval — Epoch AI | Overall pass (%) (medium) | 0.74 |
| 1stof 42 | Fictionlivebench — Epoch AI | 120k token score (medium) | 1 |
| 1stof 81 | Forecastbench — Epoch AI | Overall score | 62.5 |
| 1stof 22 | SWE-bench Multimodal | % Resolved | 35.98 |
| 2ndof 28 | LiveCodeBench Leaderboard (default window 8/1/2024–5/1/2025, 454 problems) | Pass@1 (high) | 75.8 |
| 5thof 108 | Math level 5 — Epoch AI | mean_score (high) | 97.772 |
| 6thof 69 | Aider polyglot coding leaderboard | Percent correct (high) | 81.3 |
| 6thof 72 | Aider polyglot — Epoch AI | Percent correct (high) | 81.3 |
| 6thof 49 | Lech mazur writing — Epoch AI | Mean score (medium) | 8.39 |
| 7thof 11 | Gdpval — Epoch AI | Win Rate (%) (medium) | 0.308 |
| 8thof 109 | Berkeley Function-Calling Leaderboard (BFCL) V4 | Overall Acc (prompt) | 63.05 |
| 10thof 46 | Enigma eval — Epoch AI | Accuracy (medium) | 13.09 |
| 10thof 38 | Vpct — Epoch AI | Correct (medium) | 0.52 |
| 14thof 23 | Cl bench — Epoch AI | Overall (high) | 0.178 |
| 16thof 51 | Humanity's Last Exam (Epoch AI replication) | Accuracy (high) | 20.32 |
| 16thof 20 | Os world — Epoch AI | Score (medium) | 23 |
| 17thof 50 | Metr time horizons — Epoch AI | average_score (medium) | 0.654 |
| 21stof 61 | AIME 2025 | Accuracy (± 95% CI) (high) | 89.17 |
| 21stof 38 | Gso — Epoch AI | Score OPT@1 (high) | 0.088 |
| 22ndof 80 | Simpleqa verified — Epoch AI | mean_score (high) | 49.4 |
| 26thof 34 | LMArena · Search | Score (Elo) | 1144 |
| 27thof 127 | Mystery game puzzles — Epoch AI | mean_score (high) | 29 |
| 30thof 41 | Deepresearchbench — Epoch AI | Average score (medium) | 0.452 |
| 30thof 35 | Swe bench verified — Epoch AI | mean_score (medium) | 62.319 |
| 31stof 222 | Chess puzzles — Epoch AI | mean_score (medium) | 38 |
| 36thof 123 | Lmca — Epoch AI | Score (high) | 39.7 |
| 43rdof 65 | Apex agents — Epoch AI | Pass@1 score (high) | 0.172 |
| 45thof 101 | FrontierMath (Tiers 1-3) | mean_score (high) | 18.685 |
| 46thof 161 | Dtbench — Epoch AI | Accuracy (high) | 84.8 |
| 47thof 112 | Ale bench — Epoch AI | Performance (high) | 933.55 |
| 47thof 102 | Simplebench — Epoch AI | Score (AVG@5) (high) | 0.531 |
| 58thof 72 | Frontiermath tier 4 — Epoch AI | mean_score (high) | 2.083 |
| 58thof 106 | Frontiermath tiers 1 3 v2 — Epoch AI | mean_score (high) | 33.333 |
| 64thof 170 | Weirdml — Epoch AI | Accuracy (high) | 52.42 |
| 68thof 152 | LMArena · Vision | Score (Elo) | 1215 |
| 74thof 152 | Design Arena · dataviz (models) | Elo | 1200 |
| 87thof 290 | OTIS Mock AIME 2024-2025 | mean_score (medium) | 84.444 |
| 99thof 402 | LMArena · Text | Score (Elo) | 1432 |
| 100thof 230 | ARC-AGI-1 (Semi-Private) | Score (%) (low) | 75.7 |
| 108thof 177 | Critpt — Epoch AI | Accuracy (high) | 1.4 |
| 111thof 205 | ARC-AGI-1 (Public Eval) | Score (%) (high) | 64.25 |
| 114thof 313 | GPQA Diamond | mean_score (high) | 81.818 |
| 119thof 302 | Artificial Analysis · Intelligence Index | Artificial Analysis Intelligence Index | 20 |
| 128thof 216 | ARC-AGI-2 (Public Eval) | Score (%) (medium) | 4.49 |
| 128thof 232 | ARC-AGI-2 (Semi-Private) | Score (%) (high) | 6.53 |
| 128thof 244 | Arc agi — Epoch AI | Score (high) | 60.83 |
| 129thof 150 | Design Arena · uicomponent (models) | Elo | 1032 |
| 130thof 227 | Arc agi 2 — Epoch AI | Score (high) | 6.53 |
| 130thof 155 | Design Arena · gamedev (models) | Elo | 1057 |
| 136thof 156 | Design Arena · codecategories (models) | Elo | 1035 |
| 138thof 162 | Design Arena · website (models) | Elo | 1049 |
| 267thof 656 | Epoch capabilities index — Epoch AI | ECI Score (high) | 146.91 |
About
o3 — a text model from OpenAI, sold by 8 companies from $1.5 in and $6 out per million tokens, placed 1st of 81 on Forecastbench — Epoch AI.
It takes text, images and files and returns text, with a context window of 200,000 tokens. It was published in June 2024, trained on material up to May 2024. Its sellers say it can reason step by step and call a tool. The catalogue files it under reasoning, code and agents. Eight companies sell it. The cheapest is $1.5 in and $6 out per million tokens at ElectronHub. Beside the standard rate there are batch and separately routed lanes. It has been measured on 52 boards, and stands best at 1st of 81 on Forecastbench — Epoch AI.
Every current figure
- Maker
- OpenAI
- Register
- model
- Takes
- text + image + file
- Returns
- text
- Context
- 200,000 tokens
- Longest answer
- 131,072 tokens
- Published
- June 2024
- Knowledge to
- May 2024
- Licence
- not read
- Sellers
- 8
- Price
- $1.5 in and $6 out per million tokens — ElectronHub
- Boards
- 52
- Best place
- 1st of 81 — Forecastbench — Epoch AI
Known as 10 names
o3 (high)o3-2025-04-16 (Prompt)o3 (medium)o3-searchGUIRepair + o3 (2025-04-16)o3 0416openai/o3openai/o3:batchopenai/o3:flexOpenAI: o3 (batch)