GPT-4.1 fine-tuned
GPT-4.1, OpenAI's smartest non-reasoning model for instruction following and tool calling, listed as supporting the v1/fine-tuning endpoint; snapshot gpt-4.1-2025-04-14 is a supervised fine-tuning target.
text + image → text · made by OpenAI
Sold by 5 ways
| Seller | Lane | Rate |
|---|---|---|
| Microsoft Azure AI cloud | batch | $1.1per Mtok in$4.4per Mtok outraw.githubusercontent.com · read 2026-09-16 |
| OpenAI api | batch | $1.5per Mtok in$6per Mtok outdevelopers.openai.com · read 2026-08-24 |
| Microsoft Azure AI cloud | standard | $2.2per Mtok in$8.8per Mtok out$0.55per Mtok cachedraw.githubusercontent.com · read 2026-09-16 |
| OpenAI api | fine-tuned | $3per Mtok in$12per Mtok outdevelopers.openai.com · read 2026-08-24 |
| ElectronHub aggregator | standard | $1.5per Mtok in$6per Mtok outapi.electronhub.ai · read 2026-09-16 |
Measured 30 standings
| Place | Board | Metric | Score |
|---|---|---|---|
| 7thof 81 | Forecastbench — Epoch AI | Overall score | 61.5 |
| 8thof 15 | Cad eval — Epoch AI | Overall pass (%) | 0.42 |
| 18thof 42 | Fictionlivebench — Epoch AI | 120k token score | 0.625 |
| 21stof 105 | Vectara Hallucination Leaderboard | Hallucination Rate | 5.6 |
| 34thof 35 | Swe bench verified — Epoch AI | mean_score | 48.542 |
| 35thof 108 | Math level 5 — Epoch AI | mean_score | 83.006 |
| 36thof 46 | Enigma eval — Epoch AI | Accuracy | 2.17 |
| 38thof 72 | Aider polyglot — Epoch AI | Percent correct | 52.4 |
| 45thof 51 | Humanity's Last Exam (Epoch AI replication) | Accuracy | 5.4 |
| 58thof 80 | Simpleqa verified — Epoch AI | mean_score | 31.1 |
| 67thof 72 | Frontiermath tier 4 — Epoch AI | mean_score | 0 |
| 68thof 101 | FrontierMath (Tiers 1-3) | mean_score | 5.517 |
| 71stof 152 | LMArena · Vision | Score (Elo) | 1214 |
| 77thof 123 | Lmca — Epoch AI | Score | 25.61 |
| 87thof 102 | Simplebench — Epoch AI | Score (AVG@5) | 0.27 |
| 88thof 112 | Ale bench — Epoch AI | Performance | 558.1 |
| 90thof 161 | Dtbench — Epoch AI | Accuracy | 68.27 |
| 97thof 106 | Frontiermath tiers 1 3 v2 — Epoch AI | mean_score | 5.965 |
| 117thof 170 | Weirdml — Epoch AI | Accuracy | 39.04 |
| 130thof 402 | LMArena · Text | Score (Elo) | 1414 |
| 151stof 222 | Chess puzzles — Epoch AI | mean_score | 6 |
| 185thof 313 | GPQA Diamond | mean_score | 66.919 |
| 190thof 216 | ARC-AGI-2 (Public Eval) | Score (%) | 0 |
| 193rdof 205 | ARC-AGI-1 (Public Eval) | Score (%) | 11.75 |
| 202ndof 290 | OTIS Mock AIME 2024-2025 | mean_score | 38.333 |
| 206thof 227 | Arc agi 2 — Epoch AI | Score | 0.42 |
| 211thof 232 | ARC-AGI-2 (Semi-Private) | Score (%) | 0.42 |
| 218thof 230 | ARC-AGI-1 (Semi-Private) | Score (%) | 5.5 |
| 233rdof 244 | Arc agi — Epoch AI | Score | 5.5 |
| 482ndof 656 | Epoch capabilities index — Epoch AI | ECI Score | 136.82 |
About
GPT-4.1 fine-tuned — a text model from OpenAI, sold by 3 companies from $1.5 in and $6 out per million tokens, placed 7th of 81 on Forecastbench — Epoch AI.
It takes text and images and returns text, with a context window of 1,047,576 tokens. The catalogue files it under reasoning and agents. Three companies sell it. The cheapest is $1.5 in and $6 out per million tokens at ElectronHub. Beside the standard rate there are batch and separately routed lanes. It has been measured on 30 boards, and stands best at 7th of 81 on Forecastbench — Epoch AI.
Every current figure
- Maker
- OpenAI
- Register
- model
- Takes
- text + image
- Returns
- text
- Context
- 1,047,576 tokens
- Licence
- not read
- Sellers
- 3
- Price
- $1.5 in and $6 out per million tokens — ElectronHub
- Boards
- 30
- Best place
- 7th of 81 — Forecastbench — Epoch AI
Known as 2 names
gpt-4.1-ftgpt-4.1-2025-04-14