Pass IndexThe State of AISign in

Fictionlivebench — Epoch AI

Fictionlivebench — Epoch AI is run by Epoch AI. It has ranked 42 models, of which the catalogue holds 30, scoring from 0.188 to 1.

measured by Epoch AI · the board itself

PlaceModelMetricScore
1stof 42o3120k token score (medium)1
2ndof 42GPT-5120k token score (medium)0.969
5thof 42o3-pro120k token score (medium)0.889
6thof 42Gemini 2.5 Pro Preview 06-05120k token score0.875
7thof 42Kimi K2.5120k token score0.781
8thof 42Grok 4 Fast120k token score0.75
9thof 42Gemini 2.5 Pro Preview 05-06120k token score0.719
10thof 42Gemini 2.5 Pro120k token score0.719
11thof 42Qwen3 235B A22B Thinking 2507120k token score0.688
12thof 42Gemini 2.5 Flash120k token score0.688
14thof 42Grok 3 Mini120k token score (medium)0.656
16thof 42Qwen3 Next 80B A3B Instruct120k token score0.625
17thof 42gemini-2.0-flash-001120k token score0.625
18thof 42GPT-4.1 fine-tuned120k token score0.625
19thof 42GPT-5 mini120k token score (medium)0.625
20thof 42o4-mini120k token score (medium)0.625
21stof 42MiniMax M1120k token score0.594
22ndof 42Grok 3120k token score0.583
23rdof 42Claude 3.7 Sonnet120k token score0.531
24thof 42o1120k token score (medium)0.531
25thof 42DeepSeek-V3.1120k token score0.531
28thof 42GPT-4.1 mini120k token score0.469
29thof 42parasail-qwen3-235b-a22b-instruct-2507120k token score0.444
30thof 42o3-mini120k token score (medium)0.438
32ndof 42Claude Opus 4120k token score0.375
35thof 42Llama-4-Maverick-17B-128E-Instruct120k token score0.364
36thof 42Claude Sonnet 4120k token score0.364
38thof 42DeepSeek-R1120k token score0.333
41stof 42GPT-5 nano120k token score (medium)0.219
42ndof 42GPT-4.1 nano120k token score0.188

Read from the board on 2026-09-16