Pass IndexThe State of AISign in

Enigma eval — Epoch AI

Enigma eval — Epoch AI is run by Epoch AI. It has ranked 46 models, of which the catalogue holds 35, scoring from 2.17 to 91.

measured by Epoch AI · the board itself

PlaceModelMetricScore
1stof 46Claude Fable 5Accuracy (high)39.28
2ndof 46GPT-5.6 SolAccuracy (high)37.12
3rdof 46Gemini 3.1 Pro PreviewAccuracy36.78
4thof 46Gemini 3.5 FlashAccuracy (high)25.41
5thof 46GPT-5.4 ProAccuracy23.82
6thof 46Claude Opus 4.8Accuracy (xhigh)23.51
7thof 46GPT-5 ProAccuracy18.75
8thof 46gemini-3-pro-previewAccuracy18.24
9thof 46GPT-5.4Accuracy (xhigh)15.96
10thof 46o3Accuracy (medium)13.09
11thof 46Claude Opus 4.5Accuracy11.91
13thof 46GPT-5.1Accuracy11.23
14thof 46GPT-5Accuracy10.47
15thof 46GPT-5.2Accuracy10.39
16thof 46o4-miniAccuracy (high)9.21
17thof 46GPT-5 miniAccuracy8.193
18thof 46Claude Opus 4.6Accuracy (max)7.6
19thof 46Claude Opus 4.1Accuracy7.179
22ndof 46o1-proAccuracy6.14
23rdof 46Claude Sonnet 4.5Accuracy5.997
24thof 46o1Accuracy5.65
25thof 46Claude Opus 4Accuracy5.57
26thof 46Gemini 2.5 Pro Preview 06-05Accuracy5.57
27thof 46Claude 3.7 SonnetAccuracy4.23
29thof 46Kimi K2.5Accuracy3.38
31stof 46Claude Sonnet 4Accuracy3.12
32ndof 46Gemini 3.1 Flash-LiteAccuracy3.04
33rdof 46Gemini 2.5 FlashAccuracy2.7
34thof 46Gemini 2.5 Pro Preview 05-06Accuracy2.36
36thof 46GPT-4.1 fine-tunedAccuracy2.17
39thof 46Claude 3.5 SonnetAccuracy91
41stof 46Claude 3 OpusAccuracy82
42ndof 46GPT-4o (2024-11-20)Accuracy80
45thof 46Llama-4-Maverick-17B-128E-InstructAccuracy58
46thof 46Llama-3.2-90B-Vision-InstructAccuracy38

Read from the board on 2026-09-16