LLM Leaderboard

Benchmark Scores for Frontier AI Models

Newest Model

LM Intelligence Index

62.2

#1 of 63 models

63 models ranked by Overall Index, from 62.2 to 19.0.
463 scores 63 models 8 benchmarksSource-linked scores · LM Board computes the Index

LM Board runs no evaluations. Artificial Analysis (opens in a new tab) publishes 463 of the 463 benchmark scores; no scores are vendor-published. LM Board computes the Index and ranks from those published scores.

Every score links to the measurement it came from, with the date it was retrieved and the settings it was run under. Every model name opens its complete citation record; where the full table fits, score numerals open their source details directly.

Scores retrieved Aug 5, 2026 · 11 models released in the last 45 days

Showing 63 of 63 models. Sorted by LM Intelligence Index, descending. Showing table projection.

Filters
Provider
Filter by provider
Weights
63 / 63

12 columns. Use arrow keys to move through cells, Enter to activate the primary action, and F2 then Tab to reach every action in the current cell. Escape returns to the grid.

Currently sorted descending.
Models ranked by LM Intelligence Index. Sort with a column header; open a model for sources and details.
1
#1Claude Opus 5Anthropicmax
Evidence 7 scores
62.289.130.370.052.693.255.729.1$5.0 / $25.0
2
#2GPT-5.6 SolOpenAImax
Evidence 8 scores
62.188.033.073.747.294.156.172.732.3$5.0 / $30.0
3
#3Kimi K3Moonshot AImax
Evidence 7 scores
61.785.033.474.744.493.558.723.4$3.0 / $15.0
4
#4GPT-5.5OpenAIxhigh
Evidence 8 scores
60.984.331.374.344.393.556.175.927.1$5.0 / $30.0
5
#5GPT-5.6 TerraOpenAImax
Evidence 8 scores
60.488.031.874.041.892.553.971.230.0$2.0 / $12.0
6
#6Claude Fable 5Anthropicmax
Evidence 8 scores
60.084.626.870.053.392.660.263.528.6$10.0 / $50.0
7
#7GPT-5.4OpenAIxhigh
Evidence 8 scores
58.878.330.374.041.692.056.674.023.4$2.5 / $15.0
8
#8Grok 4.5xAIhigh
Evidence 7 scores
57.681.732.667.740.393.154.115.4$2.0 / $6.0
9
#9GPT-5.3-CodexOpenAIxhigh
Evidence 6 scores
57.574.039.991.553.275.416.9$1.8 / $14.0
10
#10GPT-5.6 LunaOpenAImax
Evidence 7 scores
57.480.927.274.037.291.152.620.6$0.20 / $1.2
11
#11Claude Sonnet 5Anthropicmax
Evidence 7 scores
57.180.528.370.739.691.153.616.9$2.0 / $10.0
12
#12Gemini 3.1 Pro PreviewGoogle
Evidence 8 scores
56.973.816.572.744.794.158.977.117.7$2.0 / $12.0
13
#13Claude Opus 4.8Anthropicmax
Evidence 8 scores
56.884.627.667.745.792.053.562.220.9$5.0 / $25.0
14
#14GLM-5.2Z.aimax
Evidence 8 scores
56.377.926.871.340.189.550.573.320.9$1.4 / $4.4
15
#15Muse Spark 1.1Metaxhigh
Evidence 7 scores
56.177.925.263.345.189.858.215.1$1.3 / $4.3
16
#16Gemini 3.5 FlashGooglehigh
Evidence 8 scores
56.178.725.469.341.092.253.176.313.1$1.5 / $9.0
17
#17DeepSeek V4 Flash 0731DeepSeekmax
Evidence 7 scores
55.478.731.165.736.890.849.916.6$0.14 / $0.28
18
#18GPT-5.2OpenAIxhigh
Evidence 6 scores
55.272.735.590.352.175.411.6$1.8 / $14.0
19
#19Gemini 3.6 FlashGooglehigh
Evidence 7 scores
55.177.524.569.738.392.852.710.6$1.5 / $7.5
20
#20Claude Opus 4.7Anthropicmax
Evidence 8 scores
54.883.228.970.339.691.454.558.612.0$5.0 / $25.0
21
#21Gemini 3 Pro PreviewGooglehigh
Evidence 6 scores
54.870.737.290.856.170.49.1
22
#22Qwen3.7-MaxQwen
Evidence 8 scores
53.574.510.969.038.192.348.880.513.4$2.5 / $7.5
23
#23Kimi K2.6Moonshot AI
Evidence 8 scores
52.665.920.669.735.991.153.576.08.0$0.95 / $4.0
24
#24DeepSeek V4 ProDeepSeekmax
Evidence 8 scores
52.564.025.866.335.988.850.076.512.9$0.44 / $0.87
25
#25Muse SparkMeta
Evidence 8 scores
52.362.219.669.739.988.451.575.911.3
26
#26MiniMax M3MiniMax
Evidence 8 scores
51.865.213.074.037.192.945.482.93.7$0.30 / $1.2
27
#27Claude Opus 4.6Anthropicmax
Evidence 6 scores
51.170.736.789.651.953.112.6$5.0 / $25.0
28
#28Grok 4.20xAIreasoning
Evidence 6 scores
50.358.032.291.145.681.26.6$1.3 / $2.5
29
#29DeepSeek V4 FlashDeepSeekmax
Evidence 8 scores
50.061.822.963.032.189.444.979.27.1$0.14 / $0.28
30
#30GPT-5.4 miniOpenAIxhigh
Evidence 8 scores
49.759.221.469.326.787.549.973.310.0$0.75 / $4.5
31
#31Claude Sonnet 4.6Anthropicmax
Evidence 8 scores
49.571.230.570.730.087.546.856.63.1$3.0 / $15.0
32
#32Kimi K2.7 CodeMoonshot AI
Evidence 8 scores
49.467.418.166.332.889.647.563.110.0$0.95 / $4.0
33
#33GPT-5.4 nanoOpenAIxhigh
Evidence 8 scores
48.560.721.066.026.581.746.975.99.3$0.20 / $1.3
34
#34Claude Opus 4.5Anthropicreasoning
Evidence 6 scores
48.374.028.486.649.558.04.6$5.0 / $25.0
35
#35Qwen3.6-PlusQwen
Evidence 8 scores
47.561.416.569.725.788.240.775.22.9$0.50 / $3.0
36
#36GPT-5.1OpenAIhigh
Evidence 8 scores
47.052.414.075.026.587.343.372.94.9$1.3 / $10.0
37
#37GLM-5.1Z.aireasoning
Evidence 8 scores
46.961.811.662.328.086.843.876.34.6$1.4 / $4.4
38
#38Nemotron 3 Ultra 550B-A55BNVIDIAreasoning
Evidence 8 scores
46.553.913.867.026.686.739.981.43.1
39
#39Qwen3.5-397B-A17BQwenreasoning
Evidence 8 scores
46.251.313.465.727.389.342.078.81.7$0.60 / $3.6
40
#40GLM-5Z.aireasoning
Evidence 6 scores
46.163.327.382.046.272.32.0$1.0 / $3.2
41
#41Kimi K2.5Moonshot AIreasoning
Evidence 8 scores
45.645.714.265.329.487.949.070.23.1$0.60 / $3.0
42
#42GPT-5OpenAIhigh
Evidence 8 scores
45.535.219.675.626.585.442.973.15.7$1.3 / $10.0
43
#43o3OpenAI
Evidence 6 scores
44.769.320.082.741.071.41.1$2.0 / $8.0
44
#44Grok 4xAI
Evidence 6 scores
44.768.023.987.745.753.72.0
45
#45Kimi K2 ThinkingMoonshot AI
Evidence 6 scores
44.566.322.383.842.468.12.6
46
#46Claude Sonnet 4.5Anthropicreasoning
Evidence 8 scores
43.055.819.065.717.383.444.757.31.1$3.0 / $15.0
47
#47Grok 4.1 FastxAIreasoning
Evidence 6 scores
42.568.017.685.344.252.72.9
48
#48DeepSeek V3.2DeepSeekreasoning
Evidence 8 scores
42.446.818.865.022.284.038.960.72.9
49
#49Gemini 3.5 Flash-LiteGoogle
Evidence 7 scores
41.253.616.562.017.583.840.90.0$0.30 / $2.5
50
#50Mistral Medium 3.5Mistral AI
Evidence 8 scores
40.250.614.461.012.874.939.668.80.0$1.5 / $7.5
51
#51MiniMax M2MiniMax
Evidence 6 scores
40.161.012.577.736.172.30.9$0.30 / $1.2
52
#52Claude Opus 4.1Anthropicreasoning
Evidence 6 scores
39.766.311.980.940.955.40.0$15.0 / $75.0
53
#53Gemini 2.5 ProGoogle
Evidence 8 scores
37.928.59.366.021.184.442.848.72.6$1.3 / $10.0
54
#54Claude Haiku 4.5Anthropicreasoning
Evidence 8 scores
37.344.29.170.39.767.243.354.30.0$1.0 / $5.0
55
#55gpt-oss-120bOpenAIhigh
Evidence 8 scores
36.826.212.050.718.578.238.969.01.1
56
#56DeepSeek R1 0528DeepSeek
Evidence 6 scores
36.254.714.981.340.339.61.4
57
#57GLM-4.6Z.aireasoning
Evidence 8 scores
36.149.410.554.313.478.038.443.41.1$0.60 / $2.2
58
#58Qwen3-MaxQwen
Evidence 6 scores
32.146.711.176.438.344.20.0$1.2 / $6.0
59
#59Llama 4 MaverickMeta
Evidence 8 scores
25.77.93.946.04.867.133.143.00.0
60
#60Devstral 2Mistral AI
Evidence 8 scores
25.630.310.330.03.659.433.138.10.0
61
#61Mistral Large 3Mistral AI
Evidence 8 scores
24.612.05.834.74.168.036.236.20.0$0.50 / $1.5
62
#62Llama 3.1 Nemotron Ultra 253B v1NVIDIAreasoning
Evidence 6 scores
23.87.38.172.834.738.20.0
63
#63Llama 4 ScoutMeta
Evidence 8 scores
19.03.83.325.84.358.717.039.50.0