Skip to content
ModelCap

Benchmark leaderboard

Arena coding leaderboard

Arena's coding category scores the same blind pairwise votes restricted to programming prompts, so it reflects which models people preferred for code, not a fixed test set. This page lists every model ModelCap tracks with a published Arena coding result, its best evaluated configuration, the source's own rank, and the model's ModelCap Index position, API price and context window.

Snapshot as of 10 August 2026

Claude Fable 5 from Anthropic leads the Arena coding leaderboard among the 134 tracked models with 1554, per the source snapshot published 10 August 2026.

Tracked models with a result
134
Entries on the source board
381
Open-weight models listed
38

Arena coding standings (134 models)

Sorted by published score, one row per model. Source: Arena (LMArena).

Arena coding leaderboard: every tracked model's published score, source rank, ModelCap Index position, price and context
#ModelRatingConfigurationSource rankModelCap IndexOutput $/1MContextWeightsPublished
1Claude Fable 5Anthropic15541545–1562Default#1 / 381#1$50.001MAPI only10 August 2026
2Claude Opus 4.6Anthropic15521546–1558Thinking+1 more evaluated#2 / 381$25.001MAPI only10 August 2026
3Claude Opus 4.7Anthropic15511545–1558Thinking+1 more evaluated#3 / 381$25.001MAPI only10 August 2026
4Kimi K3Moonshot AI15411530–1552Max#6 / 381#4$15.001MRestricted license10 August 2026
5Claude Opus 4.8Anthropic15331526–1540Thinking+1 more evaluated#7 / 381$25.001MAPI only10 August 2026
6Muse Spark 1.2Meta15321512–1551X-High+1 more evaluated#8 / 381#3$4.251MAPI only10 August 2026
7Muse Spark 1.1Meta15311522–1540Default#9 / 381$4.251MAPI only10 August 2026
8Claude Opus 5Anthropic15301521–1539High+1 more evaluated#10 / 381#2$25.001MAPI only10 August 2026
9Claude Opus 4.5Anthropic15301522–1537Thinking · 32K+1 more evaluated#11 / 381$25.00200KAPI only10 August 2026
10Claude Sonnet 4.6Anthropic15291523–1535Default#13 / 381$15.001MAPI only10 August 2026
11GPT-5.6 SolOpenAI15281519–1538X-High#14 / 381#7$30.001MAPI only10 August 2026
12Qwen3.8 MaxQwen15281515–1542Default#15 / 381#5$6.001MAPI only10 August 2026
13Qwen3.7 MaxQwen15251506–1543Preview#18 / 381$4.421MAPI only10 August 2026
14Claude Sonnet 5Anthropic15231515–1532High#19 / 381#14$10.001MAPI only10 August 2026
15Gemini 3.6 FlashGoogle15231513–1532Default#21 / 381#8$7.501MAPI only10 August 2026
16Grok 4.5xAI15221513–1531Default#22 / 381#12$6.00500KAPI only10 August 2026
17GPT-5.4OpenAI15211515–1527High+1 more evaluated#23 / 381$15.001MAPI only10 August 2026
18Gemini 3.1 Pro PreviewGoogle15211515–1526Default+1 more evaluated#24 / 381#6$12.001MAPI only10 August 2026
19GPT-5.5OpenAI15201513–1526High+2 more evaluated#25 / 381#10$30.001MAPI only10 August 2026
20Claude Sonnet 4.5Anthropic15191515–1524Thinking · 32K+1 more evaluated#26 / 381$15.001MAPI only10 August 2026
21GPT-5.6 TerraOpenAI15181509–1528X-High#29 / 381#15$6.001MAPI only10 August 2026
22GLM 5.1Z.ai15151508–1522Default#30 / 381$4.40205KOpen weights10 August 2026
23GPT-5.2 ChatOpenAI15151508–1522Default#31 / 381#9$14.00128KAPI only10 August 2026
24Kimi K2.6Moonshot AI15141507–1521Default#32 / 381$4.00262KRestricted license10 August 2026
25Claude Opus 4.1Anthropic15121506–1519Thinking · 16K+1 more evaluated#39 / 381$75.00200KAPI only10 August 2026
26Grok 4.20xAI15111505–1516Beta · 0309 · Reasoning+1 more evaluated#40 / 381$2.502MAPI only10 August 2026
27Qwen3.6 Max PreviewQwen15091493–1524Default#42 / 381$6.16262KAPI only10 August 2026
28Gemini 3.5 FlashGoogle15091501–1516High+1 more evaluated#43 / 381$9.001MAPI only10 August 2026
29Grok 4.20 Multi-AgentxAI15081502–1514Beta · 0309#44 / 381#11$2.502MAPI only10 August 2026
30Gemini 3 Flash PreviewGoogle15081500–1516Default+1 more evaluated#46 / 381$3.001MAPI only10 August 2026
31Qwen3.7 PlusQwen15071500–1514Default#48 / 381#16$1.281MAPI only10 August 2026
32GLM 5.2Z.ai15061499–1514Max#49 / 381#13$2.421MOpen weights10 August 2026
33Gemini 3.5 Flash LiteGoogle15041494–1514Default#52 / 381#17$2.501MAPI only10 August 2026
34Kimi K2.5Moonshot AI15011496–1507Thinking#56 / 381$2.85262KRestricted license10 August 2026
35Hy3Tencent15011484–1517Default#57 / 381#18$0.528262KOpen weights10 August 2026
36Claude Opus 4Anthropic14991491–1507Thinking · 16K+1 more evaluated#58 / 381$75.00200KAPI only10 August 2026
37Gemma 4 31BGoogle14991483–1514Default#60 / 381#20$0.34262KOpen weights10 August 2026
38GPT-5.6 LunaOpenAI14981488–1507X-High#61 / 381#19$0.601MAPI only10 August 2026
39GLM 5Z.ai14971490–1505Default#62 / 381$2.55205KOpen weights10 August 2026
40MiniMax M3MiniMax14971490–1504Default#63 / 381#23$1.201MRestricted license10 August 2026
41GPT-5.4 MiniOpenAI14971491–1503High#64 / 381#21$4.50400KAPI only10 August 2026
42Qwen3.6 PlusQwen14951488–1501Default#66 / 381$1.951MAPI only10 August 2026
43InklingThinking Machines14941484–1503Default#67 / 381#24$4.051MOpen weights10 August 2026
44Qwen3.5 397B A17BQwen14911486–1497Default#69 / 381#22$3.60262KOpen weights10 August 2026
45GPT-5.1OpenAI14911484–1498High+1 more evaluated#71 / 381$10.00400KAPI only10 August 2026
46GLM 5V TurboZ.ai14901478–1502Default#73 / 381#26$4.00203KAPI only10 August 2026
47GPT-5.2OpenAI14901483–1496High+1 more evaluated#74 / 381$14.00400KAPI only10 August 2026
48Grok 4.3xAI14891483–1495Default#77 / 381$2.501MAPI only10 August 2026
49Kimi K2 ThinkingMoonshot AI14861481–1492Turbo#80 / 381#27$2.50262KRestricted license10 August 2026
50GLM 4.7Z.ai14851473–1497Default#81 / 381$1.75205KOpen weights10 August 2026
51Gemma 4 26B A4BGoogle14811466–1496Default#86 / 381#25$0.40262KOpen weights10 August 2026
52Claude Haiku 4.5Anthropic14791475–1484Default#89 / 381#36$5.00200KAPI only10 August 2026
53Mistral Medium 3.5Mistral AI14791468–1490Default#91 / 381#30$7.50262KAPI only10 August 2026
54Claude Sonnet 4Anthropic14731466–1481Thinking · 32K+1 more evaluated#97 / 381$15.001MAPI only10 August 2026
55Qwen3 MaxQwen14731460–1486Default#98 / 381$3.90262KAPI only10 August 2026
56Qwen3 235B A22B Instruct 2507Qwen14721468–1477Default#99 / 381#31$0.55262KOpen weights10 August 2026
57GPT-5OpenAI14691461–1477High#102 / 381$10.00400KAPI only10 August 2026
58Mistral Large 3 2512Mistral AI14681463–1474Default#104 / 381#35$1.50262KAPI only10 August 2026
59Gemini 2.5 ProGoogle14651461–1469Default#106 / 381$10.001MAPI only10 August 2026
60Qwen3 VL 235B A22B InstructQwen14641452–1477Default#108 / 381#34$1.90262KOpen weights10 August 2026
61Hy3 previewTencent14601447–1474Default#115 / 381$0.21262KRestricted license10 August 2026
62GPT-5.4 NanoOpenAI14601454–1466High#116 / 381#39$1.25400KAPI only10 August 2026
63o3OpenAI14601453–1466Default#117 / 381#29$8.00200KAPI only10 August 2026
64Qwen3.5-122B-A10BQwen14591452–1466Default#118 / 381#33$2.40262KOpen weights10 August 2026
65GLM 4.6Z.ai14581451–1465Default#120 / 381$2.00205KOpen weights10 August 2026
66Gemini 3.1 Flash Lite PreviewGoogle14571451–1463Default#122 / 381$1.501MAPI only10 August 2026
67Qwen3 Coder 480B A35BQwen14571448–1466Default#123 / 381#46$1.00262KOpen weights10 August 2026
68GPT-4.1OpenAI14571450–1463Default#124 / 381$8.001MAPI only10 August 2026
69Mistral Medium 3.1Mistral AI14551450–1460Default#126 / 381$2.00131KAPI only10 August 2026
70GLM 4.5Z.ai14551446–1464Default#127 / 381$2.20131KOpen weights10 August 2026
71Qwen3 VL 235B A22B ThinkingQwen14551440–1469Default#128 / 381#42$4.00131KOpen weights10 August 2026
72Step 3.5 FlashStepFun14511445–1457Default#130 / 381$0.30262KOpen weights10 August 2026
73Qwen3.5-27BQwen14501442–1457Default#131 / 381$1.56262KOpen weights10 August 2026
74Qwen3 Next 80B A3B InstructQwen14461437–1454Default#135 / 381#40$1.10262KOpen weights10 August 2026
75Qwen3 235B A22BQwen14461438–1453No thinking+1 more evaluated#136 / 381#51$1.82131KOpen weights10 August 2026
76Qwen3 235B A22B Thinking 2507Qwen14421428–1457Default#141 / 381#43$2.30262KOpen weights10 August 2026
77Qwen3 30B A3B Instruct 2507Qwen14401431–1448Default#143 / 381#48$0.193262KOpen weights10 August 2026
78Qwen3.5-FlashQwen14371431–1443Default#146 / 381$0.261MAPI only10 August 2026
79Qwen3.5-35B-A3BQwen14351428–1442Default#149 / 381$1.00262KOpen weights10 August 2026
80GPT-4.1 MiniOpenAI14331426–1441Default#152 / 381$1.601MAPI only10 August 2026
81Mistral Medium 3Mistral AI14331425–1441Default#153 / 381$2.00131KAPI only10 August 2026
82o1OpenAI14331423–1443Default#154 / 381$60.00200KAPI only10 August 2026
83o4 MiniOpenAI14331426–1440Default#157 / 381#47$4.40200KAPI only10 August 2026
84GPT-5 MiniOpenAI14301422–1439High#160 / 381$2.00400KAPI only10 August 2026
85GLM 4.5 AirZ.ai14261418–1434Default#164 / 381#54$0.85131KOpen weights10 August 2026
86GLM 4.7 FlashZ.ai14241413–1435Default#165 / 381#62$0.40203KOpen weights10 August 2026
87Gemini 2.5 FlashGoogle14241419–1428Default#166 / 381$2.501MAPI only10 August 2026
88Qwen3 Next 80B A3B ThinkingQwen14201409–1432Default#167 / 381#59$1.20262KOpen weights10 August 2026
89GLM 4.6VZ.ai14171392–1442Default#169 / 381#58$0.90131KOpen weights10 August 2026
90o3 MiniOpenAI14161410–1423Default#170 / 381$4.40200KAPI only10 August 2026
91MiniMax M1MiniMax14161408–1423Default#172 / 381$2.201MAPI only10 August 2026
92Trinity Large ThinkingArcee AI14141407–1422Default#173 / 381#60$0.85262KRestricted license10 August 2026
93Mistral Small 3.2 24BMistral AI14121402–1422Default#174 / 381#68$0.25256KOpen weights10 August 2026
94Nemotron 3 SuperNVIDIA14081394–1422Default#178 / 381#67$0.401MRestricted license10 August 2026
95Qwen3 32BQwen14071383–1431Default#180 / 381#72$0.28131KOpen weights10 August 2026
96GLM 4.5VZ.ai14041386–1423Default#181 / 381$1.8066KOpen weights10 August 2026
97Nova 2 LiteAmazon13951383–1406Default#188 / 381#80$2.501MAPI only10 August 2026
98Mercury 2Inception Labs13941373–1415Default#190 / 381#73$0.75128KAPI only10 August 2026
99Command ACohere13891383–1396Default#196 / 381#69$10.00256KGated access10 August 2026
100Qwen3 30B A3BQwen13861377–1395Default#201 / 381#180$0.50131KOpen weights10 August 2026
101GPT-5 NanoOpenAI13841370–1399High#205 / 381$0.40400KAPI only10 August 2026
102GPT-4.1 NanoOpenAI13741355–1393Default#211 / 381$0.401MAPI only10 August 2026
103Llama 4 MaverickMeta13731365–1380Default#212 / 381#181$0.6961MGated access10 August 2026
104GPT-4o (2024-05-13)OpenAI13691363–1376Default#216 / 381$15.00128KAPI only10 August 2026
105Olmo 3 32B ThinkAllen Institute13641345–1382Default#224 / 381#19666KOpen weights10 August 2026
106Nemotron 3 Nano 30B A3BNVIDIA13621352–1373BF16#226 / 381#195$0.20262KRestricted license10 August 2026
107Llama 4 ScoutMeta13621353–1370Default#227 / 381#188$0.301.3MGated access10 August 2026
108Mistral Small 3.1 24BMistral AI13621354–1369Default#228 / 381$0.555128KOpen weights10 August 2026
109GPT-4o (2024-08-06)OpenAI13601352–1368Default#229 / 381$10.00128KAPI only10 August 2026
110Gemma 3 27BGoogle13581351–1365Default#231 / 381#65$0.45262KGated access10 August 2026
111Qwen2.5 72B InstructQwen13561348–1363Default#234 / 381#198$0.4033KRestricted license10 August 2026
112Granite 4.1 8BIBM Granite13541334–1374Default#237 / 381#197$0.10131KOpen weights10 August 2026
113Mistral Large 2407Mistral AI13541346–1362Default#238 / 381$6.00131KAPI only10 August 2026
114GPT-4o-mini (2024-07-18)OpenAI13491342–1355Default#243 / 381$0.60128KAPI only10 August 2026
115GPT-4 TurboOpenAI13471340–1354Default#244 / 381#187$30.00128KAPI only10 August 2026
116Llama 3.3 70B InstructMeta13451339–1352Default#247 / 381#193$0.32131KGated access10 August 2026
117Nova Pro 1.0Amazon13431334–1352Default#248 / 381#199$3.20300KAPI only10 August 2026
118Qwen2.5 Coder 32B InstructQwen13421324–1361Default#251 / 381#205$1.0033KOpen weights10 August 2026
119Llama 3.1 70B InstructMeta13331326–1340Default#256 / 381$0.40131KGated access10 August 2026
120Gemma 3 12BGoogle13161293–1339Default#264 / 381#93$0.15131KGated access10 August 2026
121Mistral Small 3Mistral AI13121300–1325Default#268 / 381$0.0833KOpen weights10 August 2026
122Gemma 3n 4BGoogle13081298–1318Default#275 / 381#194$0.1233KGated access10 August 2026
123Phi 4Microsoft13061296–1316Default#276 / 381#208$0.1416KOpen weights10 August 2026
124Nova Lite 1.0Amazon13061295–1316Default#278 / 381$0.24300KAPI only10 August 2026
125Gemma 2 27BGoogle13051298–1311Default#279 / 381$0.658KGated access10 August 2026
126Claude 3 HaikuAnthropic13011294–1308Default#280 / 381$1.25200KAPI only10 August 2026
127Nova Micro 1.0Amazon12881278–1299Default#286 / 381#213$0.14128KAPI only10 August 2026
128Command R (08-2024)Cohere12811268–1295Default#288 / 381#211$0.60128KAPI only10 August 2026
129Command R+ (08-2024)Cohere12801266–1294Default#289 / 381#206$10.00128KAPI only10 August 2026
130Mixtral 8x22B InstructMistral AI12771268–1286Default#293 / 381#215$6.0066KOpen weights10 August 2026
131Gemma 3 4BGoogle12741250–1297Default#297 / 381#200$0.10131KGated access10 August 2026
132Llama 3.1 8B InstructMeta12601252–1267Default#307 / 381#219$0.08131KGated access10 August 2026
133Llama 3.2 3B InstructMeta11761160–1192Default#345 / 381#226$0.33131KGated access10 August 2026
134Llama 3.2 1B InstructMeta11491133–1165Default#360 / 381#227$0.20160KGated access10 August 2026

How ModelCap uses Arena coding

Arena coding ranks models by human preference rating on coding prompts. ModelCap ingests the board as published, matches each entry to a catalogue model with a reviewed identity, and shows the source's score, interval and rank unchanged. Where a model has an admitted result, it feeds the ModelCap Index as coding evidence alongside the other public boards; the methodology documents the weighting and the identity rules.

Arena coding leaderboard: common questions

What is the Arena coding benchmark?

Arena's coding category scores the same blind pairwise votes restricted to programming prompts, so it reflects which models people preferred for code, not a fixed test set. It is published by Arena (LMArena).

Which AI model leads Arena coding right now?

Claude Fable 5 (Anthropic) holds the top Arena coding score among the models ModelCap tracks, at 1554 as of the source snapshot published 10 August 2026.

How many models are ranked on the Arena coding leaderboard here?

134 tracked models have a published Arena coding result on ModelCap; the source board itself lists 381 entries. Each row shows the model's best evaluated configuration.

What is the best open-weight model on Arena coding?

GLM 5.1 from Z.ai is the highest-scoring model with openly downloadable weights on this board, at 1515.

Which model offers the best value on Arena coding?

Among the ten highest-scoring priced models, Muse Spark 1.2 has the lowest listed output price at $4.25 per 1M tokens while scoring 1532.

How is Arena coding scored?

The board ranks models by human preference rating on coding prompts; higher ratings are better. ModelCap shows the source's own score, interval and rank and never re-runs the evaluation.

Does Arena coding decide the ModelCap Index rank?

Not on its own. The ModelCap Index combines several public capability sources with published uncertainty; Arena coding contributes as coding evidence where a model has an admitted result. The methodology page documents the weighting.

How recent are the Arena coding results?

The newest Arena coding publication ModelCap holds is dated 10 August 2026. The page re-renders every minute from the sealed dataset, so it reflects the latest refresh of the source board.

Explore further