Skip to content
ModelCap

Benchmark leaderboard

WildClawBench OpenClaw leaderboard

WildClawBench OpenClaw is a public benchmark board published by WildClawBench. ModelCap records each tracked model's published score and position on it without re-running the evaluation. This page lists every model ModelCap tracks with a published WildClawBench OpenClaw result, its best evaluated configuration, the source's own rank, and the model's ModelCap Index position, API price and context window.

Snapshot as of 20 July 2026

GPT-5.6 Sol from OpenAI leads the WildClawBench OpenClaw leaderboard among the 24 tracked models with 67.2, per the source snapshot published 20 July 2026.

Tracked models with a result
24
Entries on the source board
34
Open-weight models listed
9

WildClawBench OpenClaw standings (24 models)

Sorted by published score, one row per model. Source: WildClawBench.

WildClawBench OpenClaw leaderboard: every tracked model's published score, source rank, ModelCap Index position, price and context
#ModelScoreConfigurationSource rankModelCap IndexOutput $/1MContextWeightsPublished
1GPT-5.6 SolOpenAI67.2Default#1 / 34#5$15.001MAPI only20 July 2026
2Claude Opus 4.8Anthropic64.7Default#2 / 34$25.001MAPI only20 July 2026
3Claude Opus 4.7Anthropic62.2Default#3 / 34$25.001MAPI only20 July 2026
4Claude Fable 5Anthropic62.0Default#4 / 34#1$50.001MAPI only20 July 2026
5GPT-5.5OpenAI58.2Default#5 / 34#8$30.001MAPI only20 July 2026
6Grok 4.5xAI57.5Default#6 / 34$6.00500KAPI only20 July 2026
7Qwen3.8 MaxQwen56.2Default#7 / 34#6$6.001MAPI only20 July 2026
8Muse Spark 1.1Meta54.8Default#8 / 34$4.251MAPI only20 July 2026
9Kimi K3Moonshot AI54.5Default#9 / 34#3$15.001MRestricted license20 July 2026
10GLM 5.2Z.ai54.2Default#10 / 34#11$3.151MOpen weights20 July 2026
11Claude Opus 4.6Anthropic51.6Default#11 / 34$25.001MAPI only20 July 2026
12GPT-5.4OpenAI50.3Default#12 / 34$15.001MAPI only20 July 2026
13Hy3Tencent49.7Default#13 / 34#18$0.528262KOpen weights20 July 2026
14GLM 5.1Z.ai48.2Default#14 / 34$3.04205KOpen weights20 July 2026
15Qwen3.8 27BQwen48.0Default#15 / 34#16$3.20262KOpen weights20 July 2026
16Muse Glimmer 30BMeta47.6Default#16 / 34#190$1.50131KOpen weights20 July 2026
17Kimi K2.7 CodeMoonshot AI46.9Default#17 / 34#26$3.50262KRestricted license20 July 2026
18Qwen3.6 27BQwen43.2Default#20 / 34$3.60262KOpen weights20 July 2026
19GLM 5Z.ai42.6Default#22 / 34$1.92205KOpen weights20 July 2026
20Gemma 4 31BGoogle37.6Default#25 / 34#30$0.34262KOpen weights20 July 2026
21GLM 5 TurboZ.ai33.9Default#28 / 34#39$4.00203KAPI only20 July 2026
22Kimi K2.5Moonshot AI30.8Default#30 / 34$2.85262KRestricted license20 July 2026
23Step 3.5 FlashStepFun26.7Default#33 / 34$0.30262KOpen weights20 July 2026
24Grok 4.20xAI19.3Default#34 / 34$2.502MAPI only20 July 2026

How ModelCap uses WildClawBench OpenClaw

WildClawBench OpenClaw ranks models by the published WildClawBench OpenClaw score. ModelCap ingests the board as published, matches each entry to a catalogue model with a reviewed identity, and shows the source's score, interval and rank unchanged. Where a model has an admitted result, it feeds the ModelCap Index as agent evidence alongside the other public boards; the methodology documents the weighting and the identity rules.

WildClawBench OpenClaw leaderboard: common questions

What is the WildClawBench OpenClaw benchmark?

WildClawBench OpenClaw is a public benchmark board published by WildClawBench. ModelCap records each tracked model's published score and position on it without re-running the evaluation. It is published by WildClawBench.

Which AI model leads WildClawBench OpenClaw right now?

GPT-5.6 Sol (OpenAI) holds the top WildClawBench OpenClaw score among the models ModelCap tracks, at 67.2 as of the source snapshot published 20 July 2026.

How many models are ranked on the WildClawBench OpenClaw leaderboard here?

24 tracked models have a published WildClawBench OpenClaw result on ModelCap; the source board itself lists 34 entries. Each row shows the model's best evaluated configuration.

What is the best open-weight model on WildClawBench OpenClaw?

GLM 5.2 from Z.ai is the highest-scoring model with openly downloadable weights on this board, at 54.2.

Which model offers the best value on WildClawBench OpenClaw?

Among the ten highest-scoring priced models, GLM 5.2 has the lowest listed output price at $3.15 per 1M tokens while scoring 54.2.

How is WildClawBench OpenClaw scored?

The board ranks models by the published WildClawBench OpenClaw score; higher scores are better. ModelCap shows the source's own score, interval and rank and never re-runs the evaluation.

Does WildClawBench OpenClaw decide the ModelCap Index rank?

Not on its own. The ModelCap Index combines several public capability sources with published uncertainty; WildClawBench OpenClaw contributes as agent evidence where a model has an admitted result. The methodology page documents the weighting.

How recent are the WildClawBench OpenClaw results?

The newest WildClawBench OpenClaw publication ModelCap holds is dated 20 July 2026. The page re-renders every minute from the sealed dataset, so it reflects the latest refresh of the source board.

Explore further