Which model tops the best-for-agents list right now?
GPT-5.6 Sol (OpenAI), as of 9 September 2026, with LMArena Agent 7.7% · AA τ²-Bench Telecom 85.1% across 2 public agent boards and ModelCap Index position #5.
Agent ranking
Language models measured on at least two public agent boards (LMArena Agent, BFCL, τ²-bench), ordered by the agent component of the ModelCap Index. Models with a single agent result follow in their own tier, ordered by that board, so one result never outranks two.
Snapshot as of 9 September 2026
GPT-5.6 Sol from OpenAI leads the best-for-agents list as of 9 September 2026 with LMArena Agent 7.7% · AA τ²-Bench Telecom 85.1% across 2 public agent boards, holding ModelCap Index position #5; 29 models qualify.
Every row is a current, canonical language model with a public position on the ModelCap Index; the list filters by a published fact and orders by a published figure.
Ranked models with a single admitted agent result, ordered by that board's normalised score. One board is a narrower claim than two, so these rows sit apart from the list above and never mix into it.
GPT-5.6 Sol (OpenAI), as of 9 September 2026, with LMArena Agent 7.7% · AA τ²-Bench Telecom 85.1% across 2 public agent boards and ModelCap Index position #5.
It contains current, canonical language models with a measured public ModelCap Index position and admitted results on at least two public agent-family boards, ordered by the agent component of the ModelCap Index (ties broken by Index rank). A second tier, "Measured on one agent board", follows: Ranked models with a single admitted agent result, ordered by that board's normalised score. One board is a narrower claim than two, so these rows sit apart from the list above and never mix into it. Nothing is estimated for the list itself: the filter is a published fact and the order is a published figure, and every row links to the model page that shows its evidence.
29 models qualify (4 in the primary tier, 25 measured on one board). The population is the live ModelCap Index, so a model appears here the moment it holds a public position and meets the filter.
GPT-6 Astra holds ModelCap Index position #1, the highest of any model in the list, and sits at #2 here by agent boards.
Solar Pro 4 lists the lowest blended API price at $0.052 per 1M tokens ($0.03 input / $0.12 output).
MiMo-V2.5-Pro from Xiaomi is the highest-placed model with openly downloadable weights (#4 here, ModelCap Index #21).
It is rendered from the sealed dataset published 9 September 2026 and re-renders within a minute of each refresh; positions, prices and context figures are the ones shown on the live rankings at the same instant.