ModelCap
Current AI model comparison
Exact identities · public evidence
Anthropic
Claude Opus 5.5
Position
#1
Index
98.6
Measured
2 public benchmark observations across 2 boards
Comparison side 1
ByteDance Seed
Seed 2.1 Turbo
Position
#23
Index
79.7
Estimated
global-corpus-prior over 209 held-out anchors (11%); launch card against 2 resolved peers on 3 rows (89%); exceeds every named peer on 1 of 3 rows; no cross-lab optimism probe available; shrunk 2.8 toward the measured corpus
Comparison side 2
Output / 1M
$20.00
vs $2.50
Context
1M
vs 262K
Weights
API only
vs API only