ModelCap
Current AI model comparison
Exact identities · public evidence
Nanbeige
Nanbeige4.1 3B
Position
#94
Index
43.5
Estimated
global-corpus-prior over 183 held-out anchors (19%); launch card against 2 resolved peers on 6 rows (81%); exceeds every named peer on 5 of 6 rows; optimism haircut 3 from cross-lab-probe; shrunk 2.3 up toward the measured corpus
Comparison side 1
OpenAI
o4 Mini
Position
#91
Index
45.4
Measured
3 public benchmark observations across 3 boards
Comparison side 2
Output / 1M
Unavailable
vs $4.40
Context
—
vs 200K
Weights
Open weights
vs API only