ModelCap
Current AI model comparison
Exact identities · public evidence
Allen Institute
Olmo 3 1125 32B
Position
#177
Index
18.5
Estimated
publisher-corpus-prior over 187 held-out anchors (9%); launch card against 3 resolved peers on 6 rows (91%); exceeds every named peer on 1 of 6 rows; no cross-lab optimism probe available; shrunk 0.2 up toward the measured corpus
Comparison side 1
Tiiuae
Falcon H1 3B Instruct
Position
#178
Index
18.2
Estimated
global-corpus-prior over 187 held-out anchors (20%); launch card against 2 resolved peers on 11 rows (80%); exceeds every named peer on 9 of 11 rows; optimism haircut 10.03 from cross-lab-probe; shrunk 8.4 up toward the measured corpus
Comparison side 2
Output / 1M
Unavailable
vs Unavailable
Context
—
vs —
Weights
Open weights
vs Restricted license