ModelCap
Current AI model comparison
Exact identities · public evidence
Allen Institute
Olmo 3 1125 32B
Position
#211
Index
18.6
Estimated
publisher-corpus-prior over 209 held-out anchors (27%); launch card against 4 resolved peers on 6 rows (73%); exceeds every named peer on 1 of 6 rows; no cross-lab optimism probe available; shrunk 0.1 toward the measured corpus
Comparison side 1
Arcee AI
Arcee Blitz
Position
#212
Index
18.5
Inherited
finetune of mistralai/mistral-small-24b-instruct-2501 · finetune (100%)
Comparison side 2
Output / 1M
Unavailable
vs Unavailable
Context
—
vs —
Weights
Open weights
vs Open weights