ModelCap
Current AI model comparison
Exact identities · public evidence
Allen Institute
Olmo Hybrid 7B
Position
#193
Index
12.6
Estimated
publisher-corpus-prior over 194 held-out anchors (23%); launch card against 2 resolved peers on 6 rows (77%); exceeds every named peer on 6 of 6 rows; optimism haircut 0 from cross-lab-probe; shrunk 1.4 up toward the measured corpus
Comparison side 1
Microsoft
Phi 4
Position
#194
Index
12.4
Measured
2 public benchmark observations across 2 boards
Comparison side 2
Output / 1M
Unavailable
vs $0.14
Context
—
vs 16K
Weights
Open weights
vs Open weights