ModelCap

Current AI model comparison

Exact identities · public evidence

Allen Institute

Olmo 3.1 32B Think

Position

#180

Index

17.7

Measured

2 public benchmark observations across 2 boards

Comparison side 1

Tiiuae

Falcon H1 3B Instruct

Position

#178

Index

18.2

Estimated

global-corpus-prior over 187 held-out anchors (20%); launch card against 2 resolved peers on 11 rows (80%); exceeds every named peer on 9 of 11 rows; optimism haircut 10.03 from cross-lab-probe; shrunk 8.4 up toward the measured corpus

Comparison side 2

Output / 1M

Unavailable

vs Unavailable

Context

vs

Weights

Open weights

vs Restricted license