ModelCap
Current AI model comparison
Exact identities · public evidence
Allen Institute
Llama 3.1 Tulu 3 70B
Position
#200
Index
22.0
Measured
1 public benchmark observation across 1 board
Comparison side 1
LiquidAI
LFM2 8B A1B
Position
#199
Index
22.2
Estimated
global-corpus-prior over 210 held-out anchors (18%); launch card against 3 resolved peers on 10 rows (82%); exceeds every named peer on 8 of 10 rows; optimism haircut 9.5 from cross-lab-probe; shrunk 7.6 up toward the measured corpus
Comparison side 2
Output / 1M
Unavailable
vs Unavailable
Context
—
vs —
Weights
Restricted license
vs Restricted license