ModelCap

Current AI model comparison

Exact identities · public evidence

Google

gemma 4 E4B it

Position

#104

Index

41.3

Estimated

publisher-corpus-prior over 185 held-out anchors (7%); launch card against 3 resolved peers on 5 rows (94%); no cross-lab optimism probe available; shrunk 1.7 up toward the measured corpus

Comparison side 1

Z.ai

GLM 4.5 Air

Position

#101

Index

42.2

Measured

2 public benchmark observations across 2 boards

Comparison side 2

Output / 1M

Unavailable

vs $0.85

Context

vs 131K

Weights

Open weights

vs Open weights