ModelCap

Current AI model comparison

Exact identities · public evidence

Internlm

Intern S2 Preview 397B

Position

#94

Index

49.4

Measured

global-corpus-prior over 195 held-out anchors (64%); 1 specialist-board observation at 1.2% support (36%); shrunk 5.9 up toward the measured corpus

Comparison side 1

StepFun

Step 3.7 Flash

Position

#97

Index

49.2

Estimated

3 reported rows against measured corpus ladders (17%); global-corpus-prior over 195 held-out anchors (18%); measured predecessor stepfun/step-3.5-flash less the succession penalty (65%); shrunk 0.8 up toward the measured corpus

Comparison side 2

Output / 1M

Unavailable

vs $1.15

Context

vs 262K

Weights

Open weights

vs Open weights