ModelCap
Current AI model comparison
Exact identities · public evidence
Stepfun Ai
Step3 VL 10B
Position
#53
Index
59.5
Estimated
global-corpus-prior over 194 held-out anchors (18%); launch card against 2 resolved peers on 4 rows (82%); exceeds every named peer on 3 of 4 rows; no cross-lab optimism probe available; shrunk 1.6 toward the measured corpus
Comparison side 1
XiaomiMiMo
MiMo V2 Flash
Position
#48
Index
61.4
Estimated
global-corpus-prior over 194 held-out anchors (14%); launch card against 4 resolved peers on 10 rows (87%); exceeds every named peer on 1 of 10 rows; optimism haircut 0 from cross-lab-probe; shrunk 1.6 toward the measured corpus
Comparison side 2
Output / 1M
Unavailable
vs Unavailable
Context
—
vs —
Weights
Open weights
vs Open weights