ModelCap
Current AI model comparison
Exact identities · public evidence
DeepSeek
DeepSeek V4.1 Flash
Position
#40
Index
65.2
Estimated
publisher-corpus-prior over 185 held-out anchors (22%); opens from measured predecessor deepseek/deepseek-v4-flash-0731 (78%); shrunk 2 toward the measured corpus
Comparison side 1
Poolside
Laguna M.1
Position
#42
Index
64.5
Estimated
global-corpus-prior over 185 held-out anchors (7%); launch card against 3 resolved peers on 3 rows (94%); optimism haircut 0 from cross-lab-probe; shrunk 0.8 toward the measured corpus
Comparison side 2
Output / 1M
$0.60
vs Unavailable
Context
1M
vs —
Weights
Open weights
vs Open weights