ModelCap

Current AI model comparison

Exact identities · public evidence

DeepSeek

DeepSeek V4 Flash 0731

Position

#25

Index

70.7

Measured

2 public benchmark observations across 2 boards

Comparison side 1

Meituan

LongCat 2.0

Position

#11

Index

79.0

Estimated

global-corpus-prior over 183 held-out anchors (7%); launch card against 4 resolved peers on 6 rows (93%); no cross-lab optimism probe available; shrunk 1.9 toward the measured corpus

Comparison side 2

Output / 1M

$0.18

vs $1.20

Context

1.3M

vs 1M

Weights

Open weights

vs Open weights