ModelCap

Current AI model comparison

Exact identities · public evidence

Arcee AI

Trinity Large Thinking

Position

#111

Index

45.9

Measured

3 public benchmark observations across 3 boards

Comparison side 1

Nanbeige

Nanbeige4.1 3B

Position

#108

Index

46.9

Estimated

global-corpus-prior over 203 held-out anchors (19%); launch card against 2 resolved peers on 6 rows (81%); exceeds every named peer on 5 of 6 rows; optimism haircut 3 from cross-lab-probe; shrunk 1.9 up toward the measured corpus

Comparison side 2

Output / 1M

$0.80

vs Unavailable

Context

262K

vs

Weights

Restricted license

vs Open weights