1Current-model rank | | 97.1Measured 2 public benchmark observations across 2 boards; Index basis Measured; 22% support; score interval 72.4–100.0. | 1M 128K max output Exact context: 1,000,000 tokens; exact maximum output: 128,000 tokens | | $20.00 |
2Current-model rank | | 88.4Measured 1 public benchmark observation across 1 board; Index basis Measured; 20% support; score interval 67.3–100.0. | 1M 128K max output Exact context: 1,050,000 tokens; exact maximum output: 128,000 tokens | | $10.00 |
3Current-model rank | | 87.8Measured 5 public benchmark observations across 5 boards; Index basis Measured; 78% support; score interval 80.6–95.0. | 1M 128K max output Exact context: 1,000,000 tokens; exact maximum output: 128,000 tokens | 1,498 LMArena Highest supported configuration: Max; source rank 5 | $50.00 |
4Current-model rank | | 87.4Measured 1 public benchmark observation across 1 board; Index basis Measured; 20% support; score interval 66.3–100.0. | 500K 450K max output Exact context: 500,000 tokens; exact maximum output: 450,000 tokens | | $4.80 |
5Current-model rank | | 87.3Measured 1 public benchmark observation across 1 board; Index basis Measured; 20% support; score interval 66.2–100.0. | 1M 131K max output Exact context: 1,048,576 tokens; exact maximum output: 131,072 tokens | | $0.87 |
6Current-model rank | | 86.5Measured 1 public benchmark observation across 1 board; Index basis Measured; 20% support; score interval 65.4–100.0. | 1M 131K max output Exact context: 1,000,000 tokens; exact maximum output: 131,072 tokens | | $6.00 |
7Current-model rank | | 86.2Measured 6 public benchmark observations across 6 boards; Index basis Measured; 68% support; score interval 77.6–94.8. | 1M 128K max output Exact context: 1,050,000 tokens; exact maximum output: 128,000 tokens | 1,480 LMArena Highest supported configuration: Max; source rank 24 | $50.00 |
8Current-model rank | | 85.6Measured 4 public benchmark observations across 4 boards; Index basis Measured; 75% support; score interval 78.0–93.2. | 1M 944K max output Exact context: 1,048,576 tokens; exact maximum output: 943,718 tokens | 1,493 LMArena Highest supported configuration: Max; source rank 8 | $4.25 |
9Current-model rank | | 84.1Measured 4 public benchmark observations across 4 boards; Index basis Measured; 75% support; score interval 76.5–91.7. | 1M 66K max output Exact context: 1,048,576 tokens; exact maximum output: 65,536 tokens | 1,493 LMArena Highest supported configuration: High; source rank 9 | $3.75 |
10Current-model rank | | 83.9Modeled · peer benchmarks publisher-corpus-prior over 223 held-out anchors (8%); launch card against 5 resolved peers on 7 rows (92%); exceeds every named peer on 1 of 7 rows; no cross-lab optimism probe available; shrunk 0.6 toward the measured corpus; Index basis Estimated; 33% support; score interval 75.4–92.3. | 1M 131K max output Exact context: 1,048,576 tokens; exact maximum output: 131,072 tokens | | $0.28 |
11Current-model rank | | 83.8Measured 4 public benchmark observations across 4 boards; Index basis Measured; 83% support; score interval 77.1–90.5. | 1.3M 131K max output Exact context: 1,310,720 tokens; exact maximum output: 131,072 tokens | 1,483 LMArena Highest supported configuration: Max; source rank 19 | |
12Current-model rank | | 83.5Measured 6 public benchmark observations across 6 boards; Index basis Measured; 89% support; score interval 77.4–89.6. | 1M 944K max output Exact context: 1,048,576 tokens; exact maximum output: 943,718 tokens | 1,485 LMArena Highest supported configuration: Max; source rank 17 | $15.00 |
13Current-model rank | | 82.2Measured 7 public benchmark observations across 7 boards; Index basis Measured; 90% support; score interval 76.2–88.2. | 1M 128K max output Exact context: 1,050,000 tokens; exact maximum output: 128,000 tokens | 1,482 LMArena Highest supported configuration: High; source rank 20 | $30.00 |
14Current-model rank | | 81.6Measured 4 public benchmark observations across 4 boards; Index basis Measured; 82% support; score interval 74.9–88.3. | 1.3M 944K max output Exact context: 1,310,720 tokens; exact maximum output: 943,718 tokens | 1,475 LMArena Highest supported configuration: Default; source rank 29 | $0.50 |
15Current-model rank | | 81.5Measured 1 public benchmark observation across 1 board; Index basis Measured; 20% support; score interval 60.4–100.0. | 1M 131K max output Exact context: 1,000,000 tokens; exact maximum output: 131,072 tokens | | $0.47 |
16Current-model rank | | 81.4Modeled · peer benchmarks global-corpus-prior over 223 held-out anchors (6%); launch card against 4 resolved peers on 6 rows (94%); no cross-lab optimism probe available; shrunk 1.6 toward the measured corpus; Index basis Estimated; 32% support; score interval 73.9–88.8. | 1M 262K max output Exact context: 1,048,756 tokens; exact maximum output: 262,144 tokens | | $1.20 |
17Current-model rank | | 81.2Measured 4 public benchmark observations across 4 boards; Index basis Measured; 86% support; score interval 74.9–87.5. | 1M 66K max output Exact context: 1,048,576 tokens; exact maximum output: 65,536 tokens | 1,487 LMArena Highest supported configuration: Default; source rank 15 | $12.00 |
18Current-model rank | | 81.0Measured 2 public benchmark observations across 2 boards; Index basis Measured; 62% support; score interval 76.0–86.0. | 2M 1.8M max output Exact context: 2,000,000 tokens; exact maximum output: 1,800,000 tokens | 1,470 LMArena Highest supported configuration: Beta · 0309; source rank 40 | $2.50 |
19Current-model rank | | 80.2Measured 2 public benchmark observations across 2 boards; Index basis Measured; 23% support; score interval 59.3–100.0. | 1M 944K max output Exact context: 1,048,576 tokens; exact maximum output: 943,718 tokens | | |
20Current-model rank | | 80.1Modeled · peer benchmarks global-corpus-prior over 223 held-out anchors (6%); launch card against 6 resolved peers on 5 rows (94%); optimism haircut 0 from cross-lab-probe; shrunk 1.6 toward the measured corpus; Index basis Estimated; 27% support; score interval 72.7–87.6. | — — max output Exact context: 0 tokens | | Self-hosted: weights only · no listed API price |
| | 79.6Modeled · peer benchmarks global-corpus-prior over 223 held-out anchors (11%); launch card against 2 resolved peers on 3 rows (90%); exceeds every named peer on 1 of 3 rows; no cross-lab optimism probe available; shrunk 2.6 toward the measured corpus; Index basis Estimated; 23% support; score interval 69.9–89.2. | 262K 236K max output Exact context: 262,144 tokens; exact maximum output: 235,929 tokens | | $2.50 |
| | 79.5Measured 6 public benchmark observations across 6 boards; Index basis Measured; 89% support; score interval 73.4–85.6. | 1M 128K max output Exact context: 1,050,000 tokens; exact maximum output: 128,000 tokens | 1,466 LMArena Highest supported configuration: X-High; source rank 45 | $12.00 |
| | 79.5Modeled · peer benchmarks global-corpus-prior over 223 held-out anchors (8%); launch card against 8 resolved peers on 8 rows (92%); exceeds every named peer on 1 of 8 rows; optimism haircut 0 from cross-lab-probe; shrunk 1.9 toward the measured corpus; Index basis Estimated; 38% support; score interval 71.2–87.8. | 262K 236K max output Exact context: 262,144 tokens; exact maximum output: 235,929 tokens | | $0.25 |
24Current-model rank | | 79.2Modeled · peer benchmarks global-corpus-prior over 223 held-out anchors (10%); launch card against 2 resolved peers on 11 rows (90%); exceeds every named peer on 3 of 11 rows; optimism haircut 0 from cross-lab-probe; shrunk 2.4 toward the measured corpus; Index basis Estimated; 40% support; score interval 69.9–88.5. | — — max output Exact context: 0 tokens | | Self-hosted: weights only · no listed API price |
25Current-model rank | | 78.5Measured 3 public benchmark observations across 3 boards; Index basis Measured; 60% support; score interval 73.3–83.7. | 1M 384K max output Exact context: 1,048,576 tokens; exact maximum output: 384,000 tokens | 1,463 LMArena Highest supported configuration: High; source rank 50 | |