Skip to content
ModelCap

Model decision surface

Compare AI models

Start with Claude Sonnet 5 and GPT-6 Astra, or choose any two current ranked language models. Compare capability evidence, price, context, provider availability, and weight access without pretending one field decides every use case.

Current public data

Claude Sonnet 5 vs GPT-6 Astra

Live dataset updated 9/9/2026, 8:46:14 PM UTC

Open 1200×630 evidence receipt
Factual comparison of Claude Sonnet 5 and GPT-6 Astra
Field
ModelCap position#18#1
Index score75.693.7
EvidenceMeasured4 public benchmark observations across 4 boardsMeasured4 public benchmark observations across 4 boards
Input / 1M$2.00$10.00
Output / 1M$10.00$50.00
Pricing statusfreshfresh
Context1M1M
Providers52
Weight accessAPI onlyAPI only

Decision facts

  • Claude Sonnet 5 is #18; GPT-6 Astra is #1 on the same current language board.
  • Index scores are 75.6 for Claude Sonnet 5 and 93.7 for GPT-6 Astra. Their published uncertainty intervals overlap, so the rank difference alone does not establish a reliable capability advantage for your workload.
  • Both positions use Measured evidence.
  • Listed output price per 1M tokens is $10.00 for Claude Sonnet 5 and $50.00 for GPT-6 Astra. For 1,000 requests with 2,000 input and 500 output tokens each (2M input + 0.5M output), the listed-rate estimate is $9.00 for Claude Sonnet 5 versus $45.00 for GPT-6 Astra. Claude Sonnet 5 costs 80.0% less in this scenario. This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the selected endpoint before budgeting.
  • Published context is 1,000,000 tokens for Claude Sonnet 5 and 1,050,000 for GPT-6 Astra.
  • ModelCap currently lists 5 providers for Claude Sonnet 5 and 2 for GPT-6 Astra.

These are separate published fields, not a synthetic winner. ModelCap does not collapse price, access, context, and capability evidence into a hidden recommendation score.

Each comparison page is a permanent, shareable URL with the same live figures as this tool: ModelCap Index position, API pricing, context window, provider count, weight access and every shared benchmark board.