Skip to content
ModelCap

Model decision surface

Compare AI models

Start with GPT-4o (2024-08-06) and gpt-oss-120b, or choose any two current ranked language models. Compare capability evidence, price, context, provider availability, and weight access without pretending one field decides every use case.

Current public data

GPT-4o (2024-08-06) vs gpt-oss-120b

Live dataset updated 9/9/2026, 9:50:24 PM UTC

Open 1200×630 evidence receipt
Factual comparison of GPT-4o (2024-08-06) and gpt-oss-120b
Field
ModelCap position#128#110
Index score30.836.8
EvidenceMeasured2 public benchmark observations across 2 boardsMeasured3 public benchmark observations across 3 boards
Input / 1M$2.50$0.037
Output / 1M$10.00$0.17
Pricing statusfreshfresh
Context128K131K
Providers218
Weight accessAPI onlyOpen weights

Decision facts

  • GPT-4o (2024-08-06) is #128; gpt-oss-120b is #110 on the same current language board.
  • Index scores are 30.8 for GPT-4o (2024-08-06) and 36.8 for gpt-oss-120b. Their published uncertainty intervals overlap, so the rank difference alone does not establish a reliable capability advantage for your workload.
  • Both positions use Measured evidence.
  • Listed output price per 1M tokens is $10.00 for GPT-4o (2024-08-06) and $0.17 for gpt-oss-120b. For 1,000 requests with 2,000 input and 500 output tokens each (2M input + 0.5M output), the listed-rate estimate is $10.00 for GPT-4o (2024-08-06) versus $0.16 for gpt-oss-120b. gpt-oss-120b costs 98.4% less in this scenario. This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the selected endpoint before budgeting.
  • Published context is 128,000 tokens for GPT-4o (2024-08-06) and 131,072 for gpt-oss-120b.
  • Weight access differs: GPT-4o (2024-08-06) is none; gpt-oss-120b is open.
  • ModelCap currently lists 2 providers for GPT-4o (2024-08-06) and 18 for gpt-oss-120b.

These are separate published fields, not a synthetic winner. ModelCap does not collapse price, access, context, and capability evidence into a hidden recommendation score.

Each comparison page is a permanent, shareable URL with the same live figures as this tool: ModelCap Index position, API pricing, context window, provider count, weight access and every shared benchmark board.