Skip to content
ModelCap

Model decision surface

Compare AI models

Start with Llama 4 Scout and GPT-5.1-Codex-Max, or choose any two current ranked language models. Compare capability evidence, price, context, provider availability, and weight access without pretending one field decides every use case.

Current public data

Llama 4 Scout vs GPT-5.1-Codex-Max

Live dataset updated 8/18/2026, 3:36:20 AM UTC

Open 1200×630 evidence receipt
Factual comparison of Llama 4 Scout and GPT-5.1-Codex-Max
Field
ModelCap position#108#113
Index score22.422.1
EvidenceMeasured4 public benchmark observations across 4 boardsEstimatedpublisher corpus prior · leave-one-anchor-out calibrated
Input / 1M$0.10$1.25
Output / 1M$0.30$10.00
Pricing statusfreshfresh
Context1.3M400K
Providers41
Weight accessGated accessAPI only

Decision facts

  • Llama 4 Scout is #108; GPT-5.1-Codex-Max is #113 on the same current language board.
  • Index scores are 22.4 for Llama 4 Scout and 22.1 for GPT-5.1-Codex-Max.
  • Evidence differs: Llama 4 Scout is Measured; GPT-5.1-Codex-Max is Estimated.
  • Listed output price per 1M tokens is $0.30 for Llama 4 Scout and $10.00 for GPT-5.1-Codex-Max.
  • Published context is 1,310,720 tokens for Llama 4 Scout and 400,000 for GPT-5.1-Codex-Max.
  • Weight access differs: Llama 4 Scout is gated; GPT-5.1-Codex-Max is none.
  • ModelCap currently lists 4 providers for Llama 4 Scout and 1 for GPT-5.1-Codex-Max.

These are separate published fields, not a synthetic winner. ModelCap does not collapse price, access, context, and capability evidence into a hidden recommendation score.

Each comparison page is a permanent, shareable URL with the same live figures as this tool: ModelCap Index position, API pricing, context window, provider count, weight access and every shared benchmark board.