Skip to content
ModelCap

Head-to-head comparison

o4 Mini vs Qwen3 235B A22B Thinking 2507

o4 Mini (OpenAI) and Qwen3 235B A22B Thinking 2507 (Qwen) compared on the ModelCap Index, API price, context window, provider availability, weight access and every public benchmark board they share. Figures are the same ones shown on the live rankings; nothing here is a hidden score.

Snapshot as of 3 September 2026

As of 3 September 2026, Qwen3 235B A22B Thinking 2507 holds the stronger ModelCap Index position (#55 vs #62); Qwen3 235B A22B Thinking 2507 is 1.9× cheaper per output token ($2.30/1M vs $4.40/1M); o4 Mini offers the longer context window (200K vs 131K tokens); and Qwen3 235B A22B Thinking 2507 ships open weights.

Which should you choose?

Choose o4 Mini if…

  • you need the longer context window — 200K tokens against 131K.

Choose Qwen3 235B A22B Thinking 2507 if…

  • you want the stronger overall ModelCap Index position — #55 against #62 (46.7 vs 43.2 points).
  • API cost matters — $2.30/1M output tokens against $4.40/1M, about 1.9× cheaper.
  • your workload is prompt-heavy — input tokens cost $0.23/1M against $1.10/1M.
  • you want to self-host: Qwen3 235B A22B Thinking 2507 ships open weights while o4 Mini is api only.
  • you want provider choice — 3 listed API providers against 1.
  • Arena coding is your yardstick — 1443 against 1433.
  • you want the more recently listed model — Qwen3 235B A22B Thinking 2507 was listed 25 July 2025, o4 Mini 16 April 2025.

o4 Mini vs Qwen3 235B A22B Thinking 2507: specs, pricing and context

Specification comparison of o4 Mini and Qwen3 235B A22B Thinking 2507
Fieldo4 MiniOpenAIQwen3 235B A22B Thinking 2507Qwen
ModelCap Index position#62#55
Index score43.246.7
EvidenceMeasuredMeasured
Input price / 1M tokens$1.10$0.23
Output price / 1M tokens$4.40$2.30
Context window200K tokens131K tokens
Max output tokens100K118K
API providers listed13
Weight accessAPI onlyOpen weights
Input modalitiesImage, Text, FileText
First listed16 April 202525 July 2025
PublisherOpenAIQwen

Evidence: 4 public benchmark observations across 4 boards · 2 public benchmark observations across 2 boards. Prices are the lowest listed API offer per million tokens observed on the OpenRouter catalogue.

Benchmark scores: o4 Mini vs Qwen3 235B A22B Thinking 2507

Public benchmark boards where o4 Mini or Qwen3 235B A22B Thinking 2507 has a published result
Boardo4 MiniQwen3 235B A22B Thinking 2507Leads
Arena codingArena (LMArena)14331426–144014431428–1457Qwen3 235B A22B Thinking 2507
ARC-AGI-2ARC Prize Foundation6.1%HighOnly one result
SWE-benchSWE-bench (Princeton / SWE-agent team)45.0%Only one result
BFCL V4UC Berkeley (Gorilla)53.2Function callingOnly one result
AA Intelligence IndexArtificial Analysis26.1intelligence-indexOnly one result
AA τ²-Bench TelecomArtificial Analysis55.6%tau2-bench-telecomOnly one result
AA Humanity's Last ExamArtificial Analysis16.5%humanitys-last-examOnly one result

Scores are the sources' own published figures for each model's best evaluated configuration; ModelCap never re-runs a benchmark.

Want a different pairing? Open the interactive comparison tool to swap either model for any current ranked language model.

o4 Mini vs Qwen3 235B A22B Thinking 2507: common questions

Is o4 Mini better than Qwen3 235B A22B Thinking 2507?

Qwen3 235B A22B Thinking 2507 ranks higher on the ModelCap Index as of 3 September 2026: #55 against #62. That is a capability ranking built from public benchmark evidence with published uncertainty; whether it is "better" for you also depends on price, context and where you can run it.

Is o4 Mini cheaper than Qwen3 235B A22B Thinking 2507?

Qwen3 235B A22B Thinking 2507 is cheaper on output tokens: $2.30/1M against $4.40/1M. Input tokens are $1.10/1M for o4 Mini and $0.23/1M for Qwen3 235B A22B Thinking 2507. Prices are the lowest listed API offer ModelCap observed, in USD per million tokens.

Which has the bigger context window, o4 Mini or Qwen3 235B A22B Thinking 2507?

o4 Mini has the larger context window: 200K tokens against 131K.

Which is better for coding, o4 Mini or Qwen3 235B A22B Thinking 2507?

Arena coding: o4 Mini 1433, Qwen3 235B A22B Thinking 2507 1443 — Qwen3 235B A22B Thinking 2507 leads.

Are o4 Mini and Qwen3 235B A22B Thinking 2507 open-weight models?

o4 Mini: API only. Qwen3 235B A22B Thinking 2507: Open weights. Open weights mean the checkpoint can be downloaded and self-hosted under its licence; API-only models are available solely through hosted endpoints.

Where can I run o4 Mini and Qwen3 235B A22B Thinking 2507?

ModelCap currently lists 1 API provider for o4 Mini and 3 for Qwen3 235B A22B Thinking 2507, from the OpenRouter catalogue snapshot the site serves; each model page lists the providers and their prices.

How current is this o4 Mini vs Qwen3 235B A22B Thinking 2507 comparison?

Every figure comes from the sealed ModelCap dataset published 3 September 2026; the page re-renders within a minute of each data refresh, and the ModelCap Index positions are the same ones shown on the live rankings.

Explore further