Skip to content
ModelCap

Head-to-head comparison

Qwen3 235B A22B Instruct 2507 vs Step3 VL 10B

Qwen3 235B A22B Instruct 2507 (Qwen) and Step3 VL 10B (Stepfun Ai) compared on the ModelCap Index, API price, context window, provider availability, weight access and every public benchmark board they share. Figures are the same ones shown on the live rankings; nothing here is a hidden score.

Snapshot as of 13 September 2026

As of 13 September 2026, Step3 VL 10B holds the stronger ModelCap Index position (#53 vs #56).

Which should you choose?

Choose Qwen3 235B A22B Instruct 2507 if…

  • you want provider choice — 10 listed API providers against 0.
  • AA τ²-Bench Telecom is your yardstick — 33.3% against 16.1%.
  • AA Humanity's Last Exam is your yardstick — 11.1% against 10.8%.

Choose Step3 VL 10B if…

  • you want the stronger overall ModelCap Index position — #53 against #56 (59.1 vs 57.6 points).
  • you want the more recently listed model — Step3 VL 10B was listed 13 January 2026, Qwen3 235B A22B Instruct 2507 21 July 2025.

Qwen3 235B A22B Instruct 2507 vs Step3 VL 10B: specs, pricing and context

Specification comparison of Qwen3 235B A22B Instruct 2507 and Step3 VL 10B
FieldQwen3 235B A22B Instruct 2507QwenStep3 VL 10BStepfun Ai
ModelCap Index position#56#53
Index score57.659.1
EvidenceMeasuredEstimated
Input price / 1M tokens$0.087Unlisted
Output price / 1M tokens$0.35Unlisted
Context window262K tokens— tokens
Max output tokens236K
API providers listed100
Weight accessOpen weightsOpen weights
Input modalitiesTextText, Image
First listed21 July 202513 January 2026
PublisherQwenStepfun Ai

Evidence: 2 public benchmark observations across 2 boards · global-corpus-prior over 186 held-out anchors (18%); launch card against 2 resolved peers on 4 rows (82%); exceeds every named peer on 3 of 4 rows; no cross-lab optimism probe available; shrunk 1.5 toward the measured corpus. Prices are the lowest listed API offer per million tokens observed on the OpenRouter catalogue.

Benchmark scores: Qwen3 235B A22B Instruct 2507 vs Step3 VL 10B

Public benchmark boards where Qwen3 235B A22B Instruct 2507 or Step3 VL 10B has a published result
BoardQwen3 235B A22B Instruct 2507Step3 VL 10BLeads
Arena codingArena (LMArena)14721468–1477Only one result
BFCL V4UC Berkeley (Gorilla)52.2PromptOnly one result
AA τ²-Bench TelecomArtificial Analysis33.3%tau2:telecom:dual-control:pass-at-1:3-repeats:source-model="Qwen3 235B A22B 2507 Instruct":reasoning=false16.1%tau2:telecom:dual-control:pass-at-1:3-repeats:source-model="Step3 VL 10B":reasoning=trueQwen3 235B A22B Instruct 2507
AA Humanity's Last ExamArtificial Analysis11.1%hle:may-2025:text-only-2158:no-tools:pass-at-1:source-model="Qwen3 235B A22B 2507 Instruct":reasoning=false10.8%hle:may-2025:text-only-2158:no-tools:pass-at-1:source-model="Step3 VL 10B":reasoning=trueQwen3 235B A22B Instruct 2507

Scores are the sources' own published figures for each model's best evaluated configuration; ModelCap never re-runs a benchmark.

Want a different pairing? Open the interactive comparison tool to swap either model for any current ranked language model.

Qwen3 235B A22B Instruct 2507 vs Step3 VL 10B: common questions

Is Qwen3 235B A22B Instruct 2507 better than Step3 VL 10B?

Step3 VL 10B ranks higher on the ModelCap Index as of 13 September 2026: #53 against #56. That is a capability ranking built from public benchmark evidence with published uncertainty; whether it is "better" for you also depends on price, context and where you can run it. Their published uncertainty intervals overlap, so the rank difference alone does not establish a reliable capability advantage for your workload.

Is Qwen3 235B A22B Instruct 2507 cheaper than Step3 VL 10B?

At least one of the two has no listed API price on ModelCap right now, so no price comparison is made.

Which has the bigger context window, Qwen3 235B A22B Instruct 2507 or Step3 VL 10B?

Both publish a 262K-token context window.

Which is better for coding, Qwen3 235B A22B Instruct 2507 or Step3 VL 10B?

The two models do not share a coding benchmark board on ModelCap yet, so no head-to-head coding score is published; the ModelCap Index position is the closest overall signal.

Are Qwen3 235B A22B Instruct 2507 and Step3 VL 10B open-weight models?

Qwen3 235B A22B Instruct 2507: Open weights. Step3 VL 10B: Open weights. Open weights mean the checkpoint can be downloaded and self-hosted under its licence; API-only models are available solely through hosted endpoints.

Where can I run Qwen3 235B A22B Instruct 2507 and Step3 VL 10B?

ModelCap currently lists 10 API providers for Qwen3 235B A22B Instruct 2507 and 0 for Step3 VL 10B, from the OpenRouter catalogue snapshot the site serves; each model page lists the providers and their prices.

How current is this Qwen3 235B A22B Instruct 2507 vs Step3 VL 10B comparison?

Every figure comes from the sealed ModelCap dataset published 13 September 2026; the page re-renders within a minute of each data refresh, and the ModelCap Index positions are the same ones shown on the live rankings.

Explore further