Skip to content
ModelCap

Head-to-head comparison

Kimi K2.7 Code vs Step3 VL 10B

Kimi K2.7 Code (Moonshot AI) and Step3 VL 10B (Stepfun Ai) compared on the ModelCap Index, API price, context window, provider availability, weight access and every public benchmark board they share. Figures are the same ones shown on the live rankings; nothing here is a hidden score.

Snapshot as of 9 September 2026

As of 9 September 2026, Kimi K2.7 Code holds the stronger ModelCap Index position (#47 vs #48); and Step3 VL 10B ships open weights.

Which should you choose?

Choose Kimi K2.7 Code if…

  • you want the stronger overall ModelCap Index position — #47 against #48 (59.5 vs 59.2 points).
  • you want provider choice — 14 listed API providers against 0.
  • AA τ²-Bench Telecom is your yardstick — 90.1% against 16.1%.
  • AA Humanity's Last Exam is your yardstick — 35.0% against 10.8%.
  • you want the more recently listed model — Kimi K2.7 Code was listed 12 June 2026, Step3 VL 10B 13 January 2026.

Choose Step3 VL 10B if…

  • you want to self-host: Step3 VL 10B ships open weights while Kimi K2.7 Code is restricted license.

Kimi K2.7 Code vs Step3 VL 10B: specs, pricing and context

Specification comparison of Kimi K2.7 Code and Step3 VL 10B
FieldKimi K2.7 CodeMoonshot AIStep3 VL 10BStepfun Ai
ModelCap Index position#47#48
Index score59.559.2
EvidenceMeasuredEstimated
Input price / 1M tokens$0.71Unlisted
Output price / 1M tokens$3.50Unlisted
Context window262K tokens— tokens
Max output tokens236K
API providers listed140
Weight accessRestricted licenseOpen weights
Input modalitiesText, ImageText, Image
First listed12 June 202613 January 2026
PublisherMoonshot AIStepfun Ai

Evidence: 2 public benchmark observations across 2 boards · global-corpus-prior over 182 held-out anchors (29%); launch card against 3 resolved peers on 4 rows (71%); exceeds every named peer on 3 of 4 rows; no cross-lab optimism probe available; shrunk 2.3 toward the measured corpus. Prices are the lowest listed API offer per million tokens observed on the OpenRouter catalogue.

Benchmark scores: Kimi K2.7 Code vs Step3 VL 10B

Public benchmark boards where Kimi K2.7 Code or Step3 VL 10B has a published result
BoardKimi K2.7 CodeStep3 VL 10BLeads
AA Intelligence IndexArtificial Analysis26.3intelligence-index:v4.3:reasoning-unspecifiedOnly one result
AA Terminal-Bench 2.1Artificial Analysis67.4%terminal-bench:2.1:terminus-2:e2b:pass-at-1:3-repeats:source-model="Kimi K2.7 Code":reasoning=trueOnly one result
AA τ²-Bench TelecomArtificial Analysis90.1%tau2:telecom:dual-control:pass-at-1:3-repeats:source-model="Kimi K2.7 Code":reasoning=true16.1%tau2:telecom:dual-control:pass-at-1:3-repeats:source-model="Step3 VL 10B":reasoning=trueKimi K2.7 Code
AA Humanity's Last ExamArtificial Analysis35.0%hle:may-2025:text-only-2158:no-tools:pass-at-1:source-model="Kimi K2.7 Code":reasoning=true10.8%hle:may-2025:text-only-2158:no-tools:pass-at-1:source-model="Step3 VL 10B":reasoning=trueKimi K2.7 Code
WildClawBench OpenClawWildClawBench46.9Only one result

Scores are the sources' own published figures for each model's best evaluated configuration; ModelCap never re-runs a benchmark.

Want a different pairing? Open the interactive comparison tool to swap either model for any current ranked language model.

Kimi K2.7 Code vs Step3 VL 10B: common questions

Is Kimi K2.7 Code better than Step3 VL 10B?

Kimi K2.7 Code ranks higher on the ModelCap Index as of 9 September 2026: #47 against #48. That is a capability ranking built from public benchmark evidence with published uncertainty; whether it is "better" for you also depends on price, context and where you can run it. Their published uncertainty intervals overlap, so the rank difference alone does not establish a reliable capability advantage for your workload.

Is Kimi K2.7 Code cheaper than Step3 VL 10B?

At least one of the two has no listed API price on ModelCap right now, so no price comparison is made.

Which has the bigger context window, Kimi K2.7 Code or Step3 VL 10B?

Both publish a 262K-token context window.

Which is better for coding, Kimi K2.7 Code or Step3 VL 10B?

The two models do not share a coding benchmark board on ModelCap yet, so no head-to-head coding score is published; the ModelCap Index position is the closest overall signal.

Are Kimi K2.7 Code and Step3 VL 10B open-weight models?

Kimi K2.7 Code: Restricted license. Step3 VL 10B: Open weights. Open weights mean the checkpoint can be downloaded and self-hosted under its licence; API-only models are available solely through hosted endpoints.

Where can I run Kimi K2.7 Code and Step3 VL 10B?

ModelCap currently lists 14 API providers for Kimi K2.7 Code and 0 for Step3 VL 10B, from the OpenRouter catalogue snapshot the site serves; each model page lists the providers and their prices.

How current is this Kimi K2.7 Code vs Step3 VL 10B comparison?

Every figure comes from the sealed ModelCap dataset published 9 September 2026; the page re-renders within a minute of each data refresh, and the ModelCap Index positions are the same ones shown on the live rankings.

Explore further