Skip to content
ModelCap

Head-to-head comparison

GPT-4 Turbo vs Inkling Small

GPT-4 Turbo (OpenAI) and Inkling Small (Thinking Machines) compared on the ModelCap Index, API price, context window, provider availability, weight access and every public benchmark board they share. Figures are the same ones shown on the live rankings; nothing here is a hidden score.

Snapshot as of 3 September 2026

As of 3 September 2026, Inkling Small holds the stronger ModelCap Index position (#42 vs #174); Inkling Small is 25× cheaper per output token ($1.20/1M vs $30.00/1M); Inkling Small offers the longer context window (1M vs 128K tokens); and Inkling Small ships open weights.

Which should you choose?

Choose GPT-4 Turbo if…

  • ModelCap publishes no field on which GPT-4 Turbo leads Inkling Small for this pair.

Choose Inkling Small if…

  • you want the stronger overall ModelCap Index position — #42 against #174 (53.9 vs 19.9 points).
  • API cost matters — $1.20/1M output tokens against $30.00/1M, about 25× cheaper.
  • your workload is prompt-heavy — input tokens cost $0.45/1M against $10.00/1M.
  • you need the longer context window — 1M tokens against 128K.
  • you want to self-host: Inkling Small ships open weights while GPT-4 Turbo is api only.
  • you want provider choice — 3 listed API providers against 1.
  • Arena coding is your yardstick — 1476 against 1347.
  • AA Intelligence Index is your yardstick — 41.2 against 7.7.
  • AA Humanity's Last Exam is your yardstick — 33.3% against 3.1%.
  • you want the more recently listed model — Inkling Small was listed 30 July 2026, GPT-4 Turbo 9 April 2024.

GPT-4 Turbo vs Inkling Small: specs, pricing and context

Specification comparison of GPT-4 Turbo and Inkling Small
FieldGPT-4 TurboOpenAIInkling SmallThinking Machines
ModelCap Index position#174#42
Index score19.953.9
EvidenceMeasuredMeasured
Input price / 1M tokens$10.00$0.45
Output price / 1M tokens$30.00$1.20
Context window128K tokens1M tokens
Max output tokens4K262K
API providers listed13
Weight accessAPI onlyOpen weights
Input modalitiesText, ImageText, Image, Audio
First listed9 April 202430 July 2026
PublisherOpenAIThinking Machines

Evidence: 3 public benchmark observations across 3 boards · 5 public benchmark observations across 5 boards. Prices are the lowest listed API offer per million tokens observed on the OpenRouter catalogue.

Benchmark scores: GPT-4 Turbo vs Inkling Small

Public benchmark boards where GPT-4 Turbo or Inkling Small has a published result
BoardGPT-4 TurboInkling SmallLeads
Arena codingArena (LMArena)13471340–135414761467–1486Inkling Small
LMArena AgentLMArena-5.3%-7.1%–-3.5%Only one result
ARC-AGI-2ARC Prize Foundation40.1%X-HighOnly one result
AA Intelligence IndexArtificial Analysis7.7intelligence-index41.2intelligence-indexInkling Small
AA Terminal-Bench 2.1Artificial Analysis55.1%terminal-bench-2.1Only one result
AA Humanity's Last ExamArtificial Analysis3.1%humanitys-last-exam33.3%humanitys-last-examInkling Small

Scores are the sources' own published figures for each model's best evaluated configuration; ModelCap never re-runs a benchmark.

Want a different pairing? Open the interactive comparison tool to swap either model for any current ranked language model.

GPT-4 Turbo vs Inkling Small: common questions

Is GPT-4 Turbo better than Inkling Small?

Inkling Small ranks higher on the ModelCap Index as of 3 September 2026: #42 against #174. That is a capability ranking built from public benchmark evidence with published uncertainty; whether it is "better" for you also depends on price, context and where you can run it.

Is GPT-4 Turbo cheaper than Inkling Small?

Inkling Small is cheaper on output tokens: $1.20/1M against $30.00/1M. Input tokens are $10.00/1M for GPT-4 Turbo and $0.45/1M for Inkling Small. Prices are the lowest listed API offer ModelCap observed, in USD per million tokens.

Which has the bigger context window, GPT-4 Turbo or Inkling Small?

Inkling Small has the larger context window: 1M tokens against 128K.

Which is better for coding, GPT-4 Turbo or Inkling Small?

Arena coding: GPT-4 Turbo 1347, Inkling Small 1476 — Inkling Small leads.

Are GPT-4 Turbo and Inkling Small open-weight models?

GPT-4 Turbo: API only. Inkling Small: Open weights. Open weights mean the checkpoint can be downloaded and self-hosted under its licence; API-only models are available solely through hosted endpoints.

Where can I run GPT-4 Turbo and Inkling Small?

ModelCap currently lists 1 API provider for GPT-4 Turbo and 3 for Inkling Small, from the OpenRouter catalogue snapshot the site serves; each model page lists the providers and their prices.

How current is this GPT-4 Turbo vs Inkling Small comparison?

Every figure comes from the sealed ModelCap dataset published 3 September 2026; the page re-renders within a minute of each data refresh, and the ModelCap Index positions are the same ones shown on the live rankings.

Explore further