Skip to content
ModelCap

Inference providers

LLM API providers compared

79 inference providers serving at least one tracked model through OpenRouter. Uptime is measured across each provider's endpoints over the last 30 minutes; each row averages that provider's deployments, and the summary gives every provider equal weight, so neither is a per-model guarantee. Entry price is the cheapest positive language-token output price the provider lists. Every provider name opens a page listing each model it serves with that provider's own price, context window, quantization and uptime.

Snapshot as of 23 September 2026

Providers
79
Mean provider uptime
98.35%
Largest catalogue
74 models
Inference providers by number of tracked models served, with entry price and measured uptime
#ProviderModels servedOn the IndexHighest-ranked modelEntry price ($/1M out)Uptime
1DeepInfra7448MiMo-V2.6-Pro$0.0395.16%
2Novita7039GLM 5.3$0.0598.29%
3Alibaba5524Qwen3.8 Max (0902)$0.1399.72%
4OpenAI4813GPT-6 Sol$0.20100.0%
5Azure4713Claude Opus 5.5$0.4099.69%
6Google4419Claude Opus 5.5$0.2596.16%
7SiliconFlow3920GLM 5.3$0.1595.09%
8Amazon Bedrock3615Claude Opus 5.5$0.1499.27%
9Venice3419GLM 5.3$0.1595.79%
10Parasail3323GLM 5.3$0.0399.75%
11AtlasCloud2912GLM 5.3$0.2898.85%
12GMICloud2310GLM 5.3$0.18296.81%
13StreamLake2313GLM 5.3 Flash$0.15899.11%
14Google AI Studio223Gemini 3.8 Flash$0.2099.84%
15Phala2012GLM 5.3$0.2098.60%
16CoreWeave1915GLM 5.3 Flash$0.1399.49%
17Cloudflare1811GLM 5.3$0.11290.26%
18DigitalOcean169GLM 5.3$0.19699.02%
19Mistral155GLM 5.3$0.1099.66%
20BaseTen149GLM 5.3$0.2699.22%
21Z.AI146GLM 5.3$0.5099.64%
22Together1311GLM 5.3$0.2598.30%
23Anthropic124Claude Opus 5.5$5.0099.93%
24Fireworks117GLM 5.3$0.5099.63%
25Baidu103GLM 5.3$0.16899.23%
26Claude Platform on AWS102Claude Opus 5.5$10.0099.94%
27DekaLLM87DeepSeek V4.1 Flash$0.0398.66%
28Minimax82MiniMax M3$1.0297.63%
29NextBit83GLM 5.3 Flash$0.2299.39%
30Wafer86GLM 5.3$0.3599.74%
31Cohere75Command A+$0.1599.33%
32Darkbloom76Qwen3.8 27B$0.0999.57%
33Friendli74GLM 5.3$0.4099.86%
34Morph74GLM 5.3$0.30897.84%
35Nebius75Qwen3 235B A22B Instruct 2507$0.2493.01%
36Sail Research75GLM 5.3$0.47599.66%
37SambaNova75Gemma 4 31B$0.9097.60%
38xAI72Grok 4.7$2.0098.62%
39AionLabs60$1.40100.0%
40AkashML66GLM 5.3$0.1099.86%
41Chutes63Kimi K3$0.3794.61%
42Crusoe65GLM 5.3$0.2097.63%
43Groq65gpt-oss-120b$0.0899.69%
44Inceptron63GLM 5.3$0.5099.71%
45Reka63GLM 5.3$0.1098.97%
46Relace63Kimi K3$0.3699.91%
47Seed61Seed 2.1 Turbo$0.30100.0%
48Makora54GLM 5.3$0.19598.99%
49Mancer 252Qwen3.8 27B$0.3098.25%
50Mara53MiniMax M3$0.7581.22%
51Meta51Muse Spark 1.3$0.20100.0%
52Perplexity50$1.00100.0%
53Tencent51Hy4 preview$0.177100.0%
54Xiaomi53MiMo-V2.6-Pro$0.28100.0%
55Modal44GLM 5.3$1.2098.83%
56OpenInference42GLM 5.3 Flash$0.5096.54%
57Sakana AI40$4.00100.0%
58Decart31GLM 5.3$2.4799.76%
59Inception31Mercury 2.5$0.15100.0%
60InferenceNet33GLM 5.3$0.2898.80%
61ModelRun33Gemma 4 31B$1.0099.90%
62Moonshot AI32Kimi K3$4.0099.72%
63Upstage31Solar Pro 4$0.2099.92%
64DeepSeek22DeepSeek V4.1 Flash$0.60100.0%
65Io Net22GLM 5.3$0.47599.82%
66Ionstream22DeepSeek V4 Pro 0813$2.3099.86%
67Nex AGI22Nex-N2.5-Pro$0.10100.0%
68Poolside22Laguna S 2.1$0.12100.0%
69Arcee AI11Trinity Large Thinking$0.80100.0%
70Cerebras11gpt-oss-120b$0.75100.0%
71Krea11DeepSeek V4.1 Flash$0.9099.84%
72Liquid10100.0%
73Near AI11GLM 5.3 Flash$0.5099.33%
74Nvidia1072.00%
75Perceptron10$1.50100.0%
76PrimeIntellect11GLM 5.3$4.40100.0%
77Stealth10100.0%
78StepFun11Step 3.7 Flash$1.1599.48%
79Unbiased10$7.50100.0%

LLM API providers: common questions

How many LLM API providers does ModelCap track?

79 inference providers with at least one live endpoint for a tracked model as of 23 September 2026, observed through OpenRouter. Each provider name links to a page listing every model it serves with that provider's own price, context window, quantization and uptime.

Which provider serves the most AI models?

DeepInfra, with 74 tracked models across 78 endpoints. Model count is unique canonical models, so a product served from several deployments counts once.

Which provider has the cheapest LLM API pricing?

The lowest positive language-token entry price is $0.03 per 1M output tokens on DeepInfra. Entry price is the cheapest model a provider lists, not a like-for-like comparison; the cheapest-LLMs ranking and the compare tool line up specific models across providers.

Which provider has the best uptime?

AionLabs shows the highest mean measured uptime, 100.0%, over the last 30 minutes at the time of the snapshot; 79 of 79 providers report a measurement. Uptime is per endpoint from OpenRouter and averaged with equal weight, so it is a live signal rather than an SLA.

How are provider prices and uptime measured?

Both come from OpenRouter's per-endpoint catalogue and telemetry, re-observed on every ModelCap refresh. Prices are USD per 1M tokens as listed by the provider; uptime is the provider's own endpoint availability over the last 30 minutes. Offers retained from a failed refresh are marked stale and excluded from every count on this page.

Explore further