- Benchmarks
- Providers
52 providers · 265 endpoints · snapshot October 6, 2026
Inference providers, priced side by side.
The same model often costs very different amounts from one provider to the next. Each dot below is one provider; the diamond is the model maker’s own list price. Prices are reported by OpenRouter’s public API; we did not measure them, and the API reported no latency.
22 models · Select a provider to pin it on every row.
- DeepSeek V4 Flash 0423Spread 12.6x: highest ÷ lowest Blended 3:1 price (calculation)
- DeepSeek V4 Pro 0423Spread 11.2x: highest ÷ lowest Blended 3:1 price (calculation)
- gpt-oss-120bSpread 6.9x: highest ÷ lowest Blended 3:1 price (calculation)
- Llama 3.3 70B InstructSpread 6.7x: highest ÷ lowest Blended 3:1 price (calculation)
- GLM 5.3Spread 4.7x: highest ÷ lowest Blended 3:1 price (calculation)
- Kimi K3Spread 1.8x: highest ÷ lowest Blended 3:1 price (calculation)
- Llama 4 MaverickSpread 1.7x: highest ÷ lowest Blended 3:1 price (calculation)
- 15 more models: one blended 3:1 price at every provider1x
- one provider (blended: a calculation, hollow)
- first-party list price
- Spread: highest ÷ lowest price of the row (calculation)
| Model | Provider | Precision | Input $/M | Output $/M | Cache read $/M | Blended $/M (calculation) | Uptime, last day |
|---|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0423 | StreamLake | fp8 | $0.042 | $0.084 | $0.0084 | $0.0525 | 99% |
| DeepSeek V4 Flash 0423 | DeepInfra | fp8 | $0.09 | $0.18 | $0.018 | $0.1125 | 100% |
| DeepSeek V4 Flash 0423 | GMICloud | fp8 | $0.091 | $0.182 | $0.0182 | $0.1138 | 100% |
| DeepSeek V4 Flash 0423 | Venice | not reported | $0.0966 | $0.1925 | $0.0196 | $0.1206 | 96% |
| DeepSeek V4 Flash 0423 | DigitalOcean | not reported | $0.098 | $0.196 | $0.0196 | $0.1225 | 100% |
| DeepSeek V4 Flash 0423 | Alibaba | fp8 | $0.134 | $0.268 | $0.0268 | $0.1675 | 99% |
| DeepSeek V4 Flash 0423 | SiliconFlow | fp8 | $0.13 | $0.28 | $0.028 | $0.1675 | 100% |
| DeepSeek V4 Flash 0423 | AtlasCloud | fp4 | $0.14 | $0.28 | $0.028 | $0.175 | 99% |
| DeepSeek V4 Flash 0423 | Baidu | fp8 | $0.14 | $0.28 | $0.028 | $0.175 | 99% |
| DeepSeek V4 Flash 0423 | Novita | fp8 | $0.14 | $0.28 | $0.028 | $0.175 | 100% |
| DeepSeek V4 Flash 0423 | Parasail | fp8 | $0.14 | $0.28 | $0.07 | $0.175 | 100% |
| DeepSeek V4 Flash 0423 | Mancer 2 | fp8 | $0.19 | $0.50 | — | $0.2675 | 95% |
| DeepSeek V4 Flash 0423 | Relace | fp4 | $0.012 | $1.28 | $0.012 | $0.329 | 100% |
| DeepSeek V4 Flash 0423 | OpenInference | fp4 | $0.0132 | $1.41 | $0.0132 | $0.3619 | 99% |
| DeepSeek V4 Flash 0423 | Cloudflare | not reported | $0.44 | $1.32 | $0.014 | $0.66 | 98% |
| DeepSeek V4 Pro 0423 | StreamLake | fp8 | $0.2088 | $0.4176 | $0.0174 | $0.261 | 99% |
| DeepSeek V4 Pro 0423 | GMICloud | fp8 | $0.957 | $1.91 | $0.0798 | $1.20 | 97% |
| DeepSeek V4 Pro 0423 | Relace | fp4 | $0.2067 | $4.20 | $0.21 | $1.21 | 100% |
| DeepSeek V4 Pro 0423 | Parasail | fp8 | $0.45 | $3.48 | $0.10 | $1.21 | 98% |
| DeepSeek V4 Pro 0423 | DigitalOcean | not reported | $1.04 | $2.09 | $0.2088 | $1.31 | 100% |
| DeepSeek V4 Pro 0423 | Cloudflare | not reported | $1.15 | $2.55 | $0.20 | $1.50 | 98% |
| DeepSeek V4 Pro 0423 | DeepInfra | fp8 | $1.30 | $2.60 | $0.10 | $1.63 | 100% |
| DeepSeek V4 Pro 0423 | Alibaba | fp8 | $1.42 | $2.83 | $0.118 | $1.77 | 90% |
| DeepSeek V4 Pro 0423 | SiliconFlow | fp8 | $1.50 | $3.13 | $0.135 | $1.91 | 99% |
| DeepSeek V4 Pro 0423 | Novita | fp8 | $1.60 | $3.20 | $0.135 | $2.00 | 100% |
| DeepSeek V4 Pro 0423 | Venice | not reported | $1.65 | $3.30 | $0.33 | $2.06 | 97% |
| DeepSeek V4 Pro 0423 | AtlasCloud | fp4 | $1.68 | $3.38 | $0.13 | $2.10 | 99% |
| DeepSeek V4 Pro 0423 | Baidu | fp8 | $1.69 | $3.38 | $0.14 | $2.11 | 100% |
| DeepSeek V4 Pro 0423 | NextBit | fp8 | $1.74 | $3.48 | $0.145 | $2.17 | 99% |
| DeepSeek V4 Pro 0423 | Reka | not reported | $0.90 | $9.00 | $0.18 | $2.92 | 99% |
| gpt-oss-120b | CoreWeave | fp4 | $0.03 | $0.17 | $0.03 | $0.065 | 99% |
| gpt-oss-120b | DekaLLM | bf16 | $0.03 | $0.18 | $0.03 | $0.0675 | 100% |
| gpt-oss-120b | DeepInfra | bf16 | $0.037 | $0.17 | — | $0.0703 | 99% |
| gpt-oss-120b | AkashML | bf16 | $0.037 | $0.187 | $0.037 | $0.0745 | 100% |
| gpt-oss-120b | Mancer 2 | fp8 | $0.045 | $0.25 | — | $0.0963 | 99% |
| gpt-oss-120b | Crusoe | bf16 | $0.05 | $0.25 | $0.05 | $0.10 | 100% |
| gpt-oss-120b | Novita | fp4 | $0.05 | $0.25 | — | $0.10 | 99% |
| gpt-oss-120b | DigitalOcean | not reported | $0.06 | $0.42 | $0.012 | $0.15 | 100% |
| gpt-oss-120b | Google Vertex | not reported | $0.09 | $0.36 | — | $0.1575 | 67% |
| gpt-oss-120b | BaseTen | fp4 | $0.10 | $0.50 | $0.10 | $0.20 | 100% |
| gpt-oss-120b | Amazon Bedrock | not reported | $0.15 | $0.60 | — | $0.2625 | 99% |
| gpt-oss-120b | Groq | not reported | $0.15 | $0.60 | $0.075 | $0.2625 | 99% |
| gpt-oss-120b | Nebius | fp4 | $0.15 | $0.60 | — | $0.2625 | 97% |
| gpt-oss-120b | Phala | not reported | $0.15 | $0.60 | — | $0.2625 | 99% |
| gpt-oss-120b | SiliconFlow | fp8 | $0.15 | $0.60 | $0.075 | $0.2625 | 83% |
| gpt-oss-120b | Together | not reported | $0.15 | $0.60 | — | $0.2625 | 87% |
| gpt-oss-120b | Parasail | fp4 | $0.10 | $0.75 | $0.055 | $0.2625 | 100% |
| gpt-oss-120b | Mara | not reported | $0.15 | $0.75 | — | $0.30 | 97% |
| gpt-oss-120b | SambaNova | not reported | $0.14 | $0.95 | — | $0.3425 | 100% |
| gpt-oss-120b | Cerebras | fp16 | $0.35 | $0.75 | $0.35 | $0.45 | 100% |
| Llama 3.3 70B Instruct | DeepInfra | fp8 | $0.10 | $0.32 | — | $0.155 | 98% |
| Llama 3.3 70B Instruct | Novita | bf16 | $0.135 | $0.40 | — | $0.2013 | 98% |
| Llama 3.3 70B Instruct | AkashML | fp8 | $0.20 | $0.52 | $0.10 | $0.28 | 99% |
| Llama 3.3 70B Instruct | Parasail | fp8 | $0.22 | $0.50 | $0.11 | $0.29 | 100% |
| Llama 3.3 70B Instruct | SambaNova | not reported | $0.45 | $0.90 | — | $0.5625 | 99% |
| Llama 3.3 70B Instruct | Groq | not reported | $0.59 | $0.79 | $0.295 | $0.64 | 100% |
| Llama 3.3 70B Instruct | CoreWeave | fp16 | $0.71 | $0.71 | $0.71 | $0.71 | 98% |
| Llama 3.3 70B Instruct | Google Vertex | not reported | $0.72 | $0.72 | — | $0.72 | — |
| Llama 3.3 70B Instruct | Cloudflare | fp8 | $0.293 | $2.25 | — | $0.783 | 99% |
| Llama 3.3 70B Instruct | Together | not reported | $1.04 | $1.04 | — | $1.04 | 93% |
| GLM 5.3 | Novita | fp8 | $0.42 | $1.32 | $0.078 | $0.645 | 97% |
| GLM 5.3 | Reka | not reported | $0.17 | $3.00 | $0.169 | $0.8775 | 100% |
| GLM 5.3 | Sail Research | fp8 | $0.20 | $3.40 | $0.15 | $1.00 | 99% |
| GLM 5.3 | Morph | fp8 | $0.179 | $3.55 | $0.137 | $1.02 | 99% |
| GLM 5.3 | DeepInfra | fp4 | $0.5625 | $2.50 | $0.125 | $1.05 | 97% |
| GLM 5.3 | SiliconFlow | fp8 | $0.70 | $2.20 | $0.13 | $1.07 | 100% |
| GLM 5.3 | InferenceNet | not reported | $0.14 | $4.40 | $0.07 | $1.21 | 100% |
| GLM 5.3 | Makora | fp4 | $0.18 | $4.40 | $0.19 | $1.24 | 96% |
| GLM 5.3 | AkashML | fp8 | $0.19 | $4.40 | $0.19 | $1.24 | 100% |
| GLM 5.3 | Phala | not reported | $0.84 | $2.64 | $0.156 | $1.29 | 99% |
| GLM 5.3 | Inceptron | fp4 | $0.60 | $3.39 | $0.20 | $1.30 | 98% |
| GLM 5.3 | DigitalOcean | not reported | $0.91 | $2.86 | $0.169 | $1.40 | 100% |
| GLM 5.3 | GMICloud | fp8 | $0.98 | $3.08 | $0.182 | $1.50 | 99% |
| GLM 5.3 | Alibaba | not reported | $1.19 | $3.74 | $0.238 | $1.83 | 100% |
| GLM 5.3 | Decart | fp4 | $1.19 | $3.74 | $0.1955 | $1.83 | 100% |
| GLM 5.3 | Wafer | not reported | $0.15 | $7.00 | $0.14 | $1.86 | 100% |
| GLM 5.3 | Friendli | not reported | $1.26 | $3.96 | $0.234 | $1.94 | 100% |
| GLM 5.3 | AtlasCloud | fp8 | $1.40 | $4.40 | $0.26 | $2.15 | 100% |
| GLM 5.3 | Baidu | fp8 | $1.40 | $4.40 | $0.26 | $2.15 | 100% |
| GLM 5.3 | BaseTen | fp4 | $1.40 | $4.40 | $0.14 | $2.15 | 98% |
| GLM 5.3 | Cloudflare | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 98% |
| GLM 5.3 | Crusoe | fp4 | $1.40 | $4.40 | $0.26 | $2.15 | 98% |
| GLM 5.3 | Fireworks | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 100% |
| GLM 5.3 | Mistral | nvfp4 | $1.40 | $4.40 | $0.14 | $2.15 | 100% |
| GLM 5.3 | Modal | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 98% |
| GLM 5.3 | Nebius | fp4 | $1.40 | $4.40 | — | $2.15 | 96% |
| GLM 5.3 | Parasail | fp8 | $1.40 | $4.40 | $0.26 | $2.15 | 100% |
| GLM 5.3 | PrimeIntellect | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 99% |
| GLM 5.3 | Together | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 96% |
| GLM 5.3 | Venice | not reported | $1.40 | $4.40 | $0.26 | $2.15 | 99% |
| GLM 5.3 | Z.AI | fp8 | $1.40 | $4.40 | $0.26 | $2.15 | 100% |
| GLM 5.3 | Relace | not reported | $0.03 | $12.00 | $0.03 | $3.02 | 100% |
| Kimi K3 | Relace | fp4 | $0.83 | $13.00 | $0.45 | $3.87 | 100% |
| Kimi K3 | Phala | not reported | $1.95 | $9.75 | $0.195 | $3.90 | 97% |
| Kimi K3 | Sail Research | fp4 | $0.84 | $13.50 | $0.30 | $4.00 | 100% |
| Kimi K3 | Decart | mxfp4 | $2.01 | $10.05 | $0.201 | $4.02 | 87% |
| Kimi K3 | InferenceNet | fp4 | $0.95 | $14.00 | $0.31 | $4.21 | 100% |
| Kimi K3 | Wafer | not reported | $0.95 | $14.00 | $0.40 | $4.21 | 99% |
| Kimi K3 | Morph | fp8 | $1.27 | $13.30 | $0.278 | $4.28 | 100% |
| Kimi K3 | Makora | not reported | $1.53 | $12.75 | $0.204 | $4.33 | 97% |
| Kimi K3 | AkashML | fp4 | $1.30 | $14.00 | $1.30 | $4.47 | 98% |
| Kimi K3 | DigitalOcean | not reported | $2.55 | $12.95 | $0.255 | $5.15 | 100% |
| Kimi K3 | Together | not reported | $2.70 | $13.50 | $0.27 | $5.40 | 99% |
| Kimi K3 | DeepInfra | mxfp4 | $2.85 | $14.25 | $0.285 | $5.70 | 99% |
| Kimi K3 | BaseTen | fp8 | $3.00 | $15.00 | $0.30 | $6.00 | 98% |
| Kimi K3 | Chutes | mxfp4 | $3.00 | $15.00 | $0.30 | $6.00 | 97% |
| Kimi K3 | Fireworks | not reported | $3.00 | $15.00 | $0.30 | $6.00 | 99% |
| Kimi K3 | Modal | mxfp4 | $3.00 | $15.00 | $0.30 | $6.00 | 98% |
| Kimi K3 | Moonshot AI | mxfp4 | $3.00 | $15.00 | $0.30 | $6.00 | 100% |
| Kimi K3 | Parasail | fp4 | $3.00 | $15.00 | $0.30 | $6.00 | 98% |
| Kimi K3 | Alibaba | not reported | $3.45 | $17.25 | $0.345 | $6.90 | 99% |
| Llama 4 Maverick | DigitalOcean | not reported | $0.1875 | $0.6525 | — | $0.3037 | 99% |
| Llama 4 Maverick | Novita | fp8 | $0.27 | $0.85 | — | $0.415 | 98% |
| Llama 4 Maverick | Parasail | fp8 | $0.35 | $1.00 | $0.17 | $0.5125 | 100% |
| Claude Haiku 4.5 | Amazon Bedrock | not reported | $1.00 | $5.00 | $0.10 | $2.00 | 100% |
| Claude Haiku 4.5 | Anthropic | not reported | $1.00 | $5.00 | $0.10 | $2.00 | 100% |
| Claude Haiku 4.5 | Azure | not reported | $1.00 | $5.00 | $0.10 | $2.00 | 100% |
| Claude Haiku 4.5 | Google Vertex | not reported | $1.00 | $5.00 | $0.10 | $2.00 | 100% |
| Claude Sonnet 5 | Amazon Bedrock | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5 | Anthropic | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5 | Azure | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5 | Claude Platform on AWS | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5 | Google Vertex | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5.5 | Amazon Bedrock | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5.5 | Anthropic | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5.5 | Azure | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5.5 | Claude Platform on AWS | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Sonnet 5.5 | Google Vertex | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| Claude Opus 4.8 | Amazon Bedrock | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 95% |
| Claude Opus 4.8 | Anthropic | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 4.8 | Azure | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 98% |
| Claude Opus 4.8 | Claude Platform on AWS | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 4.8 | Google Vertex | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 5 | Amazon Bedrock | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 5 | Anthropic | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 5 | Azure | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 5 | Claude Platform on AWS | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 99% |
| Claude Opus 5 | Google Vertex | not reported | $5.00 | $25.00 | $0.50 | $10.00 | 100% |
| Claude Opus 5.5 | Amazon Bedrock | not reported | $4.00 | $20.00 | $0.20 | $8.00 | 97% |
| Claude Opus 5.5 | Anthropic | not reported | $4.00 | $20.00 | $0.20 | $8.00 | 100% |
| Claude Opus 5.5 | Azure | not reported | $4.00 | $20.00 | $0.20 | $8.00 | 100% |
| Claude Opus 5.5 | Claude Platform on AWS | not reported | $4.00 | $20.00 | $0.20 | $8.00 | 100% |
| Claude Opus 5.5 | Google Vertex | not reported | $4.00 | $20.00 | $0.20 | $8.00 | 100% |
| Claude Fable 5.1 | Amazon Bedrock | not reported | $10.00 | $50.00 | $0.25 | $20.00 | — |
| Claude Fable 5.1 | Anthropic | not reported | $10.00 | $50.00 | $0.25 | $20.00 | 100% |
| Claude Fable 5.1 | Azure | not reported | $10.00 | $50.00 | $0.25 | $20.00 | 99% |
| Claude Fable 5.1 | Google Vertex | not reported | $10.00 | $50.00 | $0.25 | $20.00 | 100% |
| GPT-6 Sol | Azure | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| GPT-6 Sol | OpenAI | not reported | $2.00 | $10.00 | $0.20 | $4.00 | 100% |
| GPT-6 Luna | Azure | not reported | $0.10 | $0.50 | $0.01 | $0.20 | 100% |
| GPT-6 Luna | OpenAI | not reported | $0.10 | $0.50 | $0.01 | $0.20 | 100% |
| GPT-6 Astra | Azure | not reported | $10.00 | $50.00 | $1.00 | $20.00 | 100% |
| GPT-6 Astra | OpenAI | not reported | $10.00 | $50.00 | $1.00 | $20.00 | 100% |
| GPT-5.5 | Azure | not reported | $5.00 | $30.00 | $0.50 | $11.25 | 100% |
| GPT-5.5 | OpenAI | not reported | $5.00 | $30.00 | $0.50 | $11.25 | 100% |
| Gemini 3.8 Flash | Google AI Studio | not reported | $0.75 | $3.75 | $0.075 | $1.50 | 100% |
| Gemini 3.8 Flash | Google Vertex | not reported | $0.75 | $3.75 | $0.075 | $1.50 | 97% |
| Gemini 3.5 Flash | Google AI Studio | not reported | $1.50 | $9.00 | $0.15 | $3.38 | 100% |
| Gemini 3.5 Flash | Google Vertex | not reported | $1.50 | $9.00 | $0.15 | $3.38 | 99% |
| Gemini 3.5 Flash Lite | Google AI Studio | not reported | $0.30 | $2.50 | $0.03 | $0.85 | 100% |
| Gemini 3.5 Flash Lite | Google Vertex | not reported | $0.30 | $2.50 | $0.03 | $0.85 | 100% |
| Gemini 3.1 Pro Preview | Google AI Studio | not reported | $2.00 | $12.00 | $0.20 | $4.50 | 100% |
| Gemini 3.1 Pro Preview | Google Vertex | not reported | $2.00 | $12.00 | $0.20 | $4.50 | 98% |
Third-party reported values. 163 listings, 4 series: Blended 3:1, Input, Output, Cache read. Blended 3:1: highest Claude Fable 5.1 · Amazon Bedrock $20.00 (n 4). Lowest DeepSeek V4 Flash 0423 · StreamLake (fp8) $0.053 (n 15). Input: highest Claude Fable 5.1 · Amazon Bedrock $10.00 (n 4). Lowest DeepSeek V4 Flash 0423 · Relace (fp4) $0.012 (n 15).
Notesn 2–32 per row
22 models with two or more standard-tier providers, widest spread first. USD per million tokens, snapshot 2026-10-06.
Prices reported by OpenRouter’s public API (third-party-reported, not measured by Agent). One dot per provider: its cheapest standard-tier endpoint. The diamond is the first-party list price. Spread: highest ÷ lowest price of the row. A parenthesis names the quantization the provider reported. n = providers with a standard-tier endpoint.
Source: OpenRouter public API: models and provider endpoints (snapshot)
Key numbers
265
Provider endpoints in the snapshot
endpoints, 52 providers, 27 models · n = 27
12.6x
Largest standard-tier price spread (DeepSeek V4 Flash 0423)
(most expensive: Cloudflare; cheapest: StreamLake (fp8)) · n = 15
10 of 10
Models where OpenRouter’s per-token price equals the first-party list price
n = 10
5.5%
Credit-purchase fee on OpenRouter’s Standard plan (third-party-reported)
($0.80 minimum by card)
0 of 265
Endpoints with a latency figure in the keyless API
n = 265
Gateways: what OpenRouter adds
A gateway resells other providers’ endpoints. Its per-token price matches the first-party list price; the fee comes when you buy credits.
0%per-token markup on 10 of 10 models
5.5% with the card credit feeCalculation
- Claude Haiku 4.5
- Claude Sonnet 5
- Claude Sonnet 5.5
- Claude Opus 4.8
- Claude Opus 5
- Claude Opus 5.5
- Claude Fable 5.1
- GPT-6 Luna
- Gemini 3.8 Flash
- Gemini 3.5 Flash
- per-token markup over the first-party list price (input and output)
- With the 5.5% card credit fee (input) (calculation, hollow)
| Item | Per-token markup (input) | Per-token markup (output) | With the 5.5% card credit fee (input) |
|---|---|---|---|
| Claude Haiku 4.5 | 0% | 0% | 5.5% |
| Claude Sonnet 5 | 0% | 0% | 5.5% |
| Claude Sonnet 5.5 | 0% | 0% | 5.5% |
| Claude Opus 4.8 | 0% | 0% | 5.5% |
| Claude Opus 5 | 0% | 0% | 5.5% |
| Claude Opus 5.5 | 0% | 0% | 5.5% |
| Claude Fable 5.1 | 0% | 0% | 5.5% |
| GPT-6 Luna | 0% | 0% | 5.5% |
| Gemini 3.8 Flash | 0% | 0% | 5.5% |
| Gemini 3.5 Flash | 0% | 0% | 5.5% |
List-price calculation, not a run. 10 rows, 3 series: Per-token markup (input), Per-token markup (output), With the 5.5% card credit fee (input). Per-token markup (input): all at 0%. Per-token markup (output): all at 0%.
Notes
Per-token markup, and the markup after the Standard credit-purchase fee (calculation)
A calculation: OpenRouter list price ÷ first-party list price − 1, then × (1 + 5.5%) for credits bought by card on the Standard plan (purchases large enough that the minimum fee does not apply). Fees as published on 2026-10-06; first-party prices from the vendors’ list prices.
Sources: OpenRouter public API: models and provider endpoints (snapshot), OpenRouter pricing and fees, Anthropic list prices (Claude models), OpenAI list prices, Google Gemini list prices
Providers with a page
The gateways, clouds and first-party APIs most people compare. The bar counts the models where its price is the lowest, tied for the lowest, or above the lowest: a calculation on reported prices (3 input : 1 output tokens).
20 providers
Inference provider · OpenRouter
OpenRouter
A gateway that routes one API to many inference providers. Its per-token price is compared with first-party list prices; its fee is charged when credits are bought.
A gateway: it resells other providers’ endpoints. See its markup.
Inference provider · Anthropic
Anthropic
Anthropic’s own API: the first-party list price of the Claude models, and Anthropic’s endpoint as listed on OpenRouter.
7 models in the index · median 1x the cheapest
Counted over 7 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · OpenAI
OpenAI
OpenAI’s own API: the first-party list price of GPT models, and OpenAI’s endpoints as listed on OpenRouter.
4 models in the index · median 1x the cheapest
Counted over 4 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Google
Google AI Studio
Google’s Gemini API (AI Studio): the first-party list price of Gemini models, and its endpoints as listed on OpenRouter.
4 models in the index · median 1x the cheapest
Counted over 4 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Google
Google Vertex AI
Google Cloud’s model platform. An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
14 models in the index · median 1x the cheapest
Counted over 13 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Amazon
Amazon Bedrock
Amazon Web Services’ model platform. An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
13 models in the index · median 1x the cheapest
Counted over 8 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Microsoft
Azure
Microsoft’s cloud model platform. An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
13 models in the index · median 1x the cheapest
Counted over 11 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Anthropic
Claude Platform on AWS
Anthropic’s Claude platform hosted on AWS. An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
5 models in the index · median 1x the cheapest
Counted over 5 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Groq
Groq
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
2 models in the index · median 4.1x the cheapest
Counted over 2 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Together AI
Together AI
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
4 models in the index · median 3.7x the cheapest
Counted over 4 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Fireworks AI
Fireworks AI
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
2 models in the index · median 2.4x the cheapest
Counted over 2 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · DeepInfra
DeepInfra
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
6 models in the index · median 1.5x the cheapest
Counted over 6 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Cerebras
Cerebras
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
1 model in the index · median 6.9x the cheapest
Counted over 1 model with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · SambaNova
SambaNova
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
2 models in the index · median 4.4x the cheapest
Counted over 2 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Nebius
Nebius
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
2 models in the index · median 3.7x the cheapest
Counted over 2 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Parasail
Parasail
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
7 models in the index · median 3.3x the cheapest
Counted over 7 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Novita AI
Novita AI
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
6 models in the index · median 1.5x the cheapest
Counted over 6 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Baseten
Baseten
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
3 models in the index · median 3.1x the cheapest
Counted over 3 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · Cloudflare
Cloudflare Workers AI
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
4 models in the index · median 5.4x the cheapest
Counted over 4 models with 2 or more standard-tier providers · blended 3:1, a calculation
Inference provider · SiliconFlow
SiliconFlow
An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
4 models in the index · median 3.6x the cheapest
Counted over 4 models with 2 or more standard-tier providers · blended 3:1, a calculation
One model, every provider
Input, output and cache-read prices on their own strips. Point at a provider to link its three prices; select it to pin it on every chart of this page. The Table view sorts any column.
Price per provider: DeepSeek V4 Flash 0423
Reported by OpenRouter’s public API, snapshot October 6, 2026. Third-party-reported prices, not measured by Agent.
15 providers with a standard-tier endpoint · cheapest StreamLake at $0.0525 blended, priciest Cloudflare at $0.66 (12.6x) · showing 15 of 15 rows
Input
Output
Cache read
- one provider (reported price, USD per million tokens, log scale per strip)
| vs cheapest | ||||||||
|---|---|---|---|---|---|---|---|---|
| StreamLake | fp8 | 1.024M | $0.042 | $0.084 | $0.0084 | $0.0525 | 99% | 1x |
| DeepInfra | fp8 | 1.048576M | $0.09 | $0.18 | $0.018 | $0.1125 | 100% | 2.1x |
| GMICloud | fp8 | 1.048575M | $0.091 | $0.182 | $0.0182 | $0.1138 | 100% | 2.2x |
| Venice | not reported | 1M | $0.0966 | $0.1925 | $0.0196 | $0.1206 | 96% | 2.3x |
| DigitalOcean | not reported | 1.048576M | $0.098 | $0.196 | $0.0196 | $0.1225 | 100% | 2.3x |
| Alibaba | fp8 | 1M | $0.134 | $0.268 | $0.0268 | $0.1675 | 99% | 3.2x |
| SiliconFlow | fp8 | 1.048576M | $0.13 | $0.28 | $0.028 | $0.1675 | 100% | 3.2x |
| AtlasCloud | fp4 | 1.048576M | $0.14 | $0.28 | $0.028 | $0.175 | 99% | 3.3x |
| Baidu | fp8 | 1.048576M | $0.14 | $0.28 | $0.028 | $0.175 | 99% | 3.3x |
| Novita | fp8 | 1.048576M | $0.14 | $0.28 | $0.028 | $0.175 | 100% | 3.3x |
| Parasail | fp8 | 1.048576M | $0.14 | $0.28 | $0.07 | $0.175 | 100% | 3.3x |
| Mancer 2 | fp8 | 1.048576M | $0.19 | $0.50 | — | $0.2675 | 95% | 5.1x |
| Relace | fp4 | 1.048576M | $0.012 | $1.28 | $0.012 | $0.329 | 100% | 6.3x |
| OpenInference | fp4 | 1.048576M | $0.0132 | $1.41 | $0.0132 | $0.3619 | 99% | 6.9x |
| Cloudflare | not reported | 384k | $0.44 | $1.32 | $0.014 | $0.66 | 98% | 12.6x |
Blended price and “vs cheapest” are calculations on the reported prices (3 input : 1 output tokens), against the cheapest standard-tier provider. A lower price can come with lower precision (fp4, fp8) or a shorter context. Uptime is what the API reported for the last day; the API reported no latency or throughput.
Provider vs provider
All comparisonsPrice rows never name a winner: a reported list price has no interval, so a gap is stated, not ranked.
OpenRouter vs Anthropic
14 ties
OpenRouter vs Google AI Studio
4 ties
OpenRouter vs OpenAI
2 ties
Groq vs Together AI
2 ties · 2 unclear
Groq vs Cerebras
3 unclear
Together AI vs Fireworks AI
3 ties · 3 unclear
Amazon Bedrock vs Google Vertex AI
21 ties · 2 unclear
Amazon Bedrock vs Azure
21 ties
Anthropic vs Amazon Bedrock
21 ties
Where every provider’s price sits
52 providers listed for the 27 models in the index. “Sole cheapest” counts models where it alone has the lowest standard-tier blended price and at least one other provider serves the model. “Tied cheapest” counts models where it shares the lowest price with other providers.
- Sole cheapest
- Tied cheapest
- Above cheapest
- Google Vertex
- Azure
- Amazon Bedrock
- Anthropic
- Parasail
- Alibaba
- DeepInfra
- DigitalOcean
- Novita
- Claude Platform on AWS
- OpenAI
- Google AI Studio
- AkashML
- Cloudflare
- Relace
- SiliconFlow
- Together
- Mistral
- BaseTen
- AtlasCloud
- Baidu
- GMICloud
- Phala
- Venice
- Fireworks
- Wafer
- InferenceNet
- Sail Research
- CoreWeave
- Crusoe
- Decart
- Groq
- Makora
- Mancer 2
- Modal
- Morph
- Nebius
- Reka
- SambaNova
- StreamLake
- xAI
- Cerebras
- Chutes
- DekaLLM
- Friendli
- Inceptron
- Mara
- Moonshot AI
- NextBit
- OpenInference
- PrimeIntellect
- Z.AI
| Provider | Models | Endpoints | Sole cheapest | Tied cheapest | Above cheapest | Median uptime, last day |
|---|---|---|---|---|---|---|
| Google Vertex | 14 | 38 | 0 | 11 | 2 | 100% |
| Azure | 13 | 25 | 0 | 11 | 0 | 100% |
| Amazon Bedrock | 13 | 24 | 0 | 7 | 1 | 100% |
| Anthropic | 7 | 10 | 0 | 7 | 0 | 100% |
| Parasail | 7 | 7 | 0 | 0 | 7 | 100% |
| Alibaba | 6 | 7 | 0 | 0 | 4 | 99% |
| DeepInfra | 6 | 7 | 1 | 0 | 5 | 99% |
| DigitalOcean | 6 | 6 | 1 | 0 | 5 | 100% |
| Novita | 6 | 6 | 1 | 0 | 5 | 98% |
| Claude Platform on AWS | 5 | 5 | 0 | 5 | 0 | 100% |
| OpenAI | 4 | 13 | 0 | 4 | 0 | 100% |
| Google AI Studio | 4 | 12 | 0 | 4 | 0 | 100% |
| AkashML | 4 | 4 | 0 | 0 | 4 | 100% |
| Cloudflare | 4 | 4 | 0 | 0 | 4 | 98% |
| Relace | 4 | 4 | 1 | 0 | 3 | 100% |
| SiliconFlow | 4 | 4 | 0 | 0 | 4 | 99% |
| Together | 4 | 4 | 0 | 0 | 4 | 95% |
| Mistral | 3 | 8 | 0 | 0 | 1 | 100% |
| BaseTen | 3 | 7 | 0 | 0 | 3 | 99% |
| AtlasCloud | 3 | 3 | 0 | 0 | 3 | 99% |
| Baidu | 3 | 3 | 0 | 0 | 3 | 100% |
| GMICloud | 3 | 3 | 0 | 0 | 3 | 99% |
| Phala | 3 | 3 | 0 | 0 | 3 | 99% |
| Venice | 3 | 3 | 0 | 0 | 3 | 97% |
| Fireworks | 2 | 5 | 0 | 0 | 2 | 99% |
| Wafer | 2 | 4 | 0 | 0 | 2 | 100% |
| InferenceNet | 2 | 3 | 0 | 0 | 2 | 100% |
| Sail Research | 2 | 3 | 0 | 0 | 2 | 99% |
| CoreWeave | 2 | 2 | 1 | 0 | 1 | 98% |
| Crusoe | 2 | 2 | 0 | 0 | 2 | 99% |
| Decart | 2 | 2 | 0 | 0 | 2 | 93% |
| Groq | 2 | 2 | 0 | 0 | 2 | 100% |
| Makora | 2 | 2 | 0 | 0 | 2 | 97% |
| Mancer 2 | 2 | 2 | 0 | 0 | 2 | 97% |
| Modal | 2 | 2 | 0 | 0 | 2 | 98% |
| Morph | 2 | 2 | 0 | 0 | 2 | 99% |
| Nebius | 2 | 2 | 0 | 0 | 2 | 97% |
| Reka | 2 | 2 | 0 | 0 | 2 | 99% |
| SambaNova | 2 | 2 | 0 | 0 | 2 | 99% |
| StreamLake | 2 | 2 | 2 | 0 | 0 | 99% |
| xAI | 1 | 5 | 0 | 0 | 0 | 99% |
| Cerebras | 1 | 1 | 0 | 0 | 1 | 100% |
| Chutes | 1 | 1 | 0 | 0 | 1 | 97% |
| DekaLLM | 1 | 1 | 0 | 0 | 1 | 100% |
| Friendli | 1 | 1 | 0 | 0 | 1 | 100% |
| Inceptron | 1 | 1 | 0 | 0 | 1 | 98% |
| Mara | 1 | 1 | 0 | 0 | 1 | 97% |
| Moonshot AI | 1 | 1 | 0 | 0 | 1 | 100% |
| NextBit | 1 | 1 | 0 | 0 | 1 | 99% |
| OpenInference | 1 | 1 | 0 | 0 | 1 | 99% |
| PrimeIntellect | 1 | 1 | 0 | 0 | 1 | 99% |
| Z.AI | 1 | 1 | 0 | 0 | 1 | 100% |
List-price calculation, not a run. 52 providers, 3 series: Sole cheapest, Tied cheapest, Above cheapest. Sole cheapest: highest StreamLake 2. Lowest Z.AI 0. Tied cheapest: highest Google Vertex 11. Lowest Z.AI 0.
Notes
Models with two or more standard-tier providers: sole cheapest, tied for cheapest, or above the cheapest (blended 3:1)
Counts of a calculation on reported prices (blended 3 input : 1 output). Not a score: a price says nothing about speed or quality.
Source: OpenRouter public API: models and provider endpoints (snapshot)
Reported by OpenRouter’s public API, snapshot October 6, 2026. Uptime is as reported, not measured by Agent.
How these numbers were made
The rules
- One snapshot of OpenRouter’s public, keyless API on October 6, 2026. Each price is copied as the API reported it. Agent measured none of them.
- One dot per provider and model: its cheapest standard-tier endpoint. The Table views list every endpoint and tier.
- Blended price = (3 × input + output) ÷ 4. Spread = highest ÷ lowest blended price of one model. Both are calculations on the reported prices.
- Providers at the same price are tied, not ranked. A price has no interval, so no page names a winner on price.
The limits
- Every price is third-party-reported by OpenRouter’s API at 2026-10-06. Prices change often; refetch before relying on them.
- A provider listed on OpenRouter is reached through OpenRouter; its price there may differ from the price on the provider’s own site.
- The cheapest endpoint may run lower precision (fp4 or fp8) or a shorter context. Price alone does not make two endpoints equal.
- No latency or throughput figure: the keyless API returned none. A cheaper provider is not shown to be slower or faster.
- First-party prices in the product table were verified on an earlier date than the snapshot; a vendor price change in between would show as a markup.
- Comparison rows between providers name no winner: a price has no interval, so the rule for winners does not apply. The gap is stated.
The full method Download the data (JSON)Every chart point (CSV)