Inference provider · Google
Google Vertex AI
Google Cloud’s model platform. An inference provider listed on OpenRouter. Its prices here are the ones OpenRouter’s public API reported for its endpoints.
Prices reported by OpenRouter’s public API, snapshot · third-party-reported, not measured by Agent
Where its price sits
11 of 13models where its price is the cheapest, ties included
Models with two or more standard-tier providers. Ties included in “cheapest”; a calculation on reported prices, 3 input : 1 output tokens. Not a score.
At a glance
14
models in the index (of 27)
Listed under Google Vertex
1x
median price vs the cheapest provider of the same model
Calculation · n = 13 models with 2 or more providers
100%
median uptime over the last day, as reported
n = 12 endpoints; reported, not measured
Not reported
latency and throughput
The keyless API returned latency for 0 of 265 endpoints. Gateway delay is not measured: no key in the environment.
Its price vs every other provider
Its cheapest standard-tier endpoint per model, against every other provider of that model. Switch the price type; select another provider to pin it instead. Blended price, position and “vs cheapest” are calculations on the reported prices. A cheaper endpoint can run lower precision or a shorter context.
Pinned Google Vertex: on 13 of 14 rows
- Claude Haiku 4.5
- Claude Sonnet 5
- Claude Sonnet 5.5
- Claude Opus 4.8
- Claude Opus 5
- Claude Opus 5.5
- Claude Fable 5.1
- gpt-oss-120b
- Gemini 3.8 Flash
- Gemini 3.5 Flash
- Gemini 3.5 Flash Lite
- Gemini 3.1 Pro Preview
- Llama 4 Maverick
- Llama 3.3 70B Instruct
- one provider (blended: a calculation, hollow)
- first-party list price
| Model | Input · output $/M | Blended $/M | Position | vs cheapest | Uptime, last day | Precision |
|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | $1.00 · $5.00 | $2.00 | tied cheapest of 4 | 1x | 100% | not reported |
| Claude Sonnet 5 | $2.00 · $10.00 | $4.00 | tied cheapest of 5 | 1x | 100% | not reported |
| Claude Sonnet 5.5 | $2.00 · $10.00 | $4.00 | tied cheapest of 5 | 1x | 100% | not reported |
| Claude Opus 4.8 | $5.00 · $25.00 | $10.00 | tied cheapest of 5 | 1x | 100% | not reported |
| Claude Opus 5 | $5.00 · $25.00 | $10.00 | tied cheapest of 5 | 1x | 100% | not reported |
| Claude Opus 5.5 | $4.00 · $20.00 | $8.00 | tied cheapest of 5 | 1x | 100% | not reported |
| Claude Fable 5.1 | $10.00 · $50.00 | $20.00 | tied cheapest of 4 | 1x | 100% | not reported |
| gpt-oss-120b | $0.09 · $0.36 | $0.1575 | 9th of 20 | 2.4x | 67% | not reported |
| Gemini 3.8 Flash | $0.75 · $3.75 | $1.50 | tied cheapest of 2 | 1x | 97% | not reported |
| Gemini 3.5 Flash | $1.50 · $9.00 | $3.38 | tied cheapest of 2 | 1x | 99% | not reported |
| Gemini 3.5 Flash Lite | $0.30 · $2.50 | $0.85 | tied cheapest of 2 | 1x | 100% | not reported |
| Gemini 3.1 Pro Preview | $2.00 · $12.00 | $4.50 | tied cheapest of 2 | 1x | 98% | not reported |
| Llama 4 Maverickregional tier only | $0.35 · $1.15 | $0.55 | no standard-tier endpoint | — | — | not reported |
| Llama 3.3 70B Instruct | $0.72 · $0.72 | $0.72 | 8th of 10 | 4.6x | — | not reported |
Third-party reported values. 74 listings, 4 series: Blended 3:1, Input, Output, Cache read. Blended 3:1: highest Claude Fable 5.1 · Amazon Bedrock $20.00 (n 4). Lowest gpt-oss-120b · CoreWeave (fp4) $0.065 (n 20). Input: highest Claude Fable 5.1 · Amazon Bedrock $10.00 (n 4). Lowest gpt-oss-120b · DekaLLM (bf16) $0.03 (n 20).
Notesn 2–20 per row
One row per model it serves; Google Vertex AI is pinned in color, every other provider is gray. USD per million tokens, snapshot 2026-10-06.
Prices reported by OpenRouter’s public API (third-party-reported, not measured by Agent). One dot per provider: its cheapest standard-tier endpoint. The diamond is the first-party list price. Spread: highest ÷ lowest price of the row. A parenthesis names the quantization the provider reported. n = providers with a standard-tier endpoint.
Source: OpenRouter public API: models and provider endpoints (snapshot)
Price per provider: Claude Haiku 4.5
Reported by OpenRouter’s public API, snapshot October 6, 2026. Third-party-reported prices, not measured by Agent.
4 providers with a standard-tier endpoint · all 4 providers report the same price ($2.00 blended) · showing 4 of 4 rows
Input
Output
Cache read
- one provider (reported price, USD per million tokens, log scale per strip)
- first-party list price
| vs cheapest | ||||||||
|---|---|---|---|---|---|---|---|---|
| Amazon Bedrock | not reported | 200k | $1.00 | $5.00 | $0.10 | $2.00 | 100% | 1x |
| Anthropic | not reported | 200k | $1.00 | $5.00 | $0.10 | $2.00 | 100% | 1x |
| Azure | not reported | 200k | $1.00 | $5.00 | $0.10 | $2.00 | 100% | 1x |
| Google Vertex | not reported | 200k | $1.00 | $5.00 | $0.10 | $2.00 | 100% | 1x |
Blended price and “vs cheapest” are calculations on the reported prices (3 input : 1 output tokens), against the cheapest standard-tier provider. A lower price can come with lower precision (fp4, fp8) or a shorter context. Uptime is what the API reported for the last day; the API reported no latency or throughput.
Every value from the study
Each row is one of its prices on its own track, with the other providers of the same chart as muted dots (37 values). Reported prices, not measured. Use Table for the plain values; each row links to the chart it copies.
| Metric | Value | Interval or range | Context |
|---|---|---|---|
| Claude Haiku 4.5: price per million tokens by provider (Input) | $1.00 | — | Claude Haiku 4.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Haiku 4.5: price per million tokens by provider (Output) | $5.00 | — | Claude Haiku 4.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Haiku 4.5: price per million tokens by provider (Cache read) | $0.10 | — | Claude Haiku 4.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5: price per million tokens by provider (Input) | $2.00 | — | Claude Sonnet 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5: price per million tokens by provider (Output) | $10.00 | — | Claude Sonnet 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5: price per million tokens by provider (Cache read) | $0.20 | — | Claude Sonnet 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5.5: price per million tokens by provider (Input) | $2.00 | — | Claude Sonnet 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5.5: price per million tokens by provider (Output) | $10.00 | — | Claude Sonnet 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Sonnet 5.5: price per million tokens by provider (Cache read) | $0.20 | — | Claude Sonnet 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 4.8: price per million tokens by provider (Input) | $5.00 | — | Claude Opus 4.8 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 4.8: price per million tokens by provider (Output) | $25.00 | — | Claude Opus 4.8 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 4.8: price per million tokens by provider (Cache read) | $0.50 | — | Claude Opus 4.8 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5: price per million tokens by provider (Input) | $5.00 | — | Claude Opus 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5: price per million tokens by provider (Output) | $25.00 | — | Claude Opus 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5: price per million tokens by provider (Cache read) | $0.50 | — | Claude Opus 5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5.5: price per million tokens by provider (Input) | $4.00 | — | Claude Opus 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5.5: price per million tokens by provider (Output) | $20.00 | — | Claude Opus 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Opus 5.5: price per million tokens by provider (Cache read) | $0.20 | — | Claude Opus 5.5 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Fable 5.1: price per million tokens by provider (Input) | $10.00 | — | Claude Fable 5.1 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Fable 5.1: price per million tokens by provider (Output) | $50.00 | — | Claude Fable 5.1 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Claude Fable 5.1: price per million tokens by provider (Cache read) | $0.25 | — | Claude Fable 5.1 · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| gpt-oss-120b: price per million tokens by provider (Input) | $0.090 | — | gpt-oss-120b · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| gpt-oss-120b: price per million tokens by provider (Output) | $0.36 | — | gpt-oss-120b · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.8 Flash: price per million tokens by provider (Input) | $0.75 | — | Gemini 3.8 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.8 Flash: price per million tokens by provider (Output) | $3.75 | — | Gemini 3.8 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.8 Flash: price per million tokens by provider (Cache read) | $0.075 | — | Gemini 3.8 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash: price per million tokens by provider (Input) | $1.50 | — | Gemini 3.5 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash: price per million tokens by provider (Output) | $9.00 | — | Gemini 3.5 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash: price per million tokens by provider (Cache read) | $0.15 | — | Gemini 3.5 Flash · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash Lite: price per million tokens by provider (Input) | $0.30 | — | Gemini 3.5 Flash Lite · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash Lite: price per million tokens by provider (Output) | $2.50 | — | Gemini 3.5 Flash Lite · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.5 Flash Lite: price per million tokens by provider (Cache read) | $0.030 | — | Gemini 3.5 Flash Lite · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.1 Pro Preview: price per million tokens by provider (Input) | $2.00 | — | Gemini 3.1 Pro Preview · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.1 Pro Preview: price per million tokens by provider (Output) | $12.00 | — | Gemini 3.1 Pro Preview · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Gemini 3.1 Pro Preview: price per million tokens by provider (Cache read) | $0.20 | — | Gemini 3.1 Pro Preview · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Llama 3.3 70B Instruct: price per million tokens by provider (Input) | $0.72 | — | Llama 3.3 70B Instruct · reported by OpenRouter’s public API, snapshot 2026-10-06 |
| Llama 3.3 70B Instruct: price per million tokens by provider (Output) | $0.72 | — | Llama 3.3 70B Instruct · reported by OpenRouter’s public API, snapshot 2026-10-06 |
Reported by OpenRouter’s public API, not measured by AgentMuted dots: the other providers on the same chartUSD per million tokens
Google Vertex AI in Inference provider index: 27 models, 52 providers: 37 values, first Claude Haiku 4.5: price per million tokens by provider (Input) $1.00.
Compare Google Vertex AI
Price rows state the gap and name no winner: a reported price has no interval.
Amazon Bedrock vs Google Vertex AI
21 ties · 2 unclear
vs Anthropic
Anthropic vs Google Vertex AI
21 ties
vs Azure
Google Vertex AI vs Azure
21 ties
Google Vertex AI vs Claude Platform on AWS
15 ties
Google AI Studio vs Google Vertex AI
12 ties
vs DeepInfra
Google Vertex AI vs DeepInfra
4 unclear
vs Groq
Google Vertex AI vs Groq
4 unclear
vs Novita AI
Google Vertex AI vs Novita AI
4 unclear
vs Parasail
Google Vertex AI vs Parasail
4 unclear
vs SambaNova
Google Vertex AI vs SambaNova
4 unclear
vs Together AI
Google Vertex AI vs Together AI
4 unclear
vs Baseten
Google Vertex AI vs Baseten
2 unclear
vs Cerebras
Google Vertex AI vs Cerebras
2 unclear
Google Vertex AI vs Cloudflare Workers AI
2 unclear
vs Nebius
Google Vertex AI vs Nebius
2 unclear
vs SiliconFlow
Google Vertex AI vs SiliconFlow
2 unclear
Next
Write-ups that use this data
AI coding agent best practices: 12 rules, each backed by a measurement
12 rules for running AI coding agents, each with one measured number: validation, model choice, effort, caching, memory, routing, CLIs and sample size.
Claude Fable 5.1 vs Opus 5.5 vs Sonnet 5.5: speed, tokens and price tested
24 of 24: Claude Fable 5.1, Opus 5.5 and Sonnet 5.5 each passed every hard task. Fable cost 3.3x Opus and 6.5x Sonnet per pass (list-price calculation).
LLM API pricing comparison, October 2026: Claude vs GPT vs Gemini per million tokens
21 LLM API prices per million tokens, October 2026: Claude, GPT, Gemini and more. Blended prices span 100x. Price per token is not price per task.