6 measured metrics · 1 study
Fireworks AIvsDeepInfra
No row separates them: 6 unclear.
The verdict
Fireworks AI and DeepInfra share 6 reported list prices from one study. No row names a winner: a reported list price has no interval, so each gap is stated, not ranked. Prices change often; the snapshot date is in each row’s context. The rows are 6 unclear; each row says why.
Headline metrics
How far apart the two sides are on the headline metrics. Length is the ratio of the two values; it is not a winner.
Bar length is the ratio of the two values on a log scale, pointing to the larger one. Larger is not better for time, tokens or cost. A bar has a side’s color only when that side is ahead in the data; gray means the data does not separate them.
| Metric | Fireworks AI | DeepInfra | n | Interval or range | Outcome | Basis | Study |
|---|---|---|---|---|---|---|---|
| Kimi K3: price per million tokens by provider (Input) | $3.00Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $2.85mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($3.00 vs $2.85, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| Kimi K3: price per million tokens by provider (Output) | $15.00Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $14.25mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($15.00 vs $14.25, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| Kimi K3: price per million tokens by provider (Cache read) | $0.30Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.28mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($0.30 vs $0.28, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Input) | $1.40GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.56fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($1.40 vs $0.56, 2.5x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Output) | $4.40GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $2.50fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($4.40 vs $2.50, 1.8x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Cache read) | $0.26GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.13fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($0.26 vs $0.13, 2.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
n is shown per side on every row
6 headline metrics as the ratio of the two values. None of them separates the sides in the data. Widest ratio: GLM 5.3: price per million tokens by provider (Input), 2.5x (Fireworks AI larger).
Metric by metric
Both values of a row come from the same chart in the same study. Each row has its own axis. The shaded band is where the two intervals or ranges overlap: a side is ahead only when they do not.
- Fireworks AI
- DeepInfra
- where the two overlap
Inference provider index: 27 models, 52 providers
- Kimi K3: price per million tokens by provider (Input)$3.00$2.85UnclearKimi K3: price per million tokens by provider (Input): Fireworks AI $3.00; DeepInfra $2.85. Unclear.
- Kimi K3: price per million tokens by provider (Output)$15.00$14.25UnclearKimi K3: price per million tokens by provider (Output): Fireworks AI $15.00; DeepInfra $14.25. Unclear.
- Kimi K3: price per million tokens by provider (Cache read)$0.30$0.28UnclearKimi K3: price per million tokens by provider (Cache read): Fireworks AI $0.30; DeepInfra $0.28. Unclear.
- GLM 5.3: price per million tokens by provider (Input)$1.40$0.56UnclearGLM 5.3: price per million tokens by provider (Input): Fireworks AI $1.40; DeepInfra $0.56. Unclear.
- GLM 5.3: price per million tokens by provider (Output)$4.40$2.50UnclearGLM 5.3: price per million tokens by provider (Output): Fireworks AI $4.40; DeepInfra $2.50. Unclear.
- GLM 5.3: price per million tokens by provider (Cache read)$0.26$0.13UnclearGLM 5.3: price per million tokens by provider (Cache read): Fireworks AI $0.26; DeepInfra $0.13. Unclear.
| Metric | Fireworks AI | DeepInfra | n | Interval or range | Outcome | Basis | Study |
|---|---|---|---|---|---|---|---|
| Kimi K3: price per million tokens by provider (Input) | $3.00Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $2.85mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($3.00 vs $2.85, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| Kimi K3: price per million tokens by provider (Output) | $15.00Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $14.25mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($15.00 vs $14.25, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| Kimi K3: price per million tokens by provider (Cache read) | $0.30Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.28mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($0.30 vs $0.28, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Input) | $1.40GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.56fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($1.40 vs $0.56, 2.5x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Output) | $4.40GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $2.50fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($4.40 vs $2.50, 1.8x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
| GLM 5.3: price per million tokens by provider (Cache read) | $0.26GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | $0.13fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06 | — | none recorded | Unclear | A reported list price has no interval, so the gap ($0.26 vs $0.13, 2.1x) is stated, not ranked. Check quantization and context before treating the two as equal products. | Inference provider index: 27 models, 52 providers |
n is shown per side on every row
6 rows from 1 study. No row separates them: 6 unclear.
When to pick which
Only from the rows above. A tie is not a reason to pick either side.
When to pick Fireworks AI
No row in this data puts Fireworks AI ahead of DeepInfra. Pick on other grounds (price, access, the tasks you run), or measure your own workload.
When to pick DeepInfra
No row in this data puts DeepInfra ahead of Fireworks AI. Pick on other grounds (price, access, the tasks you run), or measure your own workload.
Side by side
The study charts, showing only these two. Open a study for every configuration.
Input
Output
Cache read
One panel per series, all on the same axis.
| Item | Input | Output | Cache read |
|---|---|---|---|
| DeepInfra (mxfp4) | $2.85 | $14.25 | $0.28 |
| Fireworks | $3.00 | $15.00 | $0.3 |
Third-party reported values. 2 rows, 3 series: Input, Output, Cache read. Input: highest Fireworks $3.00. Lowest DeepInfra (mxfp4) $2.85. Output: highest Fireworks $15.00. Lowest DeepInfra (mxfp4) $14.25.
Notes
Standard tier, one bar per provider (its cheapest standard endpoint); reported by OpenRouter’s public API, snapshot 2026-10-06
Prices reported by OpenRouter’s public API, snapshot 2026-10-06; third-party-reported, not measured by Agent. Sorted from the lowest to the highest blended price (3 input : 1 output). A parenthesis names the quantization the provider reported; lower precision or a shorter context can explain a lower price, so check the endpoint table. Flex, priority, fast and regional endpoints are left out here because they are priced differently on purpose.
Source: OpenRouter public API: models and provider endpoints (snapshot)
Input
Output
Cache read
One panel per series, all on the same axis.
| Item | Input | Output | Cache read |
|---|---|---|---|
| DeepInfra (fp4) | $0.56 | $2.50 | $0.13 |
| Fireworks | $1.40 | $4.40 | $0.26 |
Third-party reported values. 2 rows, 3 series: Input, Output, Cache read. Input: highest Fireworks $1.40. Lowest DeepInfra (fp4) $0.56. Output: highest Fireworks $4.40. Lowest DeepInfra (fp4) $2.50.
Notes
Standard tier, one bar per provider (its cheapest standard endpoint); reported by OpenRouter’s public API, snapshot 2026-10-06
Prices reported by OpenRouter’s public API, snapshot 2026-10-06; third-party-reported, not measured by Agent. Sorted from the lowest to the highest blended price (3 input : 1 output). A parenthesis names the quantization the provider reported; lower precision or a shorter context can explain a lower price, so check the endpoint table. Flex, priority, fast and regional endpoints are left out here because they are priced differently on purpose.
Source: OpenRouter public API: models and provider endpoints (snapshot)
How a row is called
- AheadThe 95% intervals do not overlap, or the run ranges or p50–p95 bands do not overlap with at least 5 runs per side.
- TieThe values match, or both sit at the same ceiling.
- UnclearThe intervals or ranges overlap, too few runs were recorded, no interval was recorded, or more is not better (token counts are never a win).
- CalculationDerived from list prices and recorded counts. Not a bill and not a run.
The dataset compiler makes every call; this page only draws it. No composite score, no rank.
Questions
- Which is better, Fireworks AI or DeepInfra?
- Fireworks AI and DeepInfra share 6 reported list prices from one study. No row names a winner: a reported list price has no interval, so each gap is stated, not ranked. Prices change often; the snapshot date is in each row’s context. The rows are 6 unclear; each row says why.
- How were Fireworks AI and DeepInfra measured?
- They share 6 measured metrics from 1 public study: Inference provider index: 27 models, 52 providers. Every row names its configuration, its sample size and its interval or range.
- How do Fireworks AI and DeepInfra compare on kimi K3: price per million tokens by provider (Input)?
- Fireworks AI: $3.00 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). DeepInfra: $2.85 (mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). A reported list price has no interval, so the gap ($3.00 vs $2.85, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products.
- How do Fireworks AI and DeepInfra compare on kimi K3: price per million tokens by provider (Output)?
- Fireworks AI: $15.00 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). DeepInfra: $14.25 (mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). A reported list price has no interval, so the gap ($15.00 vs $14.25, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products.
- How do Fireworks AI and DeepInfra compare on kimi K3: price per million tokens by provider (Cache read)?
- Fireworks AI: $0.30 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). DeepInfra: $0.28 (mxfp4 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). A reported list price has no interval, so the gap ($0.30 vs $0.28, 1.1x) is stated, not ranked. Check quantization and context before treating the two as equal products.
- How do Fireworks AI and DeepInfra compare on gLM 5.3: price per million tokens by provider (Input)?
- Fireworks AI: $1.40 (GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). DeepInfra: $0.56 (fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). A reported list price has no interval, so the gap ($1.40 vs $0.56, 2.5x) is stated, not ranked. Check quantization and context before treating the two as equal products.
- How do Fireworks AI and DeepInfra compare on gLM 5.3: price per million tokens by provider (Output)?
- Fireworks AI: $4.40 (GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). DeepInfra: $2.50 (fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). A reported list price has no interval, so the gap ($4.40 vs $2.50, 1.8x) is stated, not ranked. Check quantization and context before treating the two as equal products.
The studies behind this page
Inference provider index: 27 models, 52 providers
Price per million tokens for 27 models across 52 providers, the spread between them and OpenRouter’s markup over first-party prices.