6 measured metrics · 1 study

Fireworks AIvsBaseten

No row separates them: 5 ties, 1 unclear.

The verdict

Fireworks AI and Baseten share 6 reported list prices from one study. No row names a winner: a reported list price has no interval, so each gap is stated, not ranked. Prices change often; the snapshot date is in each row’s context. The rows are 5 ties and 1 unclear; each row says why.

Headline metrics

How far apart the two sides are on the headline metrics. Length is the ratio of the two values; it is not a winner.

Bar length is the ratio of the two values on a log scale, pointing to the larger one. Larger is not better for time, tokens or cost. A bar has a side’s color only when that side is ahead in the data; gray means the data does not separate them.

n is shown per side on every row

6 headline metrics as the ratio of the two values. None of them separates the sides in the data. Widest ratio: GLM 5.3: price per million tokens by provider (Cache read), 1.9x (Fireworks AI larger).

Metric by metric

Both values of a row come from the same chart in the same study. Each row has its own axis. The shaded band is where the two intervals or ranges overlap: a side is ahead only when they do not.

  • Fireworks AI
  • Baseten
  • where the two overlap

Inference provider index: 27 models, 52 providers

5 ties · 1 unclear
  • Kimi K3: price per million tokens by provider (Input): Fireworks AI $3.00; Baseten $3.00. Tie.
  • Kimi K3: price per million tokens by provider (Output): Fireworks AI $15.00; Baseten $15.00. Tie.
  • Kimi K3: price per million tokens by provider (Cache read): Fireworks AI $0.30; Baseten $0.30. Tie.
  • GLM 5.3: price per million tokens by provider (Input): Fireworks AI $1.40; Baseten $1.40. Tie.
  • GLM 5.3: price per million tokens by provider (Output): Fireworks AI $4.40; Baseten $4.40. Tie.
  • GLM 5.3: price per million tokens by provider (Cache read): Fireworks AI $0.26; Baseten $0.14. Unclear.

n is shown per side on every row

6 rows from 1 study. No row separates them: 5 ties, 1 unclear.

When to pick which

Only from the rows above. A tie is not a reason to pick either side.

When to pick Fireworks AI

No row in this data puts Fireworks AI ahead of Baseten. Pick on other grounds (price, access, the tasks you run), or measure your own workload.

When to pick Baseten

No row in this data puts Baseten ahead of Fireworks AI. Pick on other grounds (price, access, the tasks you run), or measure your own workload.

Side by side

The study charts, showing only these two. Open a study for every configuration.

Reported

Input

BaseTen (fp8)
Fireworks

Output

BaseTen (fp8)
Fireworks

Cache read

BaseTen (fp8)
Fireworks

One panel per series, all on the same axis.

Third-party reported values. 2 rows, 3 series: Input, Output, Cache read. Input: all at $3.00. Output: all at $15.00.

Notes

Standard tier, one bar per provider (its cheapest standard endpoint); reported by OpenRouter’s public API, snapshot 2026-10-06

Prices reported by OpenRouter’s public API, snapshot 2026-10-06; third-party-reported, not measured by Agent. Sorted from the lowest to the highest blended price (3 input : 1 output). A parenthesis names the quantization the provider reported; lower precision or a shorter context can explain a lower price, so check the endpoint table. Flex, priority, fast and regional endpoints are left out here because they are priced differently on purpose.

Source: OpenRouter public API: models and provider endpoints (snapshot)

Reported

Input

BaseTen (fp4)
Fireworks

Output

BaseTen (fp4)
Fireworks

Cache read

BaseTen (fp4)
Fireworks

One panel per series, all on the same axis.

Third-party reported values. 2 rows, 3 series: Input, Output, Cache read. Input: all at $1.40. Output: all at $4.40.

Notes

Standard tier, one bar per provider (its cheapest standard endpoint); reported by OpenRouter’s public API, snapshot 2026-10-06

Prices reported by OpenRouter’s public API, snapshot 2026-10-06; third-party-reported, not measured by Agent. Sorted from the lowest to the highest blended price (3 input : 1 output). A parenthesis names the quantization the provider reported; lower precision or a shorter context can explain a lower price, so check the endpoint table. Flex, priority, fast and regional endpoints are left out here because they are priced differently on purpose.

Source: OpenRouter public API: models and provider endpoints (snapshot)

How a row is called

  • AheadThe 95% intervals do not overlap, or the run ranges or p50–p95 bands do not overlap with at least 5 runs per side.
  • TieThe values match, or both sit at the same ceiling.
  • UnclearThe intervals or ranges overlap, too few runs were recorded, no interval was recorded, or more is not better (token counts are never a win).
  • CalculationDerived from list prices and recorded counts. Not a bill and not a run.

The dataset compiler makes every call; this page only draws it. No composite score, no rank.

Questions

Which is better, Fireworks AI or Baseten?
Fireworks AI and Baseten share 6 reported list prices from one study. No row names a winner: a reported list price has no interval, so each gap is stated, not ranked. Prices change often; the snapshot date is in each row’s context. The rows are 5 ties and 1 unclear; each row says why.
How were Fireworks AI and Baseten measured?
They share 6 measured metrics from 1 public study: Inference provider index: 27 models, 52 providers. Every row names its configuration, its sample size and its interval or range.
How do Fireworks AI and Baseten compare on kimi K3: price per million tokens by provider (Input)?
Fireworks AI: $3.00 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Baseten: $3.00 (fp8 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Same reported list price.
How do Fireworks AI and Baseten compare on kimi K3: price per million tokens by provider (Output)?
Fireworks AI: $15.00 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Baseten: $15.00 (fp8 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Same reported list price.
How do Fireworks AI and Baseten compare on kimi K3: price per million tokens by provider (Cache read)?
Fireworks AI: $0.30 (Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Baseten: $0.30 (fp8 · Kimi K3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Same reported list price.
How do Fireworks AI and Baseten compare on gLM 5.3: price per million tokens by provider (Input)?
Fireworks AI: $1.40 (GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Baseten: $1.40 (fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Same reported list price.
How do Fireworks AI and Baseten compare on gLM 5.3: price per million tokens by provider (Output)?
Fireworks AI: $4.40 (GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Baseten: $4.40 (fp4 · GLM 5.3 · reported by OpenRouter’s public API, snapshot 2026-10-06). Same reported list price.

The studies behind this page

  • Inference
  • Providers

Inference provider index: 27 models, 52 providers

Price per million tokens for 27 models across 52 providers, the spread between them and OpenRouter’s markup over first-party prices.

265endpoints, 52 providers, 27 models · Provider endpoints in the snapshot · n = 27

34 chartsUpdated October 6, 2026

All comparisons

Turn the numbers into shipped work.

Agent runs these choices for you: a persistent AI worker with memory and rules, on your Claude and Codex subscriptions.