GPU cloud · 5 GPU models tracked
Fireworks AI
$8.00–$20.00per GPU-hour, the cheapest row to the dearest
Billed per GPU second on an on-demand deployment, with no charge for start-up time. The figure is one GPU: the page prints no vCPU, RAM or storage beside the card. Region-restricted deployments cost 1.5x on top, which this site does not publish because the page prints no such rate.
As of , Fireworks AI lists 5 GPU models we track, from $8.00 per GPU-hour (NVIDIA H100 80GB (form factor not stated), on-demand, GPU only) to $20.00 (NVIDIA GB300 NVL72, on-demand, GPU only).
| Website | fireworks.ai/ |
|---|---|
| Pricing page | fireworks.ai/pricing |
| Regions and tiers | No region or tier named in the rates below: one global rate card |
| Models tracked | 5 |
Price by GPU
As of , per GPU-hour: Fireworks AI’s on-demand rate for each card it lists, a cheaper market it also sells, and where the on-demand rate stands among the providers we track.
| GPU | Fireworks AI on-demand, per GPU-hour | Against the market (median rate card) |
|---|---|---|
| NVIDIA GB300 NVL72 | $20.00 GPU only captured | 11% above the median $18.00 · 3 providers 2 of 3 list lower · lowest Verda $10.32 · compare |
| NVIDIA B300 | $15.00 GPU only captured | 88% above the median $7.99 · 21 providers 17 of 21 list lower · lowest Deep Infra $4.89 · compare |
| NVIDIA B200 | $13.00 GPU only captured | 89% above the median $6.89 · 32 providers 28 of 32 list lower · lowest Deep Infra $3.69 · compare |
| NVIDIA H100 80GB (form factor not stated) | $8.00 GPU only captured | 131% above the median $3.46 · 20 providers 18 of 20 list lower · lowest Gcore $1.71 · compare |
| NVIDIA H200 | $8.00 GPU only captured | 76% above the median $4.54 · 42 providers 38 of 42 list lower · lowest Beam $2.09 from · compare |
The same, in words
- Fireworks AI lists the NVIDIA GB300 NVL72 from $20.00 per GPU-hour on-demand (GPU only). The on-demand rate is unchanged since first read, Oct 2, 2026. Against the market: 11% above the median rate card ($18.00 across 3 providers); 2 of the 3 providers we track list a lower on-demand rate, the lowest Verda at $10.32 (Verda’s rates, Fireworks AI vs Verda).
- Fireworks AI lists the NVIDIA B300 from $15.00 per GPU-hour on-demand (GPU only). The on-demand rate is unchanged since first read, Oct 2, 2026. Against the market: 88% above the median rate card ($7.99 across 21 providers); 17 of the 21 providers we track list a lower on-demand rate, the lowest Deep Infra at $4.89 (Deep Infra’s rates, Fireworks AI vs Deep Infra).
- Fireworks AI lists the NVIDIA B200 from $13.00 per GPU-hour on-demand (GPU only). The on-demand rate is unchanged since first read, Oct 2, 2026. Against the market: 89% above the median rate card ($6.89 across 32 providers); 28 of the 32 providers we track list a lower on-demand rate, the lowest Deep Infra at $3.69 (Deep Infra’s rates, Fireworks AI vs Deep Infra).
- Fireworks AI lists the NVIDIA H100 80GB (form factor not stated) from $8.00 per GPU-hour on-demand (GPU only). The on-demand rate is unchanged since first read, Oct 2, 2026. Against the market: 131% above the median rate card ($3.46 across 20 providers); 18 of the 20 providers we track list a lower on-demand rate, the lowest Gcore at $1.71 (Gcore’s rates, Fireworks AI vs Gcore).
- Fireworks AI lists the NVIDIA H200 from $8.00 per GPU-hour on-demand (GPU only). The on-demand rate is unchanged since first read, Oct 2, 2026. Against the market: 76% above the median rate card ($4.54 across 42 providers); 38 of the 42 providers we track list a lower on-demand rate, the lowest Beam at $2.09 (Beam’s rates, Fireworks AI vs Beam).
All listed prices
One table per price type, never mixed. Every rate normalized to one GPU for one hour.
On-demand
| GPU | USD / GPU-hour | 7-day change | Market | Region | Min. commitment | Source, captured | Deploy |
|---|---|---|---|---|---|---|---|
| NVIDIA H100 80GB (form factor not stated) | $8.00 | — | on-demand GPU only | — | — | Price page unchanged since first read, Oct 2, 2026 | Deploy ↗ |
| NVIDIA H200 | $8.00 | — | on-demand GPU only | — | — | Price page unchanged since first read, Oct 2, 2026 | Deploy ↗ |
| NVIDIA B200 | $13.00 | — | on-demand GPU only | — | — | Price page unchanged since first read, Oct 2, 2026 | Deploy ↗ |
| NVIDIA B300 | $15.00 | — | on-demand GPU only | — | — | Price page unchanged since first read, Oct 2, 2026 | Deploy ↗ |
| NVIDIA GB300 NVL72 | $20.00 | — | on-demand GPU only | — | — | Price page unchanged since first read, Oct 2, 2026 | Deploy ↗ |
Every rate on this page as data: /providers/fireworks/prices.json — the same fields as prices.json, narrowed to this provider. CC BY 4.0, no key, no rate limit. On your own page: embed this table (/embed/providers/fireworks, an iframe, no script).
Compare Fireworks AI with
Every provider that lists at least two of the same GPUs on-demand, most shared GPUs first; each opens the side-by-side page.
- Daytona 4 shared GPUs
- Deep Infra 4 shared GPUs
- EmpirioLabs 4 shared GPUs
- fal 4 shared GPUs
- Lyceum 4 shared GPUs
- Oracle Cloud (OCI) 4 shared GPUs
- Sesterce 4 shared GPUs
- Verda 4 shared GPUs
- Amazon Web Services 3 shared GPUs
- Cerebrium 3 shared GPUs
- E2E Networks 3 shared GPUs
- Gcore 3 shared GPUs
- GMI Cloud 3 shared GPUs
- GPU.ai 3 shared GPUs
- Hugging Face 3 shared GPUs
- Hyperstack 3 shared GPUs
- Koyeb 3 shared GPUs
- Modal 3 shared GPUs
- Nebius 3 shared GPUs
- RunPod 3 shared GPUs
- Together AI 3 shared GPUs
- Baseten 2 shared GPUs
- Beam 2 shared GPUs
- Civo 2 shared GPUs
- CoreWeave 2 shared GPUs
- Dataoorts 2 shared GPUs
- Edgevana 2 shared GPUs
- Google Cloud 2 shared GPUs
- HexGrid 2 shared GPUs
- Hyperbolic 2 shared GPUs
- immers.cloud 2 shared GPUs
- Paperspace 2 shared GPUs
- Vast.ai 2 shared GPUs
Data updated: