GPU cloud · 6 GPU models tracked
Baseten
$0.63–$9.98per GPU-hour, the cheapest row to the dearest
Dedicated model deployments billed by the minute and scaled down when idle, shown here as an hour; each price is a GPU instance with its CPU and memory.
As of , Baseten lists 6 GPU models we track, from $0.63 per GPU-hour (NVIDIA T4, on-demand) to $9.98 (NVIDIA B200, on-demand).
| Website | www.baseten.co/ |
|---|---|
| Pricing page | www.baseten.co/pricing/ |
| Regions and tiers | No region or tier named in the rates below: one global rate card |
| Models tracked | 6 |
Price by GPU
As of , per GPU-hour, cheapest tier first named:
- Baseten lists the NVIDIA B200 from $9.98 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 22, 2026 (7-day change: none). Against the market: 47% above the median rate card ($6.79 across 27 providers); 22 of the 27 providers we track list a lower on-demand rate, the lowest Deep Infra at $3.69 (Deep Infra’s rates, Baseten vs Deep Infra).
- Baseten lists the NVIDIA H100 80GB (form factor not stated) from $6.50 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 22, 2026 (7-day change: none). Against the market: 133% above the median rate card ($2.79 across 19 providers); 17 of the 19 providers we track list a lower on-demand rate, the lowest Gcore at $1.74 (Gcore’s rates, Baseten vs Gcore).
- Baseten lists the NVIDIA A100 80GB from $4.00 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 22, 2026 (7-day change: none). Against the market: 121% above the median rate card ($1.81 across 40 providers); 37 of the 40 providers we track list a lower on-demand rate, the lowest Deep Infra at $0.89 (Deep Infra’s rates, Baseten vs Deep Infra).
- Baseten lists the NVIDIA A10G from $1.21 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 27, 2026. Against the market: 9% above the median rate card ($1.11 across 2 providers); 1 of the 2 providers we track list a lower on-demand rate, the lowest Amazon Web Services at $1.01 (Amazon Web Services’s rates, Baseten vs Amazon Web Services).
- Baseten lists the NVIDIA L4 from $0.85 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 22, 2026 (7-day change: none). Against the market: 6% above the median rate card ($0.80 across 17 providers); 13 of the 17 providers we track list a lower on-demand rate, the lowest Seeweb at $0.43 (Seeweb’s rates, Baseten vs Seeweb).
- Baseten lists the NVIDIA T4 from $0.63 per GPU-hour on-demand. The on-demand rate is unchanged since first read, Sep 22, 2026 (7-day change: none). Against the market: 20% above the median rate card ($0.53 across 9 providers); 7 of the 9 providers we track list a lower on-demand rate, the lowest immers.cloud at $0.23 (immers.cloud’s rates, Baseten vs immers.cloud).
All listed prices
One table per price type, never mixed. Every rate normalized to one GPU for one hour.
On-demand
| GPU | USD / GPU-hour | 7-day change | Type | Region | Min. commitment | Source, captured | Deploy |
|---|---|---|---|---|---|---|---|
| NVIDIA T4 | $0.63 | no change | on-demand | — | — | Price page unchanged since first read, Sep 22, 2026 | Deploy ↗ |
| NVIDIA L4 | $0.85 | no change | on-demand | — | — | Price page unchanged since first read, Sep 22, 2026 | Deploy ↗ |
| NVIDIA A10G | $1.21 | — | on-demand | — | — | Price page unchanged since first read, Sep 27, 2026 | Deploy ↗ |
| NVIDIA A100 80GB | $4.00 | no change | on-demand | — | — | Price page unchanged since first read, Sep 22, 2026 | Deploy ↗ |
| NVIDIA H100 80GB (form factor not stated) | $6.50 | no change | on-demand | — | — | Price page unchanged since first read, Sep 22, 2026 | Deploy ↗ |
| NVIDIA B200 | $9.98 | no change | on-demand | — | — | Price page unchanged since first read, Sep 22, 2026 | Deploy ↗ |
Every rate on this page as data: /providers/baseten/prices.json — the same fields as prices.json, narrowed to this provider. CC BY 4.0, no key, no rate limit. On your own page: embed this table (/embed/providers/baseten, an iframe, no script).
Compare Baseten with
Every provider that lists at least two of the same GPUs on-demand, most shared GPUs first; each opens the side-by-side page.
- Amazon Web Services 5 shared GPUs
- Cerebrium 5 shared GPUs
- Hugging Face Inference Endpoints 5 shared GPUs
- E2E Networks 4 shared GPUs
- EmpirioLabs 4 shared GPUs
- Google Cloud 4 shared GPUs
- Koyeb 4 shared GPUs
- Modal 4 shared GPUs
- Deep Infra 3 shared GPUs
- immers.cloud 3 shared GPUs
- Lyceum 3 shared GPUs
- Northflank 3 shared GPUs
- Replicate 3 shared GPUs
- RunPod 3 shared GPUs
- Vast.ai 3 shared GPUs
- Beam 2 shared GPUs
- Cloud.ru Evolution 2 shared GPUs
- CoreWeave 2 shared GPUs
- Daytona 2 shared GPUs
- fal 2 shared GPUs
- Gcore 2 shared GPUs
- GMI Cloud 2 shared GPUs
- HexGrid 2 shared GPUs
- Hyperstack 2 shared GPUs
- Jarvislabs 2 shared GPUs
- Lambda 2 shared GPUs
- Leafcloud 2 shared GPUs
- Massed Compute 2 shared GPUs
- Microsoft Azure 2 shared GPUs
- Oracle Cloud (OCI) 2 shared GPUs
- OVHcloud 2 shared GPUs
- packet.ai 2 shared GPUs
- Seeweb 2 shared GPUs
- Sesterce 2 shared GPUs
- Verda 2 shared GPUs
Data updated: