GPU EconomyCloud GPU prices, read every hour
updated
H100 index$3.5426 providersWidestV10015.9×Mover, 7dB300−1.4%Coverage709 quotes63 providersReads, 24 h17,34826 runs

GPU cloud · 8 GPU models tracked

Hugging Face Inference Endpoints

$0.50–$9.25per GPU-hour, the cheapest row to the dearest

A dedicated inference endpoint on the cloud named in the region column, billed by the minute at the hourly rate while it runs, CPU and memory included. It serves a model; it is not a machine you log into.

As of , Hugging Face Inference Endpoints lists 8 GPU models we track, from $0.50 per GPU-hour (NVIDIA T4, on-demand) to $9.25 (NVIDIA B200, on-demand).

Every rate ↓ · Compare with 45 providers ↓

  • Cheapest row$0.50NVIDIA T4 · on-demand
  • H200$5.00cheapest tier listed
  • B200$9.25cheapest tier listed
  • A100 80GB$2.50cheapest tier listed
Websitehuggingface.co/
Pricing pagehuggingface.co/pricing
Regions and tiersaws — as labelled in the rates below
Models tracked8

Price by GPU

As of , per GPU-hour, cheapest tier first named:

All listed prices

One table per price type, never mixed. Every rate normalized to one GPU for one hour.

On-demand

Listed pay-as-you-go rates at Hugging Face Inference Endpoints.
GPUUSD / GPU-hour7-day changeTypeRegionMin. commitmentSource, capturedDeploy
NVIDIA T4$0.50—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA L4$0.80—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA L40S$1.80—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA A100 80GB$2.50—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA RTX PRO 6000 Blackwell$2.75—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA H100 80GB (form factor not stated)$4.50—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA H200$5.00—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗
NVIDIA B200$9.25—on-demandaws—Price page

unchanged since first read, Sep 25, 2026
Deploy ↗

Every rate on this page as data: /providers/huggingface/prices.json — the same fields as prices.json, narrowed to this provider. CC BY 4.0, no key, no rate limit. On your own page: embed this table (/embed/providers/huggingface, an iframe, no script).

Compare Hugging Face Inference Endpoints with

Every provider that lists at least two of the same GPUs on-demand, most shared GPUs first; each opens the side-by-side page.

Data updated: