GPU EconomyCloud GPU prices, read every hour
updated
H100 index$3.5426 providersWidestV10015.9×Mover, 7dB300−1.4%Coverage709 quotes63 providersReads, 24 h17,43726 runs

fal vs Hugging Face Inference Endpoints

As of , fal lists the lower on-demand rate on 2 of the 4 GPUs both offer on-demand; Hugging Face Inference Endpoints wins on 1. On the priciest shared card both list on-demand, NVIDIA B200: fal $6.25 vs Hugging Face Inference Endpoints $9.25 (aws) per GPU-hour on-demand — fal is $3.00 (32%) cheaper.

Cloud GPU rental prices of fal and Hugging Face Inference Endpoints, compared per GPU-hour on the GPUs both list. Where a provider has several tiers or regions, its cheapest on-demand rate is used and labelled. Community capacity (RunPod Community Cloud) and marketplace floors are shown only where a provider has nothing else, and never decide which is cheaper.

On-demand, per GPU-hour

Each price links to the pricing page it was read from and says when it was read.
GPUfalHugging Face Inference EndpointsDifferenceCheaper
NVIDIA B200$6.25
captured
Deploy ↗
$9.25
aws
captured
Deploy ↗
$3.00 (32%)fal
NVIDIA H100 80GB (form factor not stated)$4.50
captured
Deploy ↗
$4.50
aws
captured
Deploy ↗
—Same
NVIDIA H200$4.50
captured
Deploy ↗
$5.00
aws
captured
Deploy ↗
$0.50 (10%)fal
NVIDIA RTX PRO 6000 Blackwell$2.99
captured
Deploy ↗
$2.75
aws
captured
Deploy ↗
$0.24 (8%)Hugging Face Inference Endpoints

GPU by GPU

Not compared: GPUs only one of the two lists, reserved/committed tiers, and anything either provider prices only on request. See each provider's page for its full list.

Data updated: