GPU EconomyCloud GPU prices, read every hour
updated
H100 index$3.5426 providersWidestV10015.9×Mover, 7dB300−1.4%Coverage709 quotes63 providersReads, 24 h17,43726 runs

Hugging Face Inference Endpoints vs Modal

As of , Modal lists the lower on-demand rate on 4 of the 7 GPUs both offer on-demand; Hugging Face Inference Endpoints wins on 3. On the priciest shared card both list on-demand, NVIDIA B200: Hugging Face Inference Endpoints $9.25 (aws) vs Modal $6.25 per GPU-hour on-demand — Modal is $3.00 (32%) cheaper.

Cloud GPU rental prices of Hugging Face Inference Endpoints and Modal, compared per GPU-hour on the GPUs both list. Where a provider has several tiers or regions, its cheapest on-demand rate is used and labelled. Community capacity (RunPod Community Cloud) and marketplace floors are shown only where a provider has nothing else, and never decide which is cheaper.

On-demand, per GPU-hour

Each price links to the pricing page it was read from and says when it was read.
GPUHugging Face Inference EndpointsModalDifferenceCheaper
NVIDIA B200$9.25
aws
captured
Deploy ↗
$6.25
GPU only
captured
Deploy ↗
$3.00 (32%)Modal
NVIDIA H200$5.00
aws
captured
Deploy ↗
$4.54
GPU only
captured
Deploy ↗
$0.46 (9%)Modal
NVIDIA RTX PRO 6000 Blackwell$2.75
aws
captured
Deploy ↗
$3.03
GPU only
captured
Deploy ↗
$0.28 (9%)Hugging Face Inference Endpoints
NVIDIA A100 80GB$2.50
aws
captured
Deploy ↗
$2.50
GPU only
captured
Deploy ↗
$0.00 (0%)Modal
NVIDIA L40S$1.80
aws
captured
Deploy ↗
$1.95
GPU only
captured
Deploy ↗
$0.15 (8%)Hugging Face Inference Endpoints
NVIDIA L4$0.80
aws
captured
Deploy ↗
$0.80
GPU only
captured
Deploy ↗
$0.00 (0%)Modal
NVIDIA T4$0.50
aws
captured
Deploy ↗
$0.59
GPU only
captured
Deploy ↗
$0.09 (15%)Hugging Face Inference Endpoints

GPU by GPU

Not compared: GPUs only one of the two lists, reserved/committed tiers, and anything either provider prices only on request. See each provider's page for its full list.

Data updated: