GPU EconomyCloud GPU prices, read every hour
updated
H100 index$3.5426 providersWidestV10015.9×Mover, 7dB300−1.4%Coverage709 quotes63 providersReads, 24 h17,43726 runs

Microsoft Azure vs Hugging Face Inference Endpoints

As of , Hugging Face Inference Endpoints lists the lower on-demand rate on 4 of the 4 GPUs both offer on-demand. On the priciest shared card both list on-demand, NVIDIA H200: Microsoft Azure $10.60 (westus2) vs Hugging Face Inference Endpoints $5.00 (aws) per GPU-hour on-demand — Hugging Face Inference Endpoints is $5.60 (53%) cheaper.

Cloud GPU rental prices of Microsoft Azure and Hugging Face Inference Endpoints, compared per GPU-hour on the GPUs both list. Where a provider has several tiers or regions, its cheapest on-demand rate is used and labelled. Community capacity (RunPod Community Cloud) and marketplace floors are shown only where a provider has nothing else, and never decide which is cheaper.

On-demand, per GPU-hour

Each price links to the pricing page it was read from and says when it was read.
GPUMicrosoft AzureHugging Face Inference EndpointsDifferenceCheaper
NVIDIA H200$10.60
westus2
captured
Deploy ↗
$5.00
aws
captured
Deploy ↗
$5.60 (53%)Hugging Face Inference Endpoints
NVIDIA RTX PRO 6000 Blackwell$5.50
eastus2
captured
Deploy ↗
$2.75
aws
captured
Deploy ↗
$2.75 (50%)Hugging Face Inference Endpoints
NVIDIA A100 80GB$3.67
eastus2
captured
Deploy ↗
$2.50
aws
captured
Deploy ↗
$1.17 (32%)Hugging Face Inference Endpoints
NVIDIA T4$0.53
northcentralus
captured
Deploy ↗
$0.50
aws
captured
Deploy ↗
$0.03 (5%)Hugging Face Inference Endpoints

Spot / interruptible, per GPU-hour

GPUMicrosoft AzureHugging Face Inference Endpoints
NVIDIA RTX PRO 6000 Blackwell$1.02
captured
—
NVIDIA A100 80GB$0.68
captured
—
NVIDIA T4$0.07
captured
—

GPU by GPU

Not compared: GPUs only one of the two lists, reserved/committed tiers, and anything either provider prices only on request. See each provider's page for its full list.

Data updated: