GPU Pricing

55 guides in this topic

This is the pricing cluster: what a GPU actually costs to rent, broken down provider by provider and updated as rates move. Each post pulls a named provider's published on-demand and spot rates (AWS P5, Azure ND-series, GCP A3/A4, CoreWeave, Nebius, Crusoe, and a couple dozen more) and lines them up against Spheron's live pricing so the comparison isn't hypothetical.

A second thread runs through here too: API pricing versus self-hosting. When a frontier model cuts its per-token rate, that changes the token volume where renting your own GPU cluster starts winning, and these posts do that math with the actual API prices, not estimates.

The pillar post, GPU Cloud Pricing Comparison 2026, is the widest lens: H100 and B200 rates across 15+ providers in one table, refreshed regularly. Start there if you're shopping broadly. Drop into an individual provider post (Azure H100, Lambda Cloud, Runpod, Vast.ai) if you already have a quote in hand and want to know whether it's actually competitive. Every number here traces back to a provider's own rate card or Spheron's live pricing API, and we say when a figure is spot instead of on-demand.

Start Here

All GPU Pricing Guides

Baseten Pricing 2026: Model Serving Cost vs Renting GPUs
Comparison

Baseten Pricing 2026: Model Serving Cost vs Renting GPUs

Jul 30, 2026
Genesis Cloud Pricing 2026: EU H100 Cost vs Spheron
Comparison

Genesis Cloud Pricing 2026: EU H100 Cost vs Spheron

Jul 30, 2026
Replicate Pricing 2026: Per-Second Cost vs Renting a GPU
Comparison

Replicate Pricing 2026: Per-Second Cost vs Renting a GPU

Jul 30, 2026
AMD MI325X Pricing 2026: Availability and Cost vs Spheron
Comparison

AMD MI325X Pricing 2026: Availability and Cost vs Spheron

Jul 28, 2026
Meta Muse Spark API Pricing vs Self-Hosted LLMs (2026)
Comparison

Meta Muse Spark API Pricing vs Self-Hosted LLMs (2026)

Jul 28, 2026
OVHcloud GPU Pricing 2026: H100 Cost and Sovereign AI
Comparison

OVHcloud GPU Pricing 2026: H100 Cost and Sovereign AI

Jul 27, 2026
GPU Export Controls 2026: What It Means for Cloud Pricing
Research

GPU Export Controls 2026: What It Means for Cloud Pricing

Jul 26, 2026
Mistral API Pricing vs Self-Hosted LLMs: Cost and Privacy in 2026
Comparison

Mistral API Pricing vs Self-Hosted LLMs: Cost and Privacy in 2026

Jul 23, 2026
TensorDock Pricing 2026: Cheapest H100 Rental Cost vs Spheron
Comparison

TensorDock Pricing 2026: Cheapest H100 Rental Cost vs Spheron

Jul 23, 2026
AWS Bedrock Pricing vs Self-Hosted LLMs (2026)
Comparison

AWS Bedrock Pricing vs Self-Hosted LLMs (2026)

Jul 22, 2026
Groq API Pricing 2026: Cost Per Token vs GPU Rental
Comparison

Groq API Pricing 2026: Cost Per Token vs GPU Rental

Jul 22, 2026
NVIDIA H100 Price 2026: Buy vs Rent Costs Compared
Comparison

NVIDIA H100 Price 2026: Buy vs Rent Costs Compared

Jul 20, 2026
Scaleway H100 Pricing 2026: Sovereign AI GPU Cost vs Spheron
Comparison

Scaleway H100 Pricing 2026: Sovereign AI GPU Cost vs Spheron

Jul 20, 2026
H200 NVL vs SXM5: Form Factor Decision Guide (2026)
Comparison

H200 NVL vs SXM5: Form Factor Decision Guide (2026)

Jul 19, 2026
NVIDIA DGX Cloud Lepton Pricing 2026: Marketplace Cost vs Spheron
Comparison

NVIDIA DGX Cloud Lepton Pricing 2026: Marketplace Cost vs Spheron

Jul 18, 2026
Fireworks AI Pricing 2026: Inference API Cost vs GPU Rental
Comparison

Fireworks AI Pricing 2026: Inference API Cost vs GPU Rental

Jul 17, 2026
Grok 4.5 API Pricing vs Self-Hosted LLMs (2026)
Comparison

Grok 4.5 API Pricing vs Self-Hosted LLMs (2026)

Jul 16, 2026
AWS Capacity Blocks Pricing 2026: B200/B300 Price Hike vs Spheron
Comparison

AWS Capacity Blocks Pricing 2026: B200/B300 Price Hike vs Spheron

Jul 15, 2026
Azure MI300X Pricing 2026: ND MI300X v5 Cost vs Spheron
Comparison

Azure MI300X Pricing 2026: ND MI300X v5 Cost vs Spheron

Jul 15, 2026
NVIDIA L4 vs L40S: Cheapest GPU for AI Inference (2026)
Comparison

NVIDIA L4 vs L40S: Cheapest GPU for AI Inference (2026)

Jul 13, 2026
GB300 NVL72 vs GB200 NVL72: Pricing & Availability (2026)
Comparison

GB300 NVL72 vs GB200 NVL72: Pricing & Availability (2026)

Jul 12, 2026
GMI Cloud Pricing 2026: H100 and B200 Cost vs Spheron
Comparison

GMI Cloud Pricing 2026: H100 and B200 Cost vs Spheron

Jul 12, 2026
Crusoe Cloud Pricing 2026: H100 and B200 Cost Per Hour vs Spheron
Comparison

Crusoe Cloud Pricing 2026: H100 and B200 Cost Per Hour vs Spheron

Jul 11, 2026
DeepSeek API Pricing vs Self-Hosted LLMs: Cost and Privacy (2026)
Comparison

DeepSeek API Pricing vs Self-Hosted LLMs: Cost and Privacy (2026)

Jul 11, 2026
AMD MI300X and MI355X Pricing 2026: Where to Rent Cheapest
Comparison

AMD MI300X and MI355X Pricing 2026: Where to Rent Cheapest

Jul 8, 2026
Google TPU v7 Ironwood vs NVIDIA B200: Inference Cost (2026)
Comparison

Google TPU v7 Ironwood vs NVIDIA B200: Inference Cost (2026)

Jul 7, 2026
Azure H200 Pricing 2026: ND H200 v5 Cost vs Spheron
Comparison

Azure H200 Pricing 2026: ND H200 v5 Cost vs Spheron

Jul 6, 2026
Claude Opus 4.8 API Pricing vs Self-Hosted LLMs (2026)
Comparison

Claude Opus 4.8 API Pricing vs Self-Hosted LLMs (2026)

Jul 6, 2026
Google Cloud A4 B200 Pricing 2026: GCP Cost vs Spheron
Comparison

Google Cloud A4 B200 Pricing 2026: GCP Cost vs Spheron

Jul 6, 2026
AWS EC2 G7 Pricing 2026: RTX PRO 4500 Blackwell Per-Hour Cost vs Spheron
Comparison

AWS EC2 G7 Pricing 2026: RTX PRO 4500 Blackwell Per-Hour Cost vs Spheron

Jun 29, 2026
NVIDIA B200 Cloud Pricing 2026: Per-Hour Rental Cost Across Providers vs Spheron
Comparison

NVIDIA B200 Cloud Pricing 2026: Per-Hour Rental Cost Across Providers vs Spheron

Jun 20, 2026
NVIDIA H100 News 2026: Price Changes, H200 and Successor Updates, and Cloud Availability
Research

NVIDIA H100 News 2026: Price Changes, H200 and Successor Updates, and Cloud Availability

Jun 14, 2026
GPU Cloud News June 2026: Latest Hardware Launches, Pricing Changes, and Availability Updates
Engineering

GPU Cloud News June 2026: Latest Hardware Launches, Pricing Changes, and Availability Updates

Jun 13, 2026
NVIDIA Vera Rubin NVL72 GPU Cloud: Availability, Cost Per Token, and Planning Your Rubin Rental in H2 2026
Research

NVIDIA Vera Rubin NVL72 GPU Cloud: Availability, Cost Per Token, and Planning Your Rubin Rental in H2 2026

Jun 12, 2026
NVIDIA B300 vs B200 for AI Inference: Is Blackwell Ultra Worth the Premium? (2026 Cost-Per-Token Guide)
Comparison

NVIDIA B300 vs B200 for AI Inference: Is Blackwell Ultra Worth the Premium? (2026 Cost-Per-Token Guide)

Jun 10, 2026
Huawei Ascend 950 vs NVIDIA B300 and B200 for LLM Inference: TPS, Memory Bandwidth, and Cost Comparison (2026)
Comparison

Huawei Ascend 950 vs NVIDIA B300 and B200 for LLM Inference: TPS, Memory Bandwidth, and Cost Comparison (2026)

Jun 10, 2026
SambaNova SN40L vs NVIDIA H200 and B200 on GPU Cloud: RDU Inference Benchmarks, Pricing, and Migration Guide (2026)
Comparison

SambaNova SN40L vs NVIDIA H200 and B200 on GPU Cloud: RDU Inference Benchmarks, Pricing, and Migration Guide (2026)

May 30, 2026
L40S vs H100 for AI Inference: When the L40S Wins on Cost Per Token (2026 Decision Guide)
Comparison

L40S vs H100 for AI Inference: When the L40S Wins on Cost Per Token (2026 Decision Guide)

May 26, 2026
What Is a GPU Cloud? Definition, How It Works, and When to Use One (2026)
Research

What Is a GPU Cloud? Definition, How It Works, and When to Use One (2026)

May 24, 2026
Google Cloud A3 H100 Pricing 2026: GCP Per-Hour Cost vs Spheron Breakdown
Comparison

Google Cloud A3 H100 Pricing 2026: GCP Per-Hour Cost vs Spheron Breakdown

May 22, 2026
Azure H100 Pricing 2026: ND H100 v5 Per-Hour Cost Compared to Spheron
Comparison

Azure H100 Pricing 2026: ND H100 v5 Per-Hour Cost Compared to Spheron

May 21, 2026
NVIDIA Vera Rubin NVL72 (H300): Specs, Pricing & Release Date (2026)
Engineering

NVIDIA Vera Rubin NVL72 (H300): Specs, Pricing & Release Date (2026)

May 15, 2026
DeepSeek V3.2 vs Llama 4 vs Qwen 3 (2026): Cost-per-Token from $0.04 to $0.31, Which to Pick
Comparison

DeepSeek V3.2 vs Llama 4 vs Qwen 3 (2026): Cost-per-Token from $0.04 to $0.31, Which to Pick

May 14, 2026
Intel Gaudi 3 vs NVIDIA H200 and B200: LLM Inference Benchmarks, Pricing, and Migration Guide (2026)
Comparison

Intel Gaudi 3 vs NVIDIA H200 and B200: LLM Inference Benchmarks, Pricing, and Migration Guide (2026)

May 13, 2026
Cerebras vs NVIDIA H100: Wafer-Scale vs GPU for LLM Inference (2026 Decision Guide)
Comparison

Cerebras vs NVIDIA H100: Wafer-Scale vs GPU for LLM Inference (2026 Decision Guide)

Apr 28, 2026
NVIDIA B300 (Blackwell Ultra): 288GB Specs, Pricing & Benchmarks (2026)
Engineering

NVIDIA B300 (Blackwell Ultra): 288GB Specs, Pricing & Benchmarks (2026)

Apr 16, 2026
GPU Cloud Benchmarks 2026: AI GPU Throughput, Specs, Pricing
Engineering

GPU Cloud Benchmarks 2026: AI GPU Throughput, Specs, Pricing

Apr 15, 2026
AMD MI400 vs NVIDIA B300: MI455X Specs, Price (2026)
Comparison

AMD MI400 vs NVIDIA B300: MI455X Specs, Price (2026)

Apr 6, 2026
AMD MI350X vs NVIDIA B200: Specs, Benchmarks, and Cloud Pricing (2026)
Comparison

AMD MI350X vs NVIDIA B200: Specs, Benchmarks, and Cloud Pricing (2026)

Apr 3, 2026
RTX 4090 for AI/ML: Benchmarks, Specs, and Pricing
Research

RTX 4090 for AI/ML: Benchmarks, Specs, and Pricing

Mar 20, 2026
RTX 5090 LLM Benchmarks: Cost Per Million Tokens Deep Dive
Research

RTX 5090 LLM Benchmarks: Cost Per Million Tokens Deep Dive

Mar 10, 2026
RTX 5090 vs H100 vs B200: Which GPU Is Worth It for AI in 2026?
Comparison

RTX 5090 vs H100 vs B200: Which GPU Is Worth It for AI in 2026?

Mar 8, 2026
Try It Yourself

Try It on Real GPUs

The GPUs behind these guides are the ones you can rent here: H100s, H200s, B200s, and more, billed per minute with no contracts and no minimum. Pick one and you are live in under two minutes.

Deploy Time
< 2 min
Uptime SLA
99.9%
GPU Models
10+
Billing
Per-Min