GPU Pricing
55 guides in this topicThis is the pricing cluster: what a GPU actually costs to rent, broken down provider by provider and updated as rates move. Each post pulls a named provider's published on-demand and spot rates (AWS P5, Azure ND-series, GCP A3/A4, CoreWeave, Nebius, Crusoe, and a couple dozen more) and lines them up against Spheron's live pricing so the comparison isn't hypothetical.
A second thread runs through here too: API pricing versus self-hosting. When a frontier model cuts its per-token rate, that changes the token volume where renting your own GPU cluster starts winning, and these posts do that math with the actual API prices, not estimates.
The pillar post, GPU Cloud Pricing Comparison 2026, is the widest lens: H100 and B200 rates across 15+ providers in one table, refreshed regularly. Start there if you're shopping broadly. Drop into an individual provider post (Azure H100, Lambda Cloud, Runpod, Vast.ai) if you already have a quote in hand and want to know whether it's actually competitive. Every number here traces back to a provider's own rate card or Spheron's live pricing API, and we say when a figure is spot instead of on-demand.
Start Here
All GPU Pricing Guides

Baseten Pricing 2026: Model Serving Cost vs Renting GPUs
Jul 30, 2026
Genesis Cloud Pricing 2026: EU H100 Cost vs Spheron
Jul 30, 2026
Replicate Pricing 2026: Per-Second Cost vs Renting a GPU
Jul 30, 2026
AMD MI325X Pricing 2026: Availability and Cost vs Spheron
Jul 28, 2026
Meta Muse Spark API Pricing vs Self-Hosted LLMs (2026)
Jul 28, 2026
OVHcloud GPU Pricing 2026: H100 Cost and Sovereign AI
Jul 27, 2026
GPU Export Controls 2026: What It Means for Cloud Pricing
Jul 26, 2026
Mistral API Pricing vs Self-Hosted LLMs: Cost and Privacy in 2026
Jul 23, 2026
TensorDock Pricing 2026: Cheapest H100 Rental Cost vs Spheron
Jul 23, 2026
AWS Bedrock Pricing vs Self-Hosted LLMs (2026)
Jul 22, 2026
Groq API Pricing 2026: Cost Per Token vs GPU Rental
Jul 22, 2026
NVIDIA H100 Price 2026: Buy vs Rent Costs Compared
Jul 20, 2026
Scaleway H100 Pricing 2026: Sovereign AI GPU Cost vs Spheron
Jul 20, 2026
H200 NVL vs SXM5: Form Factor Decision Guide (2026)
Jul 19, 2026
NVIDIA DGX Cloud Lepton Pricing 2026: Marketplace Cost vs Spheron
Jul 18, 2026
Fireworks AI Pricing 2026: Inference API Cost vs GPU Rental
Jul 17, 2026
Grok 4.5 API Pricing vs Self-Hosted LLMs (2026)
Jul 16, 2026
AWS Capacity Blocks Pricing 2026: B200/B300 Price Hike vs Spheron
Jul 15, 2026
Azure MI300X Pricing 2026: ND MI300X v5 Cost vs Spheron
Jul 15, 2026
NVIDIA L4 vs L40S: Cheapest GPU for AI Inference (2026)
Jul 13, 2026
GB300 NVL72 vs GB200 NVL72: Pricing & Availability (2026)
Jul 12, 2026
GMI Cloud Pricing 2026: H100 and B200 Cost vs Spheron
Jul 12, 2026
Crusoe Cloud Pricing 2026: H100 and B200 Cost Per Hour vs Spheron
Jul 11, 2026
DeepSeek API Pricing vs Self-Hosted LLMs: Cost and Privacy (2026)
Jul 11, 2026
AMD MI300X and MI355X Pricing 2026: Where to Rent Cheapest
Jul 8, 2026
Google TPU v7 Ironwood vs NVIDIA B200: Inference Cost (2026)
Jul 7, 2026
Azure H200 Pricing 2026: ND H200 v5 Cost vs Spheron
Jul 6, 2026
Claude Opus 4.8 API Pricing vs Self-Hosted LLMs (2026)
Jul 6, 2026
Google Cloud A4 B200 Pricing 2026: GCP Cost vs Spheron
Jul 6, 2026
AWS EC2 G7 Pricing 2026: RTX PRO 4500 Blackwell Per-Hour Cost vs Spheron
Jun 29, 2026
NVIDIA B200 Cloud Pricing 2026: Per-Hour Rental Cost Across Providers vs Spheron
Jun 20, 2026
NVIDIA H100 News 2026: Price Changes, H200 and Successor Updates, and Cloud Availability
Jun 14, 2026
GPU Cloud News June 2026: Latest Hardware Launches, Pricing Changes, and Availability Updates
Jun 13, 2026
NVIDIA Vera Rubin NVL72 GPU Cloud: Availability, Cost Per Token, and Planning Your Rubin Rental in H2 2026
Jun 12, 2026
NVIDIA B300 vs B200 for AI Inference: Is Blackwell Ultra Worth the Premium? (2026 Cost-Per-Token Guide)
Jun 10, 2026
Huawei Ascend 950 vs NVIDIA B300 and B200 for LLM Inference: TPS, Memory Bandwidth, and Cost Comparison (2026)
Jun 10, 2026
SambaNova SN40L vs NVIDIA H200 and B200 on GPU Cloud: RDU Inference Benchmarks, Pricing, and Migration Guide (2026)
May 30, 2026
L40S vs H100 for AI Inference: When the L40S Wins on Cost Per Token (2026 Decision Guide)
May 26, 2026
What Is a GPU Cloud? Definition, How It Works, and When to Use One (2026)
May 24, 2026
Google Cloud A3 H100 Pricing 2026: GCP Per-Hour Cost vs Spheron Breakdown
May 22, 2026
Azure H100 Pricing 2026: ND H100 v5 Per-Hour Cost Compared to Spheron
May 21, 2026
NVIDIA Vera Rubin NVL72 (H300): Specs, Pricing & Release Date (2026)
May 15, 2026
DeepSeek V3.2 vs Llama 4 vs Qwen 3 (2026): Cost-per-Token from $0.04 to $0.31, Which to Pick
May 14, 2026
Intel Gaudi 3 vs NVIDIA H200 and B200: LLM Inference Benchmarks, Pricing, and Migration Guide (2026)
May 13, 2026
Cerebras vs NVIDIA H100: Wafer-Scale vs GPU for LLM Inference (2026 Decision Guide)
Apr 28, 2026
NVIDIA B300 (Blackwell Ultra): 288GB Specs, Pricing & Benchmarks (2026)
Apr 16, 2026
GPU Cloud Benchmarks 2026: AI GPU Throughput, Specs, Pricing
Apr 15, 2026
AMD MI400 vs NVIDIA B300: MI455X Specs, Price (2026)
Apr 6, 2026
AMD MI350X vs NVIDIA B200: Specs, Benchmarks, and Cloud Pricing (2026)
Apr 3, 2026
RTX 4090 for AI/ML: Benchmarks, Specs, and Pricing
Mar 20, 2026
RTX 5090 LLM Benchmarks: Cost Per Million Tokens Deep Dive
Mar 10, 2026
RTX 5090 vs H100 vs B200: Which GPU Is Worth It for AI in 2026?
Mar 8, 2026Try It on Real GPUs
The GPUs behind these guides are the ones you can rent here: H100s, H200s, B200s, and more, billed per minute with no contracts and no minimum. Pick one and you are live in under two minutes.


