Cheapest Cloud GPU Pricing 2026, H100, H200, B200 Comparison | VoltageGPU
Live VoltageGPU pricing for NVIDIA H100, H200 and RTX PRO 6000 Blackwell confidential GPUs, side-by-side with AWS, Google Cloud and Microsoft Azure, including their B200 list prices for reference. All VoltageGPU GPUs sealed inside Intel TDX trust domains with trust-domain GPU isolation, per-second billing, no commitment, $5 referral credit.
Cloud GPU Pricing Comparison 2026, VoltageGPU vs AWS vs GCP vs Azure
| GPU | VoltageGPU (Intel TDX) | AWS on-demand | Google Cloud | Azure Confidential | VoltageGPU savings |
|---|
| NVIDIA H100 80GB | $5.00/hour | $4.30/hr (p5.48xlarge ÷ 8, no TEE) | $3.67/hr (a3-highgpu, no TEE; confidential A3 High rate by zone) | $6.98/hr (NC H100 v5, no TEE); confidential NCC40ads H100 v5 $8.90/hr | below Azure’s confidential H100; AWS p5 has no TEE |
| NVIDIA H200 141GB | $6.58/hour | $12.25/hr (p5e.48xlarge ÷ 8, no TEE) | $11.06/hr (a3-megagpu, no TEE) | $13.96/hr (ND H200 v5, no TEE) | no hyperscaler lists a confidential H200; ours is $8.08/hr as a Confidential VM |
| NVIDIA B200 180GB (listed, never available to date, not attested) | $10.60/hour | $26.32/hr (p6-b200.48xl ÷ 8) | $25.00/hr (a4-highgpu) | $28.50/hr (ND B200 v6, no TEE) | no hyperscaler lists a confidential B200 |
Comparison prices are public list prices from each provider's pricing page (April 2026). VoltageGPU prices are live from the Targon /inventory endpoint and update in real time on this page.
Best Price Per Hour Cloud GPU Providers, Why VoltageGPU is Cheapest
- Per-second billing, pay for the exact second the GPU runs, not for whole hours like AWS p5
- No reserved-instance lock-in, no 1-year or 3-year commitments to unlock the listed price
- $5 referral credit covers ~2 hours of confidential H100 with zero credit card required
- Same Intel TDX confidential computing technology as Azure / Google Confidential VMs, at a fraction of the price
- Bitcoin accepted alongside Stripe, no SaaS-style PO process
Four ways to rent compute on VoltageGPU
- Confidential VM (Intel TDX, attestation generated by the tenant): RTX 6000B $3.80/hour, H100 $6.95/hour, single-GPU H200 from $8.08/hour with both the Intel TDX quote and the NVIDIA GPU attestation, 8x H100 node $55.62/hour with NVIDIA attestation of all eight GPUs (multi-GPU Protected PCIe mode, verified 10 September 2026), 8x H200 node $64.62/hour. Self-service, boots in about two and a half minutes, one hour upfront then per second.
- Confidential containers (Intel TDX trust domain, GPU passed through, GPU confidential-computing mode off): H100 $5.00/hour, H200 $6.58/hour, B200 $10.60/hour per GPU-hour, up to 8 GPUs, persistent volumes.
- Standard GPUs (no enclave, for data that is not sensitive): RTX 3090, RTX 4090, RTX 5090, RTX PRO 6000, H100, H200 from a distributed provider network, live prices on this page, typically from under $0.50 per GPU-hour on the smallest cards.
- Confidential CPU servers (Intel TDX, no GPU) for PDF parsing, OCR, embeddings and ETL, from $0.18/hour.
Cheapest Cloud GPU for LLM Inference 2026
- Qwen3-32B (TEE): $0.15/M input · $0.44/M output, vs GPT-4o at $2.50/$10
- DeepSeek-V3.2 (TEE): $0.20/M input · $0.89/M output, vs OpenAI o1 at $15/$60
- Llama-3.3-70B (TEE): $0.35/M input · $0.40/M output
- OpenAI-compatible API at
api.voltagegpu.com/v1, drop-in for OpenAI SDK, LangChain, LlamaIndex
Price Comparison vs AWS, GCP, Azure
A single-GPU H200 Confidential VM is $6.58/hour here, against $11–14/hour for comparable confidential GPU instances at the hyperscalers, billed per second with no commitment and no minimum. We have not run a reproducible throughput benchmark of our own, so we quote no performance ratio: the comparison we can stand behind today is on price, on availability, and on who generates the attestation. See the side-by-side with vendor sources.
GPU Cloud for AI Training 2026
- Pre-training 7B–70B from scratch: H200 cluster ($6.58/hour), 141 GB HBM3e fits a 70B model with KV cache headroom
- Frontier-model training (405B–2T): the 8x H100 node ($5.00/hour per GPU), attested per GPU in Protected PCIe mode; the B200 is listed, never available to date, not attested
- LoRA / QLoRA fine-tuning: any GPU; H100 80GB is the cost-optimal pick at $5.00/hour
- RLHF / DPO: H200 for reward-model + policy in the same pod thanks to large VRAM
- All training runs inside Intel TDX, your training data and gradients are encrypted in memory
Confidential Agents Pricing
- Free: $0, 1 seat, private chat, all 9 agents, no card
- Plus: $20/month, 1 seat, 2,000 messages/month, personal Telegram agent, web search and memory
- Starter: $349/month, 2,000 requests/month, 3 seats, agent mode with tools, clause checklists, risk scoring
- Pro: $1,199/month, 5,000 requests/month, 10 seats, API access, priority support, audit log
- Enterprise: Custom, unlimited seats, SSO/SCIM, dedicated support, SLA, DPA included
Frequently Asked Questions, Cloud GPU Pricing
What is the cheapest cloud GPU per hour in 2026?
For a hardware-isolated (Intel TDX) H100 80GB, VoltageGPU charges $5.00/hr. AWS p5 ($4.30/hr) and Google a3 ($3.67/hr) are commodity instances without a trusted execution environment, so they are cheaper per hour but not comparable; Azure’s confidential NCC40ads H100 v5 lists at $8.90/hr (12 September 2026). For non-confidential workloads our Standard tier (no enclave) starts under $0.50/hr on the smallest cards, live on this page. All VoltageGPU pricing is per-second with no commitment.
How do VoltageGPU prices compare to AWS, GCP, and Azure?
The only hyperscaler SKU with a GPU inside a trust boundary and a public price is Azure’s NCC H100 v5 (AMD SEV-SNP plus one H100 NVL, $8.90/hr list on 12 September 2026); Google Cloud sells confidential H100 on A3 High with Intel TDX at per-zone rates, and AWS has no confidential GPU. Hyperscaler H200 and B200 sizes such as Azure ND H200 v5 or ND B200 v6 are not confidential VMs. Example: a Confidential VM with one H200 in NVIDIA CC mode is $8.08/hr on VoltageGPU, with both proofs verified; the H200 container is $6.58/hr without tenant-side GPU attestation. Dated, sourced comparison: four clouds compared.
Is there a minimum commitment or reserved-instance discount?
No. VoltageGPU prices listed on this page are the price you pay, with no contracts, no reserved instances, and no spot/on-demand differential. Per-second billing means you can deploy an H100 for a five-minute experiment and pay only for those five minutes. Minimum top-up is $5.
How does VoltageGPU billing work?
VoltageGPU uses per-second billing. You only pay for the exact time your GPU is running. Stop your pod and billing stops instantly.
Why is VoltageGPU cheaper than hyperscalers if it uses the same Intel TDX hardware?
Lean operations and per-second billing, zero waste on idle time. The GPUs are enterprise NVIDIA hardware (H100, H200, RTX PRO 6000 Blackwell) in professional Tier-III data centers with the same Intel TDX confidential computing stack used by Azure and Google. We pass the savings through instead of bundling them into hyperscaler ecosystem services.