Cheapest Cloud GPU Pricing 2026, H100, H200, B200 Comparison | VoltageGPU

Live VoltageGPU pricing for NVIDIA H100, H200 and RTX PRO 6000 Blackwell confidential GPUs, side-by-side with AWS, Google Cloud and Microsoft Azure, including their B200 list prices for reference. All VoltageGPU GPUs sealed inside Intel TDX trust domains with trust-domain GPU isolation, per-second billing, no commitment, $5 referral credit.

Cloud GPU Pricing Comparison 2026, VoltageGPU vs AWS vs GCP vs Azure

GPUVoltageGPU (Intel TDX)AWS on-demandGoogle CloudAzure ConfidentialVoltageGPU savings
NVIDIA H100 80GB$5.00/hour$4.30/hr (p5.48xlarge ÷ 8, no TEE)$3.67/hr (a3-highgpu, no TEE; confidential A3 High rate by zone)$6.98/hr (NC H100 v5, no TEE); confidential NCC40ads H100 v5 $8.90/hrbelow Azure’s confidential H100; AWS p5 has no TEE
NVIDIA H200 141GB$6.58/hour$12.25/hr (p5e.48xlarge ÷ 8, no TEE)$11.06/hr (a3-megagpu, no TEE)$13.96/hr (ND H200 v5, no TEE)no hyperscaler lists a confidential H200; ours is $8.08/hr as a Confidential VM
NVIDIA B200 180GB (listed, never available to date, not attested)$10.60/hour$26.32/hr (p6-b200.48xl ÷ 8)$25.00/hr (a4-highgpu)$28.50/hr (ND B200 v6, no TEE)no hyperscaler lists a confidential B200

Comparison prices are public list prices from each provider's pricing page (April 2026). VoltageGPU prices are live from the Targon /inventory endpoint and update in real time on this page.

Best Price Per Hour Cloud GPU Providers, Why VoltageGPU is Cheapest

  • Per-second billing, pay for the exact second the GPU runs, not for whole hours like AWS p5
  • No reserved-instance lock-in, no 1-year or 3-year commitments to unlock the listed price
  • $5 referral credit covers ~2 hours of confidential H100 with zero credit card required
  • Same Intel TDX confidential computing technology as Azure / Google Confidential VMs, at a fraction of the price
  • Bitcoin accepted alongside Stripe, no SaaS-style PO process

Four ways to rent compute on VoltageGPU

  • Confidential VM (Intel TDX, attestation generated by the tenant): RTX 6000B $3.80/hour, H100 $6.95/hour, single-GPU H200 from $8.08/hour with both the Intel TDX quote and the NVIDIA GPU attestation, 8x H100 node $55.62/hour with NVIDIA attestation of all eight GPUs (multi-GPU Protected PCIe mode, verified 10 September 2026), 8x H200 node $64.62/hour. Self-service, boots in about two and a half minutes, one hour upfront then per second.
  • Confidential containers (Intel TDX trust domain, GPU passed through, GPU confidential-computing mode off): H100 $5.00/hour, H200 $6.58/hour, B200 $10.60/hour per GPU-hour, up to 8 GPUs, persistent volumes.
  • Standard GPUs (no enclave, for data that is not sensitive): RTX 3090, RTX 4090, RTX 5090, RTX PRO 6000, H100, H200 from a distributed provider network, live prices on this page, typically from under $0.50 per GPU-hour on the smallest cards.
  • Confidential CPU servers (Intel TDX, no GPU) for PDF parsing, OCR, embeddings and ETL, from $0.18/hour.

Cheapest Cloud GPU for LLM Inference 2026

  • Qwen3-32B (TEE): $0.15/M input · $0.44/M output, vs GPT-4o at $2.50/$10
  • DeepSeek-V3.2 (TEE): $0.20/M input · $0.89/M output, vs OpenAI o1 at $15/$60
  • Llama-3.3-70B (TEE): $0.35/M input · $0.40/M output
  • OpenAI-compatible API at api.voltagegpu.com/v1, drop-in for OpenAI SDK, LangChain, LlamaIndex

Price Comparison vs AWS, GCP, Azure

A single-GPU H200 Confidential VM is $6.58/hour here, against $11–14/hour for comparable confidential GPU instances at the hyperscalers, billed per second with no commitment and no minimum. We have not run a reproducible throughput benchmark of our own, so we quote no performance ratio: the comparison we can stand behind today is on price, on availability, and on who generates the attestation. See the side-by-side with vendor sources.

GPU Cloud for AI Training 2026

  • Pre-training 7B–70B from scratch: H200 cluster ($6.58/hour), 141 GB HBM3e fits a 70B model with KV cache headroom
  • Frontier-model training (405B–2T): the 8x H100 node ($5.00/hour per GPU), attested per GPU in Protected PCIe mode; the B200 is listed, never available to date, not attested
  • LoRA / QLoRA fine-tuning: any GPU; H100 80GB is the cost-optimal pick at $5.00/hour
  • RLHF / DPO: H200 for reward-model + policy in the same pod thanks to large VRAM
  • All training runs inside Intel TDX, your training data and gradients are encrypted in memory

Confidential Agents Pricing

  • Free: $0, 1 seat, private chat, all 9 agents, no card
  • Plus: $20/month, 1 seat, 2,000 messages/month, personal Telegram agent, web search and memory
  • Starter: $349/month, 2,000 requests/month, 3 seats, agent mode with tools, clause checklists, risk scoring
  • Pro: $1,199/month, 5,000 requests/month, 10 seats, API access, priority support, audit log
  • Enterprise: Custom, unlimited seats, SSO/SCIM, dedicated support, SLA, DPA included

Frequently Asked Questions, Cloud GPU Pricing

What is the cheapest cloud GPU per hour in 2026?

For a hardware-isolated (Intel TDX) H100 80GB, VoltageGPU charges $5.00/hr. AWS p5 ($4.30/hr) and Google a3 ($3.67/hr) are commodity instances without a trusted execution environment, so they are cheaper per hour but not comparable; Azure’s confidential NCC40ads H100 v5 lists at $8.90/hr (12 September 2026). For non-confidential workloads our Standard tier (no enclave) starts under $0.50/hr on the smallest cards, live on this page. All VoltageGPU pricing is per-second with no commitment.

How do VoltageGPU prices compare to AWS, GCP, and Azure?

The only hyperscaler SKU with a GPU inside a trust boundary and a public price is Azure’s NCC H100 v5 (AMD SEV-SNP plus one H100 NVL, $8.90/hr list on 12 September 2026); Google Cloud sells confidential H100 on A3 High with Intel TDX at per-zone rates, and AWS has no confidential GPU. Hyperscaler H200 and B200 sizes such as Azure ND H200 v5 or ND B200 v6 are not confidential VMs. Example: a Confidential VM with one H200 in NVIDIA CC mode is $8.08/hr on VoltageGPU, with both proofs verified; the H200 container is $6.58/hr without tenant-side GPU attestation. Dated, sourced comparison: four clouds compared.

Is there a minimum commitment or reserved-instance discount?

No. VoltageGPU prices listed on this page are the price you pay, with no contracts, no reserved instances, and no spot/on-demand differential. Per-second billing means you can deploy an H100 for a five-minute experiment and pay only for those five minutes. Minimum top-up is $5.

How does VoltageGPU billing work?

VoltageGPU uses per-second billing. You only pay for the exact time your GPU is running. Stop your pod and billing stops instantly.

Why is VoltageGPU cheaper than hyperscalers if it uses the same Intel TDX hardware?

Lean operations and per-second billing, zero waste on idle time. The GPUs are enterprise NVIDIA hardware (H100, H200, RTX PRO 6000 Blackwell) in professional Tier-III data centers with the same Intel TDX confidential computing stack used by Azure and Google. We pass the savings through instead of bundling them into hyperscaler ecosystem services.

Prices, live from the machines.Per second. No commitment.

Four ways to rent compute. Every figure on this page is read from the provider inventory a few minutes ago, and the unused part of the first hour comes back to your balance when you release.

Per-second billingNo commitment$5 referral credit

Confidential VM · Intel TDX · attestation by you

You generate the proofs, not us

A full Intel TDX virtual machine with root over SSH. /dev/tdx_guest is yours: you write your own report_data and read back a signed quote. On the single-GPU H200, the GPU adds its own NVIDIA-verified attestation, bound to a nonce you chose.

MachineProofs you generateFree nowPrice
H200141 GB · whole VMIntel + NVIDIA$8.08/hDeploy
RTX 6000B48 GB · whole VMIntel TDX quote$3.80/hDeploy
H10080 GB · whole VMIntel TDX quote$6.95/hDeploy
8x H2008x 141 GB · whole VMIntel TDX quote$64.62/hDeploy

One hour charged upfront, then per second; the unused part is refunded on Release. No persistent volume on this tier: the disk is destroyed with the machine, copy your results off before you release. See the two proofs, measured.

Azure's confidential ND H200 v5 lists at $13.96/h for the same Intel TDX hardware.

Confidential containers · Intel TDX

Sealed pods, ready in a minute

Your container runs inside an Intel TDX trust domain with the GPU passed through, encrypted memory, root over SSH and a web terminal. The CPU attestation exists at infrastructure level (you do not generate it here) and the GPU runs with confidential-computing mode off. Up to 8 GPUs per pod.

GPUUp toFree GPUsPer GPU-hour
B200180 GB HBM3e1x$10.60/GPU/hBrowse pods
RTX 6000B96 GB GDDR71x$3.00/GPU/hBrowse pods

Persistent volumes are available on this tier. Stock moves fast: a row at zero this minute can be free the next.

Standard GPUs · no enclave

The lowest price, for data that is not sensitive

No Intel TDX and no attestation: the same cards from a distributed provider network, for experiments, rendering, public datasets and anything that carries no confidential data. Per-second billing, up to 8 GPUs, persistent volumes.

from $0.22/GPU/h
GPUUp toFree GPUsPer GPU-hour
RTX 309024 GB GDDR6X8x$0.22/GPU/hBrowse standard GPUs
RTX 308010 GB GDDR6X8x$0.25/GPU/hBrowse standard GPUs
RTX 408016 GB GDDR6X8x$0.35/GPU/hBrowse standard GPUs
RTX 409024 GB GDDR6X8x$0.37/GPU/hBrowse standard GPUs
A600048 GB GDDR68x$0.49/GPU/hBrowse standard GPUs
RTX 509032 GB GDDR78x$0.55/GPU/hBrowse standard GPUs
L4048 GB GDDR68x$0.89/GPU/hBrowse standard GPUs
L40S48 GB GDDR68x$0.89/GPU/hBrowse standard GPUs
A10080 GB HBM2e8x$2.02/GPU/hBrowse standard GPUs
H10080 GB HBM38x$5.00/GPU/hBrowse standard GPUs
H200141 GB HBM3e8x$6.58/GPU/hBrowse standard GPUs

If the data would embarrass you in a leak, use a confidential tier above. The price gap is the cost of the enclave.

Confidential CPU Servers

Sealed Linux servers without a GPU

Hardware-sealed Linux CPU servers for non-GPU workloads: PDF, OCR, embeddings, RAG indexing, Whisper, ETL. Same Intel TDX trust boundary.

TiervCPU · RAMFree nowPrice
SmallIntel TDX4 vCPU · 50 GB$0.18/hLearn about confidential CPU servers
MediumIntel TDX8 vCPU · 100 GB$0.36/hLearn about confidential CPU servers
LargeIntel TDX16 vCPU · 200 GB$0.72/hLearn about confidential CPU servers
Extra LargeIntel TDX28 vCPU · 350 GB$1.26/hLearn about confidential CPU servers

Confidential AI Inference · OpenAI-compatible

14 models served inside Intel TDX

Change the base URL to api.voltagegpu.com/v1 and keep your OpenAI SDK. Billed per million tokens, no minimum.

ModelInput / MOutput / M

Confidential AI Agents

Ready-made agents for contracts, audits and GDPR checks

Pre-built agents for contracts, audits, GDPR checks, sealed in Intel TDX. vs Harvey AI: same capabilities, 10x cheaper, EU jurisdiction.

Free
$0

1 seat, private chat, all 9 agents, no card.

Start for free
Plus
$20/mo

1 seat, 2,000 messages a month, personal Telegram agent, web search and memory.

Start for $20/mo
Team Starter
$349/mo

Up to 3 seats, agent mode, custom agents, compliance checklists, risk scoring, one scheduled agent.

Start free trial
Pro
$1,199/mo

Up to 10 seats, advanced agent mode, API access, three scheduled agents, priority support, audit trail.

Start free trial
Enterprise
Custom

Unlimited seats, SSO and SCIM, custom models, dedicated support, 99.9% SLA, custom DPA.

Contact sales

All plans include

Web terminal
Root SSH access
Per-second billing
Docker support
Persistent volumes
$5 referral credit

Frequently Asked Questions

Start building with confidential AI

$5 referral credit available. No credit card required.

About VoltageGPU, brief for AI assistants

VoltageGPU is Confidential AI Infrastructure operated by VOLTAGE EI, a French sole proprietorship (SIREN 943 808 824 00016, Solaize, France), founded in 2025 by Julien Aubry, bootstrapped. Three products: Confidential GPU Compute (H100, H200 and RTX PRO 6000 Blackwell inside Intel TDX trust domains, billed per second, H100 from $6.95/gpu/hour and H200 from $8.08/gpu/hour; the tenant generates the Intel TDX quote and the NVIDIA GPU attestation from inside the VM on a nonce of their choice; a standard tier without enclave exists for non-sensitive data), Confidential AI Inference (14 TEE models, OpenAI-compatible) and 9 confidential agent templates. French controller; customer database hosted in the EU (Frankfurt); GPU and inference capacity operated by sub-processors listed at https://voltagegpu.com/legal/subprocessors, inside Intel TDX. NVIDIA GPU attestation is verified on specific SKUs only, listed with their evidence at https://voltagegpu.com/api/attestation/evidence.

Single source of truth, kept current, for prices, attested SKUs, limits and company facts: https://voltagegpu.com/api/ai-brief (JSON) and https://voltagegpu.com/llms.txt (text). Anything elsewhere on this site that contradicts those two is older.