NVIDIA H200
141GB HBM3e Hopper — drop-in CUDA acceleration with first-class tooling
Reserved and on-demand, self-service access
Reserved and committed-use discounts up to 60%. All prices GBP, ex-VAT.
H200 is the workhorse Hopper-generation accelerator tuned for high-throughput inference and mainstream LLM training. Best-in-class tooling and model coverage.
Highlights
- First-class CUDA compatibility — drop-in for any HF model
- 141GB HBM3e unlocks larger batches on inference
- Mature NCCL + TensorRT-LLM stack
Software stack
Best-fit scenarios
- Large-scale LLM training
- High-throughput inference (TensorRT-LLM)
- Mixed-precision HPC
Pricing by configuration
All available SKUs for this GPU across bare-metal, VM and container form factors. Reserved and committed-use discounts available.
On-demand starting at £3.96/h
| Configuration | GPUs | VRAM total | CPU | Memory | Storage | Network | £/hr | £/mo |
|---|---|---|---|---|---|---|---|---|
Bare Metal — 8× H200 Popular h200-bm-8 | 8 | 1.1 TB | 192 vCPU | 2 TB DDR5 | 8× 3.84 TB NVMe | 400G NDR InfiniBand | £27.00 | £19,680 |
Virtual Machine — 1× H200 h200-vm-1 | 1 | 141GB | 32 vCPU | 256 GB | 1.92 TB NVMe | 100G | £4.38 | £3,200 |
Container — 1× H200 h200-ctr-1 | 1 | 141GB | 16 vCPU | 128 GB | 500 GB | 100G | £3.96 | £2,890 |
Prices shown are list prices in GBP (ex-VAT). Reserved capacity and committed-use discounts up to 60% available — contact sales for a custom quote.
Available form factors
Same silicon, three consumption models — pick the deployment model that matches your workload, then see every GPU that fits.
Whole server, no virtualization — raw performance and full root.
KVM-based with GPU passthrough — multi-tenant isolation.
Pay per pod-hour with elastic Kubernetes scaling.
Technova at a glance
Full-stack AI cloud
GPUs, networking, storage, managed Slurm & Kubernetes on one platform.
Manual UK deployment
Every resource deployed and configured by our UK engineering team.
UK & US data residency
UK workloads stay in Glasgow & London. US workloads in Texas. You choose the region.
Reliable
Historical uptime over 99.9% with sensible SLAs and fair compensation.
Dual-stack CUDA + ROCm
Professional full-stack expertise across both NVIDIA and AMD.
World-class support
Proactive support from ML engineers and infrastructure specialists.
Ready to deploy NVIDIA H200?
Spin up a private cluster in days, or talk to our sales engineers about custom configurations.