NVIDIA RTX PRO 6000
96GB GDDR7 Blackwell-class workstation GPU — low barrier inference and AIGC
Reserved and on-demand, self-service access
Reserved and committed-use discounts up to 60%. All prices GBP, ex-VAT.
RTX PRO 6000 is a high-memory workstation-class NVIDIA GPU tuned for cost-efficient LLM inference, AIGC generation and small-model fine-tuning where NVLink scale-out is not required.
Highlights
- Largest single-GPU GDDR7 memory in its class
- Lowest entry price for NVIDIA CUDA inference
- Drop-in for any HuggingFace model
Software stack
Best-fit scenarios
- LLM inference (up to ~15B)
- AIGC image / video generation
- Small-model fine-tuning
Pricing by configuration
All available SKUs for this GPU across bare-metal, VM and container form factors. Reserved and committed-use discounts available.
On-demand starting at £1.30/h
| Configuration | GPUs | VRAM total | CPU | Memory | Storage | Network | £/hr | £/mo |
|---|---|---|---|---|---|---|---|---|
Bare Metal — 8× RTX PRO 6000 rtxpro6000-bm-8 | 8 | 768GB | 192 vCPU | 1 TB DDR5 | 8× 1.92 TB NVMe | 800G RoCE | £11.60 | £8,400 |
Virtual Machine — 1× RTX PRO 6000 Popular rtxpro6000-vm-1 | 1 | 96GB | 30 vCPU | 90 GB DDR5 | 960 GB NVMe | 10Gbps | £1.45 | £1,050 |
Container — 1× RTX PRO 6000 rtxpro6000-ctr-1 | 1 | 96GB | 16 vCPU | 64 GB | 500 GB | 10Gbps | £1.30 | £950 |
Prices shown are list prices in GBP (ex-VAT). Reserved capacity and committed-use discounts up to 60% available — contact sales for a custom quote.
Available form factors
Same silicon, three consumption models — pick the deployment model that matches your workload, then see every GPU that fits.
Whole server, no virtualization — raw performance and full root.
KVM-based with GPU passthrough — multi-tenant isolation.
Pay per pod-hour with elastic Kubernetes scaling.
Technova at a glance
Full-stack AI cloud
GPUs, networking, storage, managed Slurm & Kubernetes on one platform.
Manual UK deployment
Every resource deployed and configured by our UK engineering team.
UK & US data residency
UK workloads stay in Glasgow & London. US workloads in Texas. You choose the region.
Reliable
Historical uptime over 99.9% with sensible SLAs and fair compensation.
Dual-stack CUDA + ROCm
Professional full-stack expertise across both NVIDIA and AMD.
World-class support
Proactive support from ML engineers and infrastructure specialists.
Ready to deploy NVIDIA RTX PRO 6000?
Spin up a private cluster in days, or talk to our sales engineers about custom configurations.