AMD Instinct MI325X
128GB HBM3 — lowest-cost ROCm inference and AIGC
Reserved and on-demand, self-service access
Reserved and committed-use discounts up to 60%. All prices GBP, ex-VAT.
MI325X is the cost-efficient entry into the AMD Instinct family — a single-GPU ROCm card engineered for low-volume inference, AIGC generation and small-team fine-tuning without the footprint of an 8-way bare-metal node.
Highlights
- 128GB HBM3 fits modern 7B-13B models in a single card
- Lowest entry price for AMD ROCm inference
- Mature RCCL + vLLM-ROCm stack
Software stack
Best-fit scenarios
- ROCm low-cost inference
- AIGC generation
- Small-model fine-tuning
Pricing by configuration
All available SKUs for this GPU across bare-metal, VM and container form factors. Reserved and committed-use discounts available.
On-demand starting at £1.10/h
| Configuration | GPUs | VRAM total | CPU | Memory | Storage | Network | £/hr | £/mo |
|---|---|---|---|---|---|---|---|---|
Bare Metal — 8× MI325X mi325x-bm-8 | 8 | 1 TB | 192 vCPU | 1 TB DDR5 | 8× 1.92 TB NVMe | 800G RoCE | £9.76 | £7,120 |
Virtual Machine — 1× MI325X Popular mi325x-vm-1 | 1 | 128GB | 30 vCPU | 120 GB DDR5 | 960 GB NVMe | 10Gbps | £1.22 | £890 |
Container — 1× MI325X mi325x-ctr-1 | 1 | 128GB | 16 vCPU | 64 GB | 500 GB | 10Gbps | £1.10 | £800 |
Prices shown are list prices in GBP (ex-VAT). Reserved capacity and committed-use discounts up to 60% available — contact sales for a custom quote.
Available form factors
Same silicon, three consumption models — pick the deployment model that matches your workload, then see every GPU that fits.
Whole server, no virtualization — raw performance and full root.
KVM-based with GPU passthrough — multi-tenant isolation.
Pay per pod-hour with elastic Kubernetes scaling.
Technova at a glance
Full-stack AI cloud
GPUs, networking, storage, managed Slurm & Kubernetes on one platform.
Manual UK deployment
Every resource deployed and configured by our UK engineering team.
UK & US data residency
UK workloads stay in Glasgow & London. US workloads in Texas. You choose the region.
Reliable
Historical uptime over 99.9% with sensible SLAs and fair compensation.
Dual-stack CUDA + ROCm
Professional full-stack expertise across both NVIDIA and AMD.
World-class support
Proactive support from ML engineers and infrastructure specialists.
Ready to deploy AMD Instinct MI325X?
Spin up a private cluster in days, or talk to our sales engineers about custom configurations.