Model-to-hardware discovery

What do you want to run?

Most people arrive knowing the model, not the memory requirement. Start here and we will work backwards to the infrastructure.

Describe the model

Precision

Indicative memory needed

Estimate, not a guarantee

~90 GB

Based on weights at the chosen precision plus working memory for context and concurrency. Real requirements vary with context length, batching and the serving framework. Treat this as a starting point for a conversation, not a specification.

Hardware with enough memory

Professional AI GPUs
Representative

NVIDIA RTX PRO 6000 Blackwell 96 GB

A large-memory workstation card. It fits into a professional tower and lets you run or fine-tune substantial models on your own hardware.

Memory

96 GB

Deployment

Deskside workstation

ProfessionalHigh-Memory

£7,495

ex VAT

In stock

Data Centre AI Accelerators
Representative

NVIDIA H200 NVL 141 GB

A very capable data centre GPU in card form, so it can be fitted into common rack servers rather than requiring a specialist chassis.

Memory

141 GB

Deployment

Rack server

DatacentreHigh-Memory

£24,950

ex VAT

Low stock

Data Centre AI Accelerators
Representative

NVIDIA B200 SXM Accelerator

The mainstream high-end training GPU. Sold as part of an eight-GPU server rather than as a card you fit yourself.

Memory

180 GB

Deployment

Rack server

DatacentreHigh-Memory

Price on request

Available to order

Data Centre AI Accelerators
Representative

AMD Instinct MI300X

A proven AMD accelerator with 192 GB of memory, frequently chosen for serving large language models economically.

Memory

192 GB

Deployment

Rack server

DatacentreHigh-Memory

Price on request

Available to order

Data Centre AI Accelerators
Representative

AMD Instinct MI355X

AMD's flagship accelerator. Very large memory per GPU means fewer GPUs are needed to hold a big model.

Memory

288 GB

Deployment

Datacentre

DatacentreHigh-Memory

Price on request

Available to order

Data Centre AI Accelerators
Representative

NVIDIA HGX B300 8-GPU Baseboard

The GPU engine used inside eight-way AI servers. It is fitted into a server chassis by the manufacturer, not installed like a plug-in card.

Memory

2304 GB

Deployment

Datacentre

DatacentreHigh-Memory

Price on request

Allocation only

Deployment classes shown on each card: Deskside workstation, Rack server, Datacentre.