More memory for large AI models

NVIDIA H200 for LLM and AI infrastructure

NVIDIA H200 is designed for applications that benefit from greater memory capacity and bandwidth. We will configure the server, accelerator count and network architecture for your performance target.

Specifications

Key specifications

Review the main product parameters for a quick comparison. Before delivery, we confirm the exact configuration, compatibility, warranty and required quantity.

ArchitectureHopper
Memory141 GB HBM3e
Memory bandwidth4.8 TB/s
Form factorSXM / PCIe NVL by SKU
MIGup to 7 MIG instances
Workloadsmemory-bound LLM / inference / training

Best-fit applications

When to choose this model

A high-performance option for large-language-model training, inference, retrieval, analytics and long-context applications.

  • large-context LLM workloads;
  • memory-bound inference and training;
  • HGX, MGX and OEM server nodes for H200.

Configuration and delivery

What we include in the configuration

  • H200 SXM or H200 NVL;
  • server QVL and firmware;
  • cooling, power and rack readiness;
  • networking and storage profile for the workload.

Important for a reliable deployment

NVIDIA H200

H200 requires compatible server nodes, cooling and power. We include these requirements in the project configuration and delivery plan.

Confirmed product data

Review the main product parameters for a quick comparison. Before delivery, we confirm the exact configuration, compatibility, warranty and required quantity.

Other solutions

Other NVIDIA models

Compare similar solutions or describe your application and we will recommend the best-fit model.

Back to section

Tailored commercial offer

Get price and delivery time

Send the model and quantity or attach an equipment list. If the exact configuration is not yet selected, describe your application and we will prepare a suitable offer.