HGX H200
Largest NVIDIA memory for LLM training and inference
- GPU memory
- 141 GB HBM3e
- Power
- 700 W
- Form factor
- SXM5 · NVLink
- Per server
- Up to 8
Eight NVIDIA HGX H200 or H100 GPUs with NVLink, or eight AMD Instinct MI300X or Intel Gaudi 3 accelerators, with two 5th Gen Intel Xeon processors in an air-cooled 6U chassis. Sized, quoted and delivered by Forte Tech across Egypt and the GCC.
NVIDIA H200 / H100, AMD MI300X or Intel Gaudi 3
GPU memory with 8 × AMD Instinct MI300X
DDR5 across 32 DIMMs at up to 5600 MT/s
PCIe x16 slots for high-speed GPU networking
One XE9680 trains, fine-tunes and serves large models. Several of them form the core of an on-premises AI cluster.
Train or fine-tune large language models on eight tightly connected GPUs.
Serve LLMs and RAG assistants to thousands of users with up to 1.5 TB of GPU memory.
Accelerate scientific computing, risk modelling and engineering simulation.
Build AI platforms for government, banking and telecom with data kept on-premises.
The XE9680 links eight accelerators over a high-bandwidth fabric, so a large model can run across all of them as one system, inside your own data center.
Arabic and English assistants trained on your own documents, served from your data center.
Train and fine-tune foundation models on your own data.
City-scale video analytics, document processing and inspection.
Fraud detection, risk modelling and scientific workloads on GPUs.
Every XE9680 uses eight of the same accelerator. Choose the platform and we'll quote a complete, validated system.
Largest NVIDIA memory for LLM training and inference
LLM training, fine-tuning and HPC
Very large models and memory-bound inference
Training and inference with built-in 800 Gb Ethernet
Tell us your model, data and number of users. Our engineers will recommend the right GPU and a complete, compatible build.
Ask our engineersThe XE9680 pairs eight data-center accelerators with Intel Xeon processors, DDR5 and PCIe Gen5 networking in one air-cooled system.
NVIDIA HGX H200 or H100 with NVLink, AMD Instinct MI300X with Infinity Fabric, or Intel Gaudi 3 with RoCE.
Up to 64 cores per processor (or 4th Gen with up to 56) to keep the GPUs fed with data.
32 RDIMM slots at up to 5600 MT/s for data pipelines and pre-processing.
Or 8 × 2.5" NVMe/SAS/SATA (122.88 TB), with BOSS-N1 mirrored boot.
For high-speed InfiniBand or Ethernet GPU networking, plus OCP 3.0 and 2 × 1 GbE.
Signed firmware, Secure Boot and TPM 2.0, managed with iDRAC9, Redfish and OpenManage Enterprise.
Reference configurations our engineers adjust to your models, users, networking and data-center limits.
Eight NVIDIA H200 GPUs for training, fine-tuning and large-model inference.
1.5 TB of GPU memory for very large models and high-concurrency inference.
Intel Gaudi 3 accelerators with built-in 800 Gb Ethernet for scale-out AI.
Need something different? Send us your requirements or BOM and our engineers will propose the right build.
We help banks, telecom operators, government entities and enterprises specify, source and deploy Dell PowerEdge servers.
Factory-configured PowerEdge servers with Dell manufacturer warranty.
Engineers who size CPU, memory, storage and power for your actual workload.
Logistics to your data center or branch, with optional rack installation.
Formal quotations, compliance documents and invoicing for procurement teams.
XE9680 systems are built to order, and lead time depends on accelerator availability. We confirm the delivery date with your quote.
Choose the NVIDIA H200 for the widest software support and large-model training, the H100 for proven training and fine-tuning, the AMD MI300X when you need the most GPU memory per server, and Intel Gaudi 3 for Ethernet-based scale-out AI. Tell us your models and number of users and we'll recommend the right platform.
The XE9680 is an air-cooled 6U server weighing up to 114 kg. It needs a rack and floor rated for its weight, high-power circuits for its 2800 W or 3200 W Titanium power supplies, and strong data-center cooling. Our engineers review your site before delivery.
Dell requires ProDeploy Plus or Dell Customer Deployment services for XE-series servers, including GPU subsystem testing. We include the required deployment in your quotation.
The R770 is a 2U server with up to two GPUs, ideal for inference in a standard rack. The XE9680 is a 6U, 8-GPU platform for model training and large-scale inference.
Some high-end accelerators are subject to U.S. export regulations. We confirm availability and any licensing requirements for your country with your quotation.
Every server is supplied with Dell's manufacturer warranty. Extended Dell ProSupport or ProSupport Plus coverage can be added, and the exact warranty terms are stated on your quotation.
Yes. We supply customers across Egypt and the GCC, including Saudi Arabia, the UAE, Qatar, Kuwait, Oman and Bahrain.
Tell us your models, users and timeline. Our engineers will size the XE9680 system, networking and data-center requirements.