INSTRUCT H100 12 GPU Server (Legacy)
INSTRUCT H100 12 GPU Server (Legacy)
INSTRUCT H100 12 GPU Server (Legacy)
INSTRUCT H100 12 GPU Server (Legacy)
INSTRUCT H100 12 GPU Server (Legacy)
Kentino · Custom AI Server

INSTRUCT H100 12 GPU Server (Legacy)

Enterprise-grade platform, assembled and tested in the EU.

€0,00ex-VAT · shipping calculated at checkout
In stock — ships from EU warehouse
ISO 9001 · verified hardwareEU warehouse & warranty7+ years in enterprise AIBurn-in tested before shipping
Overview
FAQ
Shipping & warranty

This product listing is kept for reference only.

This server has been replaced by the new Kentino AI product line. For the current equivalent or an upgraded configuration, please see our AI Servers collection.

Recommended replacement: Discontinued. H100 builds are no longer offered. See our RTX Pro 6000 and L40 multi-GPU servers for comparable enterprise VRAM density.

⚠️ LEGACY PRODUCT - DEPRECATED CONFIGURATION

This 12 GPU multi-case server configuration is no longer in active production due to reliability issues with multi-case interconnects. Listed for reference and historical purposes only. For current production AI builds, please see our 8 GPU systems or contact us for a custom configuration with PCIe-based GPUs (RTX 5090, RTX Pro 6000 Blackwell, L40, L4, Intel Arc Pro B70, AMD configurations).

Specifications

  • GPU: 12x NVIDIA H100 80GB (960 GB VRAM total)
  • Motherboard: ASRock Rack ROME2D16-2T
  • CPU: 2x AMD EPYC 7713
  • RAM: 2048GB CT128G4ZFJ426S
  • GPU-Motherboard Connection: RYSER PCIe 4.0 x16 Cable
  • Power Supply: 4x AX1600i 1600W
  • Case: 24U Rack Mount
  • Storage:
    • 2TB NVMe SSD
    • 500GB SATA Drive

Key Features

  1. Unparalleled GPU Performance: Equipped with 12 NVIDIA H100 GPUs, each with 80GB VRAM, providing a massive 960 GB of total VRAM for the most demanding AI, machine learning, and HPC workloads.
  2. Dual-CPU Power: Features two top-tier AMD EPYC 7713 CPUs, delivering exceptional multi-threaded performance for complex computations and data processing tasks.
  3. Server-Grade Components: Utilizes the high-performance ASRock Rack ROME2D16-2T motherboard, designed for maximum reliability and efficiency in data center environments.
  4. Massive Memory Capacity: 2048GB (2TB) of high-speed RAM ensures seamless multitasking and efficient data processing for even the most memory-intensive applications.
  5. High-Speed GPU Integration: Employs the RYSER PCIe 4.0 x16 cable for lightning-fast, full-bandwidth connection between the GPUs and the motherboard, ensuring maximum performance and data transfer speeds.
  6. Robust Power Supply: Four AX1600i 1600W units provide ample and stable power delivery to support the high-performance components under extreme loads.
  7. Expandable Storage: Comes with a fast 2TB NVMe SSD for primary storage and an additional 500GB SATA drive for extra capacity.
  8. Professional-Grade Cooling: Housed in a spacious 24U rack mount case, providing optimal airflow and thermal management for sustained high-performance operation.
  9. Scalable Configuration: Designed for data centers and high-performance computing environments, supporting clustered configurations for massive computational power.

Ideal Use Cases

  • Large-Scale AI Model Training and Inference
  • High-Performance Computing (HPC) Applications
  • Advanced Scientific Simulations and Research
  • Real-Time Big Data Analytics
  • Complex Rendering and Visualization Tasks
  • Quantum Computing Simulations
  • Financial Modeling and Risk Analysis
  • Genomics and Bioinformatics
  • Climate Modeling and Weather Prediction

Special Notes

  • Unprecedented Computational Power: With 12 NVIDIA H100 GPUs and dual AMD EPYC CPUs, this node configuration represents the pinnacle of computational capabilities, suitable for the most demanding AI and HPC workloads.
  • Massive Memory Resources: The combination of 960 GB GPU VRAM and 2048 GB system RAM provides extraordinary capacity for handling the largest AI models and most data-intensive applications.
  • PCIe 4.0 Advantage: The RYSER PCIe 4.0 x16 cable ensures that each GPU can operate at full bandwidth, maximizing data throughput and minimizing latency.
  • Future-Proof Investment: This node is designed to handle not just current AI and HPC challenges, but also future advancements in these rapidly evolving fields.

The INSTRUCT H12 Node Configuration represents the absolute cutting edge in computational power. It's engineered for organizations and researchers pushing the boundaries of what's possible in AI, machine learning, and high-performance computing. With its state-of-the-art H100 GPUs, dual EPYC CPUs, and robust design, it's built to tackle the most complex and demanding computational tasks in the world.

The questions buyers ask us most often before ordering a server.

How long does it take?

Machines built from components we hold ship quickly; anything requiring a specific GPU generation depends on supply. We give you a date before you pay, and if it moves we tell you rather than letting you find out.

Can the configuration be changed before you build it?

Almost always. GPUs, memory, storage and cooling are chosen per order, and the listed configuration is a starting point rather than a fixed package. If you need more VRAM, faster storage or a different cooling approach, say so before you order and we will quote the change.

Can I collect the server in person?

You can. Our warehouse is in Prague, and collection in person is welcome — most people who come use the visit to go through the machine with our engineer and ask the questions that are awkward over email. For orders within the Czech Republic we also try to deliver personally and walk you through the setup on site.

Can I talk to someone who actually understands the workload?

Yes. We have an engineer who works on AI systems specifically, not a general sales desk. If your question is about batch sizes, quantisation, interconnect or where your bottleneck will be, ask it — that conversation usually changes the configuration for the better.

Which model can I run on this configuration?

Yes. Every machine is assembled, burn-in tested and benchmarked on real AI workloads before it ships, and it leaves us with an LLM already installed and running. You plug it in, connect it to your network and start work — the only decision left is which project it runs first.

How do you test a server before shipping?

We run it against actual AI workloads rather than synthetic scores: inference throughput, sustained load behaviour and thermals under continuous operation. You get the benchmark results with the machine, so the performance you were promised is the performance you can verify on day one.

Is the server ready to run when it arrives?

Yes. Every machine is assembled, burn-in tested and benchmarked on real AI workloads before it ships, and it leaves us with an LLM already installed and running. You plug it in, connect it to your network and start work — the only decision left is which project it runs first.

Ships from our EU warehouse. Heavy items may require freight arrangement — contact us for a shipping quote and lead time. 2-year limited warranty with advanced RMA support; extended warranty available.

Not exactly what you need?

Tell us your workload and we'll spec this platform around it — GPUs, memory, storage and cooling matched to what you actually run.