GPU Servers
8 articlesPCIe Gen6 and 800G SuperNICs: What the 2026 Server Refresh Changes
ConnectX-8 combines an 800G NIC with an integrated PCIe Gen6 switch in one device, exposing 48 Gen6 lanes from an x16 connector. That consolidation removes components from AI server designs…
The Air-Cooled Middle: RTX PRO Servers Between Workstation and HGX
MGX-based RTX PRO servers pack up to eight 96 GB GDDR7 GPUs - 768 GB aggregate - into standard air-cooled racks, serving multiple 70B-class replicas plus rendering and vGPU duty…
NVIDIA GB300 NVL72 Supercluster: Inside the 8-Rack Containerised AI Factory Node
The NVIDIA GB300 NVL72 supercluster packs eight NVL72 racks — 576 Blackwell Ultra B300 GPUs and 288 Grace CPUs — into one containerised AI factory node delivering ~11.5 EFLOPS FP4,…
H100 vs H200 vs B200: Which GPU for Your Workload?
Pick by memory and workload: the H100 (80 GB, 3.35 TB/s) is the proven workhorse for training and inference where 80 GB fits; the H200 (141 GB HBM3e, 4.8 TB/s)…
InfiniBand vs Spectrum-X vs Ethernet for AI Clusters
Choose the fabric by workload: InfiniBand for training and HPC, where GPU-to-GPU latency dominates all-reduce; Spectrum-X or RoCE Ethernet for inference and multi-tenant clouds, where cost and interoperability matter more…
Air-Cooled vs Liquid-Cooled GPU Racks: When to Switch
Air cooling runs out of headroom at roughly 35 kW per rack. Below that, well-designed airflow is fine; above it, direct-to-chip liquid cooling becomes necessary, and beyond ~100 kW per…
What Is an AI Factory? Rack-Scale AI Explained for Buyers
An AI factory is a rack-scale computing system engineered to do one thing at industrial scale: turn electricity and data into AI output (tokens). Instead of a single server, it…
GB300 NVL72: Anatomy of a 120 kW Rack-Scale AI Factory
Overview The NVIDIA GB300 NVL72 (Blackwell Ultra) marks the point where the rack, not the GPU, becomes the unit of compute. Seventy-two Blackwell Ultra (B300) GPUs and 36 Grace CPUs…