GPUSmith

Articles (English) - Page 2 / 3

Why 800 VDC Data Center Power Delivery Is Replacing 48V

Why 800 VDC Data Center Power Delivery Is Replacing 48V

A 2026 technical explainer on why AI data centers are shifting from 48V/54V to 800 VDC power delivery, covering NVIDIA's Kyber rack architecture, copper and efficiency gains, vendor timelines from Vertiv, Eaton, and Delta, and named deployments.

7/22/202637 min read
800 vdc800 vdc data centerdata center power delivery
H200 vs B200 vs B300 vs MI325X: Full 2026 GPU Comparison

H200 vs B200 vs B300 vs MI325X: Full 2026 GPU Comparison

A data-driven comparison of NVIDIA's H200, B200, B300 and AMD's Instinct MI325X covering specs, MLPerf benchmarks, cloud pricing, availability, real-world deployments and buyer guidance for AI infrastructure decisions in 2026.

7/22/202638 min read
h200b200b300
What Is a Good GPU Utilization Percentage? The 80% Target

What Is a Good GPU Utilization Percentage? The 80% Target

A 2026 data-driven guide to good GPU utilization percentages, covering NVIDIA's 80% SM Activity benchmark, Cast AI's 5% enterprise average, Model FLOPs Utilization, and how to monitor and raise GPU utilization.

7/22/202642 min read
gpu utilizationgpu utilization percentagenvidia-smi
Kimi K2 Hardware Requirements: VRAM, Nodes and Cost to Run

Kimi K2 Hardware Requirements: VRAM, Nodes and Cost to Run

A 2026 guide to Kimi K2 hardware requirements covering VRAM and GPU needs, quantization tiers from 245GB to 1.09TB, Moonshot's 16-GPU node configuration, and self-hosting versus API cost.

7/22/202641 min read
kimi k2kimi k2 hardware requirementskimi k2 vram
Air-Gapped LLM Deployment: Running Frontier Models Offline

Air-Gapped LLM Deployment: Running Frontier Models Offline

A 2026 analyst guide to air-gapped LLM deployment: DoD IL5/IL6 and FedRAMP frameworks, open-weight model licensing, GPU hardware costs from $30K H200s to $467K DGX clusters, and update pipelines.

7/22/202642 min read
air-gapped llm deploymentoffline ai modelsopen-weight models
DeepSeek R1 Self-Hosting Cost per Million Tokens (2026 Guide)

DeepSeek R1 Self-Hosting Cost per Million Tokens (2026 Guide)

Analyzes DeepSeek R1 self-hosting costs per million tokens in 2026: GPU hardware requirements (8x H200 to 16x H100), benchmarked $0.11 to $16 per million token ranges, API pricing, and real deployment case studies.

7/22/202629 min read
deepseek r1self-hosting costllm pricing
SXM vs PCIe GPUs: Form Factor, Cooling and Use Cases

SXM vs PCIe GPUs: Form Factor, Cooling and Use Cases

A 2026 comparison of SXM and PCIe NVIDIA GPUs covering HGX architecture, NVLink bandwidth, cooling requirements, H100 and A100 specs, cloud pricing, and real-world deployment case studies.

7/22/202639 min read
SXMPCIeGPU
NVIDIA Nemotron 3 Ultra Hardware Requirements: VRAM and Cost

NVIDIA Nemotron 3 Ultra Hardware Requirements: VRAM and Cost

A 2026 analyst guide to NVIDIA Nemotron 3 Ultra hardware requirements: VRAM, minimum GPU counts for BF16 and NVFP4, DGX Spark and DGX Station paths, node configurations, and per-token self-hosting versus API costs.

7/22/202636 min read
nemotron 3 ultranvidia nemotrongpu requirements
InfiniBand vs Spectrum-X vs RoCE vs Ethernet for AI Clusters

InfiniBand vs Spectrum-X vs RoCE vs Ethernet for AI Clusters

A 2026 analyst comparison of InfiniBand, NVIDIA Spectrum-X, RoCEv2, and Ultra Ethernet for AI clusters, with Dell'Oro and 650 Group market data, latency benchmarks, and case studies from Meta, xAI, and Oracle.

7/22/202639 min read
infinibandspectrum-xroce
400 VDC vs 800 VDC vs 415V AC: AI Datacenter Power Architecture

400 VDC vs 800 VDC vs 415V AC: AI Datacenter Power Architecture

A 2026 comparison of 400 VDC, 800 VDC, and 415V AC data center power architectures, covering NVIDIA's monopolar Kyber design, OCP's bipolar Diablo 400 spec, copper and efficiency data, and rollout timelines.

7/21/202635 min read
800vdc400vdc415v ac
How to Build an AI Datacenter: GPU Cluster Engineering Guide

How to Build an AI Datacenter: GPU Cluster Engineering Guide

A 2026 engineering guide to building an AI datacenter, covering GPU cluster architecture, InfiniBand vs Ethernet, power and cooling density, cost per megawatt, and grid interconnection bottlenecks.

7/21/202633 min read
ai datacentergpu cluster architectureinfiniband vs ethernet
NVIDIA AI Server OEM Comparison: Supermicro vs Dell vs HPE

NVIDIA AI Server OEM Comparison: Supermicro vs Dell vs HPE

2026 comparison of NVIDIA AI server OEMs: Dell, Supermicro, HPE, Lenovo, and Foxconn. Covers IDC market share, GB200/GB300 NVL72 racks, pricing, and 5 named deployments.

7/21/202634 min read
nvidia ai server oemsupermicro vs dellhpe nvidia ai server
800 VDC Datacenter Power: Why AI Racks Are Going High-Voltage DC

800 VDC Datacenter Power: Why AI Racks Are Going High-Voltage DC

An 2026 analyst breakdown of 800 VDC datacenter power architecture: why NVIDIA and OCP are moving AI racks to high-voltage DC, monopolar vs bipolar standards, costs, and safety gaps.

7/21/202637 min read
800 vdchvdc power architecturedata center power distribution
NVIDIA NVLink vs AMD Infinity Fabric: 2026 Interconnect Guide

NVIDIA NVLink vs AMD Infinity Fabric: 2026 Interconnect Guide

Compares NVIDIA NVLink and AMD Infinity Fabric bandwidth, NVSwitch, UALink, and Ultra Ethernet in 2026, with GB200/GB300 NVL72 specs, MI300X data, and pricing.

7/21/202638 min read
nvidia nvlinkamd infinity fabricnvswitch
Who Owns ZT Systems Now After the AMD-Sanmina Split?

Who Owns ZT Systems Now After the AMD-Sanmina Split?

Explains who owns ZT Systems as of 2026: AMD kept its design team while Sanmina bought its manufacturing plants for up to $3 billion, closing October 2025.

7/21/202635 min read
zt systemsamdsanmina
Is Kaytus on the Entity List? Inspur's Rebrand Explained

Is Kaytus on the Entity List? Inspur's Rebrand Explained

A 2026 analyst report tracing whether Kaytus and Aivres are on the US Entity List, their ownership by Entity Listed Inspur Group and IEIT Systems, the suspended BIS Affiliates Rule, and procurement compliance guidance for buyers.

7/21/202637 min read
kaytusinspuraivres
NVIDIA Groq 3 LPU Explained: Vera Rubin Inference Chip

NVIDIA Groq 3 LPU Explained: Vera Rubin Inference Chip

A 2026 technical explainer on the NVIDIA Groq 3 LPU: the $20 billion Groq licensing deal, Vera Rubin platform architecture, LP30 chip specs, and how LPUs compare to GPUs and TPUs.

7/20/202639 min read
nvidia groq 3 lpugroq lpuvera rubin platform
Cerebras Wafer-Scale Engine: WSE-3 Architecture Explained (2026)

Cerebras Wafer-Scale Engine: WSE-3 Architecture Explained (2026)

Deep dive into the Cerebras Wafer-Scale Engine WSE-3 architecture, CS-3 system specs, inference pricing, NVIDIA H100/B200 comparisons, and the May 2026 IPO, covering 2026 data.

7/20/202638 min read
cerebras wafer-scale enginewafer-scale enginecerebras wse-3
NVIDIA vs AMD vs Intel vs Cerebras AI Accelerator Comparison

NVIDIA vs AMD vs Intel vs Cerebras AI Accelerator Comparison

A 2026 analyst comparison of NVIDIA, AMD, Intel, and Cerebras AI accelerators covering architecture, MLPerf benchmarks, power efficiency, pricing, market share, and named deployments.

7/19/202636 min read
ai accelerator comparisonnvidia vs amd vs intelcerebras wse-3
GLM-5.2 Hardware Requirements: VRAM, Nodes, and Costs

GLM-5.2 Hardware Requirements: VRAM, Nodes, and Costs

A 2026 guide to GLM-5.2 hardware requirements covering VRAM by quantization, minimum RAM, multi-GPU node sizing, self-hosting costs, and API pricing versus DeepSeek.

7/19/202637 min read
glm-5.2 hardware requirementsglm-5.2 vram requirementsglm-5.2 gpu requirements
Fiber Optic Cabling Best Practices in the Data Center

Fiber Optic Cabling Best Practices in the Data Center

A 2026 analyst guide to fiber optic cabling best practices in the data center, covering TIA-942, BICSI 002, MPO vs MTP connectors, MPO polarity methods A/B/C, and fiber loss budget calculations.

7/18/202644 min read
fiber optic cablingdata center cablingTIA-942