GPUSmith

Articles (English) - Page 4 / 5

What Is a Good GPU Utilization Percentage? The 80% Target

What Is a Good GPU Utilization Percentage? The 80% Target

A 2026 data-driven guide to good GPU utilization percentages, covering NVIDIA's 80% SM Activity benchmark, Cast AI's 5% enterprise average, Model FLOPs Utilization, and how to monitor and raise GPU utilization.

7/22/2026• 42 min read
gpu utilizationgpu utilization percentagenvidia-smi
Kimi K2 Hardware Requirements: VRAM, Nodes and Cost to Run

Kimi K2 Hardware Requirements: VRAM, Nodes and Cost to Run

A 2026 guide to Kimi K2 hardware requirements covering VRAM and GPU needs, quantization tiers from 245GB to 1.09TB, Moonshot's 16-GPU node configuration, and self-hosting versus API cost.

7/22/2026• 41 min read
kimi k2kimi k2 hardware requirementskimi k2 vram
Air-Gapped LLM Deployment: Running Frontier Models Offline

Air-Gapped LLM Deployment: Running Frontier Models Offline

A 2026 analyst guide to air-gapped LLM deployment: DoD IL5/IL6 and FedRAMP frameworks, open-weight model licensing, GPU hardware costs from $30K H200s to $467K DGX clusters, and update pipelines.

7/22/2026• 42 min read
air-gapped llm deploymentoffline ai modelsopen-weight models
DeepSeek R1 Self-Hosting Cost per Million Tokens (2026 Guide)

DeepSeek R1 Self-Hosting Cost per Million Tokens (2026 Guide)

Analyzes DeepSeek R1 self-hosting costs per million tokens in 2026: GPU hardware requirements (8x H200 to 16x H100), benchmarked $0.11 to $16 per million token ranges, API pricing, and real deployment case studies.

7/22/2026• 29 min read
deepseek r1self-hosting costllm pricing
SXM vs PCIe GPUs: Form Factor, Cooling and Use Cases

SXM vs PCIe GPUs: Form Factor, Cooling and Use Cases

A 2026 comparison of SXM and PCIe NVIDIA GPUs covering HGX architecture, NVLink bandwidth, cooling requirements, H100 and A100 specs, cloud pricing, and real-world deployment case studies.

7/22/2026• 39 min read
SXMPCIeGPU
NVIDIA Nemotron 3 Ultra Hardware Requirements: VRAM and Cost

NVIDIA Nemotron 3 Ultra Hardware Requirements: VRAM and Cost

A 2026 analyst guide to NVIDIA Nemotron 3 Ultra hardware requirements: VRAM, minimum GPU counts for BF16 and NVFP4, DGX Spark and DGX Station paths, node configurations, and per-token self-hosting versus API costs.

7/22/2026• 36 min read
nemotron 3 ultranvidia nemotrongpu requirements
InfiniBand vs Spectrum-X vs RoCE vs Ethernet for AI Clusters

InfiniBand vs Spectrum-X vs RoCE vs Ethernet for AI Clusters

A 2026 analyst comparison of InfiniBand, NVIDIA Spectrum-X, RoCEv2, and Ultra Ethernet for AI clusters, with Dell'Oro and 650 Group market data, latency benchmarks, and case studies from Meta, xAI, and Oracle.

7/22/2026• 39 min read
infinibandspectrum-xroce
400 VDC vs 800 VDC vs 415V AC: AI Datacenter Power Architecture

400 VDC vs 800 VDC vs 415V AC: AI Datacenter Power Architecture

A 2026 comparison of 400 VDC, 800 VDC, and 415V AC data center power architectures, covering NVIDIA's monopolar Kyber design, OCP's bipolar Diablo 400 spec, copper and efficiency data, and rollout timelines.

7/21/2026• 35 min read
800vdc400vdc415v ac
How to Build an AI Datacenter: GPU Cluster Engineering Guide

How to Build an AI Datacenter: GPU Cluster Engineering Guide

A 2026 engineering guide to building an AI datacenter, covering GPU cluster architecture, InfiniBand vs Ethernet, power and cooling density, cost per megawatt, and grid interconnection bottlenecks.

7/21/2026• 33 min read
ai datacentergpu cluster architectureinfiniband vs ethernet
NVIDIA AI Server OEM Comparison: Supermicro vs Dell vs HPE

NVIDIA AI Server OEM Comparison: Supermicro vs Dell vs HPE

2026 comparison of NVIDIA AI server OEMs: Dell, Supermicro, HPE, Lenovo, and Foxconn. Covers IDC market share, GB200/GB300 NVL72 racks, pricing, and 5 named deployments.

7/21/2026• 34 min read
nvidia ai server oemsupermicro vs dellhpe nvidia ai server
800 VDC Datacenter Power: Why AI Racks Are Going High-Voltage DC

800 VDC Datacenter Power: Why AI Racks Are Going High-Voltage DC

An 2026 analyst breakdown of 800 VDC datacenter power architecture: why NVIDIA and OCP are moving AI racks to high-voltage DC, monopolar vs bipolar standards, costs, and safety gaps.

7/21/2026• 37 min read
800 vdchvdc power architecturedata center power distribution
NVIDIA NVLink vs AMD Infinity Fabric: 2026 Interconnect Guide

NVIDIA NVLink vs AMD Infinity Fabric: 2026 Interconnect Guide

Compares NVIDIA NVLink and AMD Infinity Fabric bandwidth, NVSwitch, UALink, and Ultra Ethernet in 2026, with GB200/GB300 NVL72 specs, MI300X data, and pricing.

7/21/2026• 38 min read
nvidia nvlinkamd infinity fabricnvswitch
Who Owns ZT Systems Now After the AMD-Sanmina Split?

Who Owns ZT Systems Now After the AMD-Sanmina Split?

Explains who owns ZT Systems as of 2026: AMD kept its design team while Sanmina bought its manufacturing plants for up to $3 billion, closing October 2025.

7/21/2026• 35 min read
zt systemsamdsanmina
Is Kaytus on the Entity List? Inspur's Rebrand Explained

Is Kaytus on the Entity List? Inspur's Rebrand Explained

A 2026 analyst report tracing whether Kaytus and Aivres are on the US Entity List, their ownership by Entity Listed Inspur Group and IEIT Systems, the suspended BIS Affiliates Rule, and procurement compliance guidance for buyers.

7/21/2026• 37 min read
kaytusinspuraivres
NVIDIA Groq 3 LPU Explained: Vera Rubin Inference Chip

NVIDIA Groq 3 LPU Explained: Vera Rubin Inference Chip

A 2026 technical explainer on the NVIDIA Groq 3 LPU: the $20 billion Groq licensing deal, Vera Rubin platform architecture, LP30 chip specs, and how LPUs compare to GPUs and TPUs.

7/20/2026• 39 min read
nvidia groq 3 lpugroq lpuvera rubin platform
Cerebras Wafer-Scale Engine: WSE-3 Architecture Explained (2026)

Cerebras Wafer-Scale Engine: WSE-3 Architecture Explained (2026)

Deep dive into the Cerebras Wafer-Scale Engine WSE-3 architecture, CS-3 system specs, inference pricing, NVIDIA H100/B200 comparisons, and the May 2026 IPO, covering 2026 data.

7/20/2026• 38 min read
cerebras wafer-scale enginewafer-scale enginecerebras wse-3
NVIDIA vs AMD vs Intel vs Cerebras AI Accelerator Comparison

NVIDIA vs AMD vs Intel vs Cerebras AI Accelerator Comparison

A 2026 analyst comparison of NVIDIA, AMD, Intel, and Cerebras AI accelerators covering architecture, MLPerf benchmarks, power efficiency, pricing, market share, and named deployments.

7/19/2026• 36 min read
ai accelerator comparisonnvidia vs amd vs intelcerebras wse-3
GLM-5.2 Hardware Requirements: VRAM, Nodes, and Costs

GLM-5.2 Hardware Requirements: VRAM, Nodes, and Costs

A 2026 guide to GLM-5.2 hardware requirements covering VRAM by quantization, minimum RAM, multi-GPU node sizing, self-hosting costs, and API pricing versus DeepSeek.

7/19/2026• 37 min read
glm-5.2 hardware requirementsglm-5.2 vram requirementsglm-5.2 gpu requirements
Fiber Optic Cabling Best Practices in the Data Center

Fiber Optic Cabling Best Practices in the Data Center

A 2026 analyst guide to fiber optic cabling best practices in the data center, covering TIA-942, BICSI 002, MPO vs MTP connectors, MPO polarity methods A/B/C, and fiber loss budget calculations.

7/18/2026• 44 min read
fiber optic cablingdata center cablingTIA-942
Hot Aisle vs Cold Aisle Containment for GPU Racks (2026)

Hot Aisle vs Cold Aisle Containment for GPU Racks (2026)

Compares hot aisle vs cold aisle containment for data centers in 2026, with PUE data, GPU rack density thresholds up to 120 kW, liquid cooling market data, and 6 real-world case studies.

7/18/2026• 40 min read
hot aisle containmentcold aisle containmentdata center cooling
GPU Data Center Cooling Design: CRAC, CRAH and Liquid Cooling

GPU Data Center Cooling Design: CRAC, CRAH and Liquid Cooling

A 2026 analyst guide to GPU data center cooling design covering CRAC vs CRAH, hot/cold aisle containment, direct-to-chip liquid cooling, chiller plants, and capacity calculation for AI racks.

7/17/2026• 43 min read
gpu data center cooling designliquid cooling data centercrac vs crah