Build Wiki
Build Wiki
A reference series on building, networking, powering, and operating AI compute — for buyers and integrators sizing their next 4-GPU box, 8-GPU server, or robotics lab.
Every article is written from real Kentino builds. No filler. Opinionated where the engineering demands it. Honest about limits.
Foundational AI Server W series
If you are spec-ing a multi-GPU box, read these first. Memory, PCIe, power, thermals, storage, and the GPU shortlist.
Token Economics T series
The money math. Tokens per euro, on-prem vs cloud cost per million tokens, and when each model actually wins.
Linux / OS / Software L series
The software stack under the GPUs. Driver pinning, CUDA setup, kernel tuning, filesystems, and monitoring.
Networking N series
NVLink reality, cluster topologies (leaf-spine, fat-tree, dragonfly, switchless), latency dissection, routing, and RDMA setup in practice.
Clustering K series
When one node isn't enough. Single-vs-multi-node decision, distributed training, inference clusters, shared storage, scheduling, and failure handling.
Integration I series
Putting it all together — inference-server setup, lab network and power budgets, a reference build, and fleet deployment.
Power Delivery P series
Getting clean power to the rack. Phases, PDUs, balancing, breakers, UPS, and generators for AI compute.
Robotics R series · blog
A modern humanoid is six or seven engineering disciplines bolted together. Anatomy, sensors, compute placement, SDKs, networking, buying, and the bleeding-edge VLM world-model stack.
Case Studies C series · blog
Real Kentino builds with real measured numbers. Photographs, BOMs, benchmarks, and honest post-mortems.
New articles every Tuesday and Thursday
This wiki is a growing library — new build, networking, clustering, power, and robotics articles publish through 2026, each drawn from a real Kentino build. If you want a specific topic prioritized, write to info@kentino.com.