Free manuals

Deployment guides: local LLMs on real hardware, step by step

Free manuals written only for stacks we actually deployed and serve from. Every trap marked is one we hit ourselves; performance numbers stay in the measured reports, one link away.

Strix Halo (Ryzen AI Max+ 395)

DGX Spark (GB10)

Apple Silicon

No guide yet — honestly. The M1 Max and M4 Pro machines are declared lanes on our testbed with zero published runs. The guide will be written while deploying on them for real, not before.

Consumer GPU

No guide yet — honestly. No CUDA-lane runs published yet. When the discrete-GPU lane starts, its guide will be written from the deployment, not from other people’s posts.