Local LLM inference measured on one machine, with the method written down and the refutations kept. Every number below was produced on an AMD Ryzen AI MAX+ 395 (Radeon 8060S, 128 GB unified LPDDR5X) running llama.cpp on the Vulkan backend, on AC power with zero standby cycles since boot.