LLMTuning an LLM workstation: what actually moved the needle on CPU and GPU
A practical account of tuning a shared Ryzen 9950X / dual RTX 5090 box that runs CI, local LLM inference, and production web at once - which CPU and GPU settings measurably changed throughput and thermals, which ones did nothing, and the drift bug that made a frequency cap silently protect nothing.
August 15, 202612 min read