Home lab enthusiasts are reporting significant thermal drift when running multi-GPU setups for local LLMs. These temperature fluctuations cause clock speed instability and inconsistent inference latency. Users on Lobste.rs suggest custom fan curves and improved spacing to mitigate the heat. Practitioners should prioritize active cooling to maintain stable token generation speeds.