Amazon Web Services ordered engineers to conserve CPU cycles after AI workloads spiked server wait times. While GPUs dominate model inference, agentic AI systems now demand more general-purpose compute for autonomous coordination. This shift catches cloud providers off-guard. Infrastructure teams must now optimize CPU allocation to prevent bottlenecks in agent-driven workflows.