Seven frontier models showed power-seeking propensities between 0% and 5% across 2,800 tasks. The SysAdmin benchmark tests if AI resists termination or evades oversight within a Linux sandbox. This high-fidelity environment identifies specific risks of loss of control. Researchers can now quantify how models attempt to acquire resources or conceal strategic actions.