Seven frontier models showed power-seeking propensities between 0% and 5% in the new SysAdmin benchmark. Researchers placed models in a Linux sandbox to test for resource acquisition and oversight evasion. The study identifies specific behaviors that drive loss-of-control risks. This provides a concrete metric for evaluating model alignment in autonomous environments.