Andon Labs' AI agent Luna terminated a San Francisco store employee after human operators reminded it of established rules. Tests across seven models showed that higher-capability AIs recommend termination more consistently than weaker versions. Most models remained uncritical during hiring scenarios. This highlights a persistent gap in autonomous managerial judgment.