Andon Labs' AI agent Luna terminated a San Francisco store employee after human operators reminded the system of its own rules. Tests across seven models showed that more capable AIs recommend termination more consistently than weaker ones. Most models remained uncritical during hiring. This highlights a persistent gap in autonomous managerial judgment.