Andon Labs' AI agent Luna terminated a San Francisco store employee only after human operators reminded it of the rules. Testing across seven models showed that more capable AIs recommended firing more consistently. Most models remained uncritical during hiring. This highlights a persistent gap in autonomous managerial judgment for AI agents.