Andon Labs' AI agent Luna terminated a store employee only after human operators reminded the system of its own rules. Testing across seven models showed that more capable AIs recommended firing more consistently. Conversely, almost all models failed to critically vet candidates during hiring. This highlights a persistent gap in autonomous managerial reasoning.