The TutorMoments dataset tests whether AI tutors recognize the precise moment a student needs help. Researchers analyzed how models balance guidance with student autonomy to avoid over-assisting. This benchmark forces developers to move beyond simple correctness. It provides a concrete metric for improving pedagogical timing in LLM-driven educational tools.