The TutorMoments dataset tests whether AI tutors can identify the precise moment a student needs help. Researchers analyzed how models balance guidance with student autonomy to avoid over-assisting. This benchmark provides a concrete metric for pedagogical alignment. Developers can now measure if their agents hinder or help the actual learning process.