TutorMoments: Do AI tutors know when to help and when to hold back?
Hugging Face 3 weeks ago 35 ● 2 sources
Allen AI introduced TutorMoments, a framework measuring whether large language models can appropriately balance giving help versus pushing students to do their own work during tutoring. Testing seven LLMs on 462 real math tutoring transcripts with 1,500 teacher-annotated decision points, researchers found models tend to over-help when given generic prompts but improve when instructions explicitly describe the scaffolding-versus-rigor trade-off. Despite improvements, no model matched human tutors' ability to make contextually appropriate pedagogical choices, and the framework is being released openly for researchers and AI tutoring developers.