← all papers · overview

Teachbench: A Syllabus-grounded Framework For Evaluating Teaching Ability In Large Language Models

Abstract

Large language models (LLMs) show promise as teaching assistants, yet their teaching capability remains insufficiently evaluated. Existing benchmarks mainly focus on problem-solving or problem-level guidance, leaving knowledge-centered teaching underexplored. We propose a syllabus-grounded evaluation framework that measures LLM teaching capability via student performance improvement after multi-tu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).