← all papers · overview

Unifying AI Tutor Evaluation: An Evaluation Taxonomy For Pedagogical Ability Assessment Of Llm-powered AI Tutors

Abstract

In this paper, we investigate whether current state-of-the-art large language models (LLMs) are effective as AI tutors and whether they demonstrate pedagogical abilities necessary for good AI tutoring in educational dialogues. Previous efforts towards evaluation have been limited to subjective protocols and benchmarks. To bridge this gap, we propose a unified evaluation taxonomy with eight pedagog

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).