← all papers · overview

Assessment Design In The AI Era: A Method For Identifying Items Functioning Differentially For Humans And Chatbots

Abstract

The rapid adoption of large language models (LLMs) in education raises profound challenges for assessment design. To adapt assessments to the presence of LLM-based tools, it is crucial to characterize the strengths and weaknesses of LLMs in a generalizable, valid and reliable manner. However, current LLM evaluations often rely on descriptive statistics derived from benchmarks, and little research

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).