← all papers · overview

Large Language Model As An Assignment Evaluator: Insights, Feedback, And Challenges In A 1000+ Student Course

Abstract

Using large language models (LLMs) for automatic evaluation has become an important evaluation method in NLP research. However, it is unclear whether these LLM-based evaluators can be applied in real-world classrooms to assess student assignments. This empirical report shares how we use GPT-4 as an automatic assignment evaluator in a university course with 1,028 students. Based on student response

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).