← all papers · overview

Confusion-aware Rubric Optimization For Llm-based Automated Grading

Abstract

Accurate and unambiguous guidelines are critical for large language model (LLM) based graders, yet manually crafting these prompts is often sub-optimal as LLMs can misinterpret expert guidelines or lack necessary domain specificity. Consequently, the field has moved toward automated prompt optimization to refine grading guidelines without the burden of manual trial and error. However, existing fra

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).