← all papers · overview

Ahp-powered LLM Reasoning For Multi-criteria Evaluation Of Open-ended Responses

Abstract

Question answering (QA) tasks have been extensively studied in the field of natural language processing (NLP). Answers to open-ended questions are highly diverse and difficult to quantify, and cannot be simply evaluated as correct or incorrect, unlike close-ended questions with definitive answers. While large language models (LLMs) have demonstrated strong capabilities across various tasks, they e

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).