← all papers · overview

TALEC: Teach Your LLM To Evaluate In Specific Domain With In-house Criteria By Criteria Division And Zero-shot Plus Few-shot

Abstract

With the rapid development of large language models (LLM), the evaluation of LLM becomes increasingly important. Measuring text generation tasks such as summarization and article creation is very difficult. Especially in specific application domains (e.g., to-business or to-customer service), in-house evaluation criteria have to meet not only general standards (correctness, helpfulness and creativ

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).