← all papers · overview

Training An Llm-as-a-judge Model: Pipeline, Insights, And Practical Lessons

Abstract

The rapid advancement of large language models (LLMs) has opened new possibilities for their adoption as evaluative judges. This paper introduces Themis, a fine-tuned LLM judge that delivers sophisticated context-aware evaluations. We provide a comprehensive overview of the development pipeline for Themis, highlighting its scenario-dependent evaluation prompts and two novel methods for controlled

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).