← all papers · overview

MARCH: Multi-agent Reinforced Self-check For LLM Hallucination

Abstract

Hallucination remains a critical bottleneck for large language models (LLMs), undermining their reliability in real-world applications, especially in Retrieval-Augmented Generation (RAG) systems. While existing hallucination detection methods employ LLM-as-a-judge to verify LLM outputs against retrieved evidence, they suffer from inherent confirmation bias, where the verifier inadvertently reprodu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).