← all papers · overview

Remodetect: Reward Models Recognize Aligned Llm's Generations

Abstract

The remarkable capabilities and easy accessibility of large language models (LLMs) have significantly increased societal risks (e.g., fake news generation), necessitating the development of LLM-generated text (LGT) detection methods for safe usage. However, detecting LGTs is challenging due to the vast number of LLMs, making it impractical to account for each LLM individually; hence, it is crucial

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).