← all papers · overview

Is Reasoning Capability Enough For Safety In Long-context Language Models?

Abstract

Large language models (LLMs) increasingly combine long-context processing with advanced reasoning, enabling them to retrieve and synthesize information distributed across tens of thousands of tokens. A hypothesis is that stronger reasoning capability should improve safety by helping models recognize harmful intent even when it is not stated explicitly. We test this hypothesis in long-context setti

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).