← all papers · overview

Ccr-bench: A Comprehensive Benchmark For Evaluating Llms On Complex Constraints, Control Flows, And Real-world Cases

Abstract

Enhancing the ability of large language models (LLMs) to follow complex instructions is critical for their deployment in real-world applications. However, existing evaluation methods often oversimplify instruction complexity as a mere additive combination of atomic constraints, failing to adequately capture the high-dimensional complexity arising from the intricate interplay of content and format,

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).