← all papers · overview

INDICT: Code Generation With Internal Dialogues Of Critiques For Both Security And Helpfulness

Abstract

Large language models (LLMs) for code are typically trained to align with natural language instructions to closely follow their intentions and requirements. However, in many practical scenarios, it becomes increasingly challenging for these models to navigate the intricate boundary between helpfulness and safety, especially against highly complex yet potentially malicious instructions. In this wor

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).