← all papers · overview

Operational Robustness Of Llms On Code Generation

Abstract

It is now common practice in software development for large language models (LLMs) to be used to generate program code. It is desirable to evaluate the robustness of LLMs for this usage. This paper is concerned in particular with how sensitive LLMs are to variations in descriptions of the coding tasks. However, existing techniques for evaluating this robustness are unsuitable for code generation b

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).