← all papers · overview

Analyzing The Performance Of Large Language Models On Code Summarization

Abstract

Large language models (LLMs) such as Llama 2 perform very well on tasks that involve both natural language and source code, particularly code summarization and code generation. We show that for the task of code summarization, the performance of these models on individual examples often depends on the amount of (subword) token overlap between the code and the corresponding reference natural languag

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).