← all papers · overview

L2ceval: Evaluating Language-to-code Generation Capabilities Of Large Language Models

Abstract

Recently, large language models (LLMs), especially those that are pretrained on code, have demonstrated strong capabilities in generating programs from natural language inputs in a few-shot or even zero-shot manner. Despite promising results, there is a notable lack of a comprehensive evaluation of these models language-to-code generation capabilities. Existing studies often focus on specific task

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).