← all papers · overview

Unitsyn: A Large-scale Dataset Capable Of Enhancing The Prowess Of Large Language Models For Program Testing

Abstract

The remarkable capability of large language models (LLMs) in generating high-quality code has drawn increasing attention in the software testing community. However, existing code LLMs often demonstrate unsatisfactory capabilities in generating accurate and complete tests since they were trained on code snippets collected without differentiating between code for testing purposes and other code. In

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).