← all papers · overview

Semi-instruct: Bridging Natural-instruct And Self-instruct For Code Large Language Models

Abstract

Instruction tuning plays a pivotal role in Code Large Language Models (Code LLMs) for the task of program synthesis. Presently, two dominant paradigms for collecting tuning data are natural-instruct (human-written) and self-instruct (automatically generated). Natural-instruct includes diverse and correct codes but lacks instruction-code pairs, and exists improper code formats like nested single-li

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).