← all papers · overview

Kun: Answer Polishment For Chinese Self-alignment With Instruction Back-translation

Abstract

In this paper, we introduce Kun, a novel approach for creating high-quality instruction-tuning datasets for large language models (LLMs) without relying on manual annotations. Adapting a self-training algorithm based on instruction back-translation and answer polishment, Kun leverages unlabelled data from diverse sources such as Wudao, Wanjuan, and SkyPile to generate a substantial dataset of over

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).