← all papers · overview

From Base To Conversational: Japanese Instruction Dataset And Tuning Large Language Models

Abstract

Instruction tuning is essential for large language models (LLMs) to become interactive. While many instruction tuning datasets exist in English, there is a noticeable lack in other languages. Also, their effectiveness has not been well verified in non-English languages. We construct a Japanese instruction dataset by expanding and filtering existing datasets and apply the dataset to a Japanese pre-

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).