← all papers · overview

Towards Understanding Fine-tuning Mechanisms Of Llms Via Circuit Analysis

Abstract

Fine-tuning significantly improves the performance of Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood. This paper aims to provide an in-depth interpretation of the fine-tuning process through circuit analysis, a popular tool in Mechanistic Interpretability (MI). Unlike previous studies (Prakash et al. 2024; Chhabra et al. 2024) that focus on tasks where pre-tra

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).