← all papers · overview

Re-examining Learning Linear Functions In Context

Abstract

In-context learning (ICL) has emerged as a powerful paradigm for easily adapting Large Language Models (LLMs) to various tasks. However, our understanding of how ICL works remains limited. We explore a simple model of ICL in a controlled setup with synthetic training data to investigate ICL of univariate linear functions. We experiment with a range of GPT-2-like transformer models trained from scr

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).