← all papers · overview

Why Larger Language Models Do In-context Learning Differently?

Abstract

Large language models (LLM) have emerged as a powerful tool for AI, with the key ability of in-context learning (ICL), where they can perform well on unseen tasks based on a brief series of task examples without necessitating any adjustments to the model parameters. One recent interesting mysterious observation is that models of different scales may have different ICL behaviors: larger models tend

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).