← 返回论文检索
ICLR 2026PosterAccept (Poster)

Comparing the learning dynamics of in-context learning and fine-tuning in language models

Basile Confavreux, Aaditya Singh, Jin Hwa Lee, Amaury Sabran, Andrew Saxe

University College London · University College London, University of London · Waveforms AI · Harvard University

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

Pretrained language models can acquire novel tasks either through in-context learning (ICL)---adapting behavior via activations without weight updates---or through supervised fine-tuning (SFT), where parameters are explicitly updated. Prior work has reported differences in their generalization performance and inductive biases, but the origins of these differences remain poorly understood. In this work, we treat ICL and SFT as distinct learning algorithms and directly compare the learning dynamics they induce across medium-sized models, analyzing both the evolution of their inductive biases and the underlying internal representations. We find that ICL preserves rich input representations but imposes stronger priors inherited from pretraining, whereas SFT suppresses task-irrelevant features---potentially explaining its weaker generalization in few-shot regimes. These results highlight a mechanistic distinction between context-driven and weight-driven learning.