Old Habits Die Hard: How Conversational History Geometrically Traps LLMs
Computer Science Department, Technion - Israel Institute of Technology · University of Oxford / Martian · University of Zagreb, FER · Technion, Technion · University of Edinburgh
PDF 由论文原始站点提供,PaperCompass 不保存论文文件。
摘要
How does the conversational past of large language models (LLMs) influence their future performance? Recent work suggests that LLMs are affected by their conversational history in unexpected ways. For instance, hallucinations in prior interactions may influence subsequent model responses. In this work, we introduce History Echoes, a framework that investigates how conversational history biases subsequent generations. The framework explores this bias from two perspectives: probabilistically, we model conversations as Markov chains to quantify state consistency; geometrically, we measure the consistency of consecutive hidden representations. Across three model families and six datasets spanning diverse phenomena, our analysis reveals a strong correlation between the two perspectives. By bridging these perspectives, we demonstrate that behavioral persistence manifests as a geometric trap, where gaps in the latent space confine the model's trajectory.