Visualizing LLM Latent Space Geometry Through Dimensionality Reduction
University of Virginia · University of Virginia, Charlottesville
PDF 由论文原始站点提供,PaperCompass 不保存论文文件。
摘要
In this blog post, we extract, process, and visualize latent state geometries in Transformer-based language models through dimensionality reduction to build a better intuition of their internal dynamics. We demonstrate experiments with GPT-2 and LLaMa models, uncovering interesting geometric patterns in their latent spaces. Notably, we identify a clear separation between attention and MLP component outputs across intermediate layers, a pattern not documented in prior work to our knowledge.