Depth-Wise Emergence of Prediction-Centric Geometry in Large Language Models

Shahar Haim; Daniel C McNamee

arXiv:2602.04931·cs.LG·February 6, 2026

Depth-Wise Emergence of Prediction-Centric Geometry in Large Language Models

Shahar Haim, Daniel C McNamee

PDF

Open Access

TL;DR

This paper reveals how large language models transition from processing context to predicting tokens through a depth-wise geometric reorganization of their internal representations, enabling causal control over predictions.

Contribution

It introduces a unified geometric framework to analyze the depth-wise transition in LLMs and demonstrates how late-layer representations facilitate prediction control.

Findings

01

Representation geometry parametrizes prediction similarity.

02

Representation norms encode context-specific information.

03

A mechanistic account of context-to-prediction transformation.

Abstract

We show that decoder-only large language models exhibit a depth-wise transition from context-processing to prediction-forming phases of computation accompanied by a reorganization of representational geometry. Using a unified framework combining geometric analysis with mechanistic intervention, we demonstrate that late-layer representations implement a structured geometric code that enables selective causal control over token prediction. Specifically, angular organization of the representation geometry parametrizes prediction distributional similarity, while representation norms encode context-specific information that does not determine prediction. Together, these results provide a mechanistic-geometric account of the dynamics of transforming context into predictions in LLMs.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsTopic Modeling · Machine Learning in Materials Science · Multimodal Machine Learning Applications