Transformer representations evolve through learned additive transformations that either preserve their current direction or redirect it. We study this evolution as a functional geometry, decomposing learned updates into parallel and perpendicular components. Across pretrained models, we find substan
SPICE: Simple Polysemantic Feature Interpretation via Clustering-based Explanation