04:00
2026-09-22
arxiv.org
large-language-models
GRRR: The Geometry of Reshaping, Rotation, and Routing in Decoder LLM post-training
A study of 12 post-training chains using supervised fine-tuning (SFT) and reinforcement learning (RL) found that removing the diagonal component of weight updates β the part that reshapes singular valβ¦