{"slug": "request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory", "title": "Request for arXiv cs.LG endorsement – interpretability paper (residual trajectory geometry)", "summary": "An independent researcher is seeking endorsement on arXiv for a cs.LG paper on interpretability of residual trajectory geometry in a chess-playing LLM trained with a latent-reasoning curriculum and GRPO. The study found that legal-move rate improved from 48% to 61% and checkmate confabulation dropped to zero after reinforcement learning, but causal tests showed neither checkpoint depends on the content of its latent thoughts, challenging the assumption that latent thoughts function as an active scratchpad.", "body_md": "Hi,\n\nThis seems like a really great research, I hope you received the endorsement for your paper!\n\nI’m an independent researcher, no institutional affiliation, and I’ve just finished my first paper as well. I’m stuck at arXiv’s endorsement wall for cs.LG too and would really appreciate some help getting past it.\n\nShort version of what it’s about: I trained a chess-playing LLM through a latent-reasoning curriculum (Coconut-style continuous “thoughts” instead of written-out reasoning), then applied GRPO on top of it. To figure out where the resulting gain actually came from, I ran a six-condition causal battery on the same model before and after RL: swapping the thought vectors for a fixed placeholder, adding noise, removing them, zeroing them out. Legal-move rate climbs from 48% to 61% and checkmate confabulation drops to zero, but the causal tests show neither checkpoint actually depends on the content of its latent thoughts. What changes with RL is robustness to disruption, not reliance on the thoughts themselves, which pushes back a bit on the usual assumption that latent thoughts function as an active scratchpad the model consults during inference.\n\nEndorsement code: WSI8UA\n\nLink: [Log in to arXiv | arXiv e-print repository](https://arxiv.org/auth/endorse?x=WSI8UA)\n\nIf you’re registered as an endorser for cs.LG and willing to take a look, I’d be really grateful. And if you’re not able to, no worries at all, thanks for reading this far either way!", "url": "https://wpnews.pro/news/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory", "canonical_source": "https://discuss.huggingface.co/t/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory-geometry/173697#post_2", "published_at": "2026-07-20 11:27:01+00:00", "updated_at": "2026-07-20 13:39:55.511462+00:00", "lang": "en", "topics": ["artificial-intelligence", "machine-learning", "large-language-models", "ai-research"], "entities": ["arXiv"], "alternates": {"html": "https://wpnews.pro/news/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory", "markdown": "https://wpnews.pro/news/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory.md", "text": "https://wpnews.pro/news/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory.txt", "jsonld": "https://wpnews.pro/news/request-for-arxiv-cs-lg-endorsement-interpretability-paper-residual-trajectory.jsonld"}}