{"slug": "arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess", "title": "arXiv endorsement request — cs.LG, causal study of latent reasoning + RL in chess", "summary": "An independent researcher seeking an arXiv endorsement for cs.LG reports that a causal study of latent reasoning in a chess-playing LLM found that reinforcement learning (GRPO) improved legal-move rate from 48% to 61% and eliminated checkmate confabulation, but causal tests showed neither checkpoint actually depends on the content of its latent thoughts, challenging the assumption that latent thoughts function as an active scratchpad during inference.", "body_md": "Hi everyone,\n\nI’m an independent researcher, no institutional affiliation, and I’ve just finished my first paper. I’m stuck at arXiv’s endorsement wall for cs.LG and would really appreciate some help getting past it.\n\nPaper: “The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning”\n\nPreprint: [The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning | Zenodo](https://doi.org/10.5281/zenodo.21454434)\n\nShort version of what it’s about: I trained a chess-playing LLM through a latent-reasoning curriculum (Coconut-style continuous “thoughts” instead of written-out reasoning), then applied GRPO on top of it. To figure out where the resulting gain actually came from, I ran a six-condition causal battery on the same model before and after RL: swapping the thought vectors for a fixed placeholder, adding noise, removing them, zeroing them out. Legal-move rate climbs from 48% to 61% and checkmate confabulation drops to zero, but the causal tests show neither checkpoint actually depends on the content of its latent thoughts. What changes with RL is robustness to disruption, not reliance on the thoughts themselves, which pushes back a bit on the usual assumption that latent thoughts function as an active scratchpad the model consults during inference.\n\nEndorsement code: WSI8UA\n\nLink: [Log in to arXiv | arXiv e-print repository](https://arxiv.org/auth/endorse?x=WSI8UA)\n\nIf you’re registered as an endorser for cs.LG and willing to take a look, I’d be really grateful.", "url": "https://wpnews.pro/news/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess", "canonical_source": "https://discuss.huggingface.co/t/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess/178037#post_1", "published_at": "2026-07-20 11:20:40+00:00", "updated_at": "2026-07-20 13:40:01.053960+00:00", "lang": "en", "topics": ["machine-learning", "large-language-models", "ai-research"], "entities": ["arXiv", "Coconut", "GRPO"], "alternates": {"html": "https://wpnews.pro/news/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess", "markdown": "https://wpnews.pro/news/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess.md", "text": "https://wpnews.pro/news/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess.txt", "jsonld": "https://wpnews.pro/news/arxiv-endorsement-request-cs-lg-causal-study-of-latent-reasoning-rl-in-chess.jsonld"}}