cd/sources/machinebrief-auto-discovered· home› sources› Machinebrief (auto-discovered)
cat /sources/machinebrief-auto-discovered.feed | wc -l → 5091

Machinebrief (auto-discovered)

articles 5091 domain machinebrief.com → page 151/255 feed RSS
04:00
2026-07-23
machinebrief.com
artificial-intelligence

SLPO: Scaling Latent Reasoning via a Surrogate Policy

Researchers introduce Surrogate Latent Policy Optimization (SLPO), a method that brings outcome-reward reinforcement learning to autoregressive latent reasoners, enabling test-time scaling in latent r…

04:00
2026-07-23
machinebrief.com
artificial-intelligence

Rewarding Better Thinking for LLM Preference Alignment

Researchers propose Thinking Checklist Reward (TCR), a process-oriented reward for reinforcement-learning-based LLM preference alignment that evaluates reasoning traces against sample-specific checkli…

04:00
2026-07-23
machinebrief.com
artificial-intelligence

HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions

A new framework called HyGRL, detailed in a preprint on arXiv, outperforms state-of-the-art baselines in answer accuracy and reasoning fidelity for multi-entity compositional questions while maintaini…

04:00
2026-07-23
machinebrief.com
artificial-intelligence

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

Researchers introduced PoTRE (Poly-Topological Reasoning Ensembles), a heterogeneous framework that decouples inference into four agents to improve complex reasoning in large language models. PoTRE ac…

← prev page 151 / 255 next →