Gradients Leakage in Split Language Models
A paper submitted to arXiv on 2 October 2026 reports that an observer at the split in split learning can rebuild most of a client's text from the traffic, with gradients adding measurable leakage: on …
A paper submitted to arXiv on 2 October 2026 reports that an observer at the split in split learning can rebuild most of a client's text from the traffic, with gradients adding measurable leakage: on …
A 2 October 2026 arXiv paper evaluating fifteen prompt-injection detectors, including Meta's Prompt Guard 2, found that public benchmark scores transfer poorly to real LLM agent deployments: the best …
A paired-replay testbed measuring task-scoped authorization in tool-using LLM agents found that broad-bearer credentials permitted harmful tool execution in 8.9% to 37.8% of valid attacked post-exposu…
Researchers introduced Queen, a 4B-parameter chess-language model that plays at Grandmaster level and explains its moves, gaining over 900 Elo points (1782 to 2697) across seven iterations of training…
Researchers Shashank Kirtania and co-authors introduced BREW (Bootstrapping expeRientially-learned Environmental knoWledge), a framework that distills an LLM agent's past interaction trajectories into…
A stand-alone technical guide argues that external regulatory audits and internal validation harnesses catch distinct AI failure modes, and that only a combined "dual-audit" strategy guarantees robust…
A September 17, 2026 arXiv paper proposes the Output-Space Hypothesis, an enumerative equivalence-checking approach implemented in a system called Dirigo that flips the standard quantifier order to ch…
A pair of arXiv pre-prints, Spatial Memory Intelligence (SMI) and Latent Spatial Memory (LSM), introduce an understanding-driven long-term memory that couples semantic meaning with 3-D location, repla…
Researchers reported in an arXiv preprint (2609.16247v2) that large language models contain a "pain axis" — an internal activation pattern they call a "pain vector" that can be mapped in the model's a…
ArXiv announced a new rate-limit policy on October 1, 2026, capping researchers at two submissions per calendar month and three active submissions at any time, after the preprint repository received 4…
ArXiv capped author submissions at two per calendar month this week after the preprint server recorded 40,363 submissions in September 2026, double September 2024's 20,569 and quadruple September 2016…
A 2 October 2026 arXiv paper introduces STEER-Bench, a 101-task benchmark across 9 domains, and reports that branch steering attacks succeed against 94.4% of standard Computer Use Agents and 89.5% of …
A hybrid digital-analog deepfake video detection framework combining a lightweight digital front-end with a spatially multiplexed optical decoding back-end achieved 97.79% average detection accuracy, …
A new method called CITA improves tool-use agents by estimating the likelihood that a tool invocation will lead to task success, consistently raising Tool F1 and task success across three benchmarks, …
A new arXiv paper (2610.02478v1) proposes Tropical Reinforcement Learning, which replaces the standard sum of trajectory probabilities with a maximum over the tropical semiring, and reports that its T…
HakemBench, a Turkish benchmark of typed decisions released under CC BY 4.0 with 2,346 items and 4,275 choice, yes/no and score questions across seven tracks, scores a leader at a composite of 0.888 w…
A new arXiv paper (2610.02254v1) reports that an integrated LLM-ISM approach can replace repeated subject-matter-expert interviews in Interpretive Structural Modeling, with rowwise causal graph discov…
A new arXiv paper (arXiv:2610.02255v1) introduces MACTS-EM, a multi-agent collaborative time series forecasting framework that combines domain-specialised forecasting agents, a meta-cognitive allocati…
Researchers introduced MintFlow, a training-free constrained sampling framework that enforces constraints on flow matching models by applying a minimal perturbation to an intermediate flow state while…
Researchers proposed Comparative Inference for Tool-use Agents (CITA), a method that trains a Comparative Inference Model (CIM) to estimate the long-horizon value of a possible next tool invocation be…