A Bellman Optimality Equation for Plasticity
A new arXiv paper (arXiv:2609.10776v1) presents preliminary work showing that a Bellman optimality equation exists for optimizing plasticity within Markov decision processes, building on Abel et al. (…
A new arXiv paper (arXiv:2609.10776v1) presents preliminary work showing that a Bellman optimality equation exists for optimizing plasticity within Markov decision processes, building on Abel et al. (…
A systematic literature mapping published on arXiv reviewed 96 articles from 2015 onward on AI and deep learning for lung cancer detection in radiology, finding convolutional neural networks with tran…
Researchers proposed the Adaptive Margin Ordinal Loss (AMOL), a multiplicative per-class weighting scheme that suppresses "center-class hedging" in ordinal classification, according to an arXiv paper …
Researchers proposed counterfactual (CF) marginalisation, a test-time evaluation procedure for assessing how robust classification models are to nuisance variables such as age or sex, according to a p…
Researchers introduced Graph-Guided Quasimetric Dense Reward (G2QDR), a framework that predicts pairwise state connectivity strength in asymmetric environments and converts those strengths into scalar…
A new arXiv paper (arXiv:2609.10866v1) extends certification methods for reinforcement learning to risk-sensitive objectives, establishing lower bounds on the exponential utility of cumulative rewards…
A new arXiv paper (2609.10863v1) identifies a duality between continuous and discrete flow matching, showing that projecting continuous convex-interpolant paths with one-hot targets through a position…
Researchers introduced MHE-Former, a Transformer-based multi-hypothesis framework that uses entropy maximization to generate diverse 3D hand and body mesh recovery predictions from monocular input, ac…
A few-shot regression framework combining Vision Transformer (ViT) feature embeddings, fuzzy c-means clustering-based task construction, and gradient-based meta-learning enables reliable plant growth …
Researchers introduced CamPilot, a multi-agent framework that integrates cinematographic planning and camera-work control to generate more coherent, logically structured movies, trained via a GRPO-bas…
A new arXiv paper (2609.11022v1) introduces a controlled benchmark testing whether vision language models can decide when to answer a physical reasoning question immediately versus when to select the …
A new arXiv paper (2609.11227v1) shows that time-to-first-spike spiking neural networks can generate richer partitions of the input space than conventional feedforward ReLU networks. The authors deriv…
A Kubernetes Dynamic Resource Allocation (DRA) driver makes composable CXL memory a schedulable cluster resource for cross-node KV-cache reuse in LLM serving, according to an arXiv paper (2609.10790v1…
A production skill router covering 34,396 skills shows that fine-tuning on synthetic data improves in-distribution skill retrieval for LLM agents but causes catastrophic forgetting on real and out-of-…
A new arXiv paper (2609.10613v1) proposes a posterior reweighting framework that models safety-aligned multimodal large language models (MLLMs) as implicitly operating over competing behavioral modes,…
Researchers introduced RDDMPI, a conditional residual diffusion framework for probabilistic multivariate time series imputation, detailed in arXiv paper 2609.11648v1. RDDMPI reformulates imputation as…
A new arXiv paper (2609.11780v1) reports that spectral metrics from the WeightWatcher framework can predict membership inference attack (MIA) vulnerability in machine learning models without training …
A new arXiv paper (2609.11538v1) shows that Generative Marginalization Models (MaMs) are equivalent to Generative Flow Networks (GFlowNets), and introduces Particle GFlowNets, which extends MaMs' samp…
ObstaDiff, a decomposed diffusion-policy framework with a lightweight obstacle-aware visual encoder, achieved 75.41% average task success and an 8.20% average obstacle collision rate across 61 real-ro…
A large-scale empirical study of 265,363 Python notebooks submitted to Kaggle competitions found that general Python code quality is decoupled from machine learning performance, showing negligible or …