{"slug": "nomadd-numerical-optimization-of-models-adapting-to-data-drift", "title": "NOMADD: Numerical Optimization of Models Adapting to Data Drift", "summary": "Researchers propose NOMADD, a post-hoc method to reduce concept drift in tabular models, applicable to trees, neural networks, and tabular foundation models. On the 18-dataset Drift-Resilient TabPFN benchmark, NOMADD improves every base family and achieves performance competitive with state-of-the-art Drift-Resilient TabPFN in seconds of training, whereas Drift-Resilient TabPFN requires pre-training on millions of synthetic datasets over approximately 1,300 GPU-hours.", "body_md": "arXiv:2608.02845v1 Announce Type: new\nAbstract: Tabular model performance degrades when feature distributions change over time or the relationship between features and outcome variables change over time, known as data drift and concept drift, respectively. These issues are challenging to mitigate in real time because labeled data may not be immediately available, or re-training a model could be impractical. While tools exist to reduce drift, they are typically bespoke to neural network architectures and adapt how models are trained. In this paper, we offer an alternative post-hoc method to reduce concept drift, which is applicable to a variety of models, from trees to neural networks to tabular foundation models. This new tool is especially useful when constraints, such as high model accuracy, bounded inference time, or model size requires users to choose between different models for their specific use-cases. Our algorithm fits the base model separately on each labeled training period, measures how its parameters evolve against a single anchor model pooled over all of those periods, compresses those changes with a low-rank factorization, and extrapolates each latent factor forward with a damped, regularized forecast. On the 18-dataset Drift-Resilient TabPFN benchmark, evaluated under that benchmark's own protocol and metric, the extrapolation improves every base family it is applied to, and achieves performance competitive with the state-of-the-art Drift-Resilient TabPFN with seconds of training. In contrast, Drift-Resilient TabPFN requires pre-training on millions of synthetic datasets over approximately 1,300 GPU-hours, and is orders of magnitude slower in inference (depending on the model). In the discussion, we explore the promise and challenges of extending this tool to other modalities.", "url": "https://wpnews.pro/news/nomadd-numerical-optimization-of-models-adapting-to-data-drift", "canonical_source": "https://arxiv.org/abs/2608.02845", "published_at": "2026-08-05 04:00:00+00:00", "updated_at": "2026-08-05 04:02:46.653799+00:00", "lang": "en", "topics": ["machine-learning", "artificial-intelligence"], "entities": ["NOMADD", "Drift-Resilient TabPFN"], "alternates": {"html": "https://wpnews.pro/news/nomadd-numerical-optimization-of-models-adapting-to-data-drift", "markdown": "https://wpnews.pro/news/nomadd-numerical-optimization-of-models-adapting-to-data-drift.md", "text": "https://wpnews.pro/news/nomadd-numerical-optimization-of-models-adapting-to-data-drift.txt", "jsonld": "https://wpnews.pro/news/nomadd-numerical-optimization-of-models-adapting-to-data-drift.jsonld"}}