Skip to content

Category

data science

2,518 papers

#machine learning Preprint Open access Oct 2026

The Geometry of Randomized Smoothing on Feasible Sets

Randomized smoothing certifies the probability of a fixed output event as the center of Gaussian noise moves. Feasibility or confidence filtering reports label probabilities only among retained proposals, producing a ratio. Its numerator is a fixed Gaussian event mass, while its denominator is the probability of retent...

Syed Izhan Khilji, Alireza Furutanpey, Schahram Dustdar · 0 citations
#machine learning Preprint Open access Oct 2026

Raw-Routed Mixture of Adapters: A Causal Intervention for Routing Collapse in Time Series Foundation Models

Time series foundation models (TSFMs) commonly adapt to new data by attaching a single trainable head to a frozen backbone, a one-size-fits-all setup that underfits heterogeneous regimes. Replacing the head with a mixture of experts is the standard upgrade, but on instance-normalized backbones (the dominant TSFM design...

Hung Phan, Thuy T. Nguyen, Minh Ngoc Dinh et al. · 0 citations
#machine learning Preprint Sep 2026

Awakening of the Buddha: Subspace Learning During Population-Loss Plateaus

Population loss can remain nearly constant while a neural network learns a substantially more predictive representation. We establish this separation for two-layer ReLU and leaky-ReLU networks trained on Gaussian inputs by simultaneous fixed-step population gradient descent on all parameters. For structured additive te...

Akash Kumar · 0 citations
#machine learning Preprint Sep 2026

A Rank Graduation metric for Algorithmic fairness

Fairness assessment in algorithmic decisions that affect individuals, such as credit scoring, often relies on parity measures calculated at the aggregate group level. Such measures may not reveal which individuals experience unfairness or which explanatory factors contribute to it. In this paper, we propose a rank-base...

Dalia Atif, Paolo Giudici · 0 citations
#machine learning Preprint Sep 2026

Flow Matching under Noisy Latent Structure: Beyond Exact Low-Dimensional Support

Flow Matching (FM) learns a velocity field whose ODE transports a simple source distribution to a target law. Existing finite-sample theory largely treats ambient-space regularity or data supported exactly on low-dimensional sets. We study linear FM under a noisy latent-generator model, where a low-dimensional H\"older...

Li-Feng Hao, Shao-Lin Ji · 0 citations
#machine learning Preprint Open access Oct 2026

Amortized Data Borrowing with Exchangeability-Aware Neural Posterior Estimation

Augmenting small concurrent studies with external or historical cohorts is attractive in drug development, where enrollment is slow, follow-up is expensive, and closely related trial or real-world data are often already available. Bayesian dynamic borrowing (BDB) provides a principled framework for adaptively controlli...

Chin-Hung Huang, JooChul Lee, Huan He · 0 citations
#machine learning Preprint Sep 2026

Learning-Enabled Estimation: Tight Characterizations under Sample Selection Biases

When can we learn from biased samples? We study regression when outcomes are observed only after passing through selection filters that depend on both covariates and outcomes themselves, a ubiquitous challenge spanning clinical trials with patient dropout, labor markets with self-selection, and auctions with strategic...

Vikram Kher, Jane H. Lee, Anay Mehrotra et al. · 0 citations
#machine learning Preprint Sep 2026

Grokking through the Lens of Minimum-Norm Interpolation

Grokking shows that fitting the training data and learning the underlying signal can occur at very different stages. However, existing theories offer limited quantitative insight into how this delayed generalization depends on inductive bias and signal structure. Our work addresses the gap by developing a statistical t...

Gil Kur, Ileana Rugina, C. Dominé et al. · 0 citations
#machine learning Preprint Sep 2026

Learning to Plan from Random Exploration

Random exploration reveals how an environment can be traversed before a goal is specified. Can this experience support long-range planning without policy-improvement training? Our random-walk analysis explains what temporal relations contain: short horizons reveal geodesic geometry in the diffusion limit, while longer...

De-Qian Kong, Guang-Yan Sun, Sheng Cheng et al. · 0 citations
#artificial intelligence Preprint Aug 2026

ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Learning

Time series classification underpins applications in healthcare, sensing, and industrial monitoring. Although time series foundation models support forecasting and transferable representation learning, classification still typically requires fitting a task-specific classifier on each target dataset, while individual ch...

Jun Fang, Shi-Feng Xie, Rui-Chu Cai et al. · 0 citations
#artificial intelligence Preprint Open access Oct 2026

From Dataset Spectral Geometry to Network Weights: A Geometry-Aware Initialization for Sigmoidal MLPs in Image Classification

Classical universal approximation theorems (UAT) establish the expressive power of sigmoidal multilayer perceptrons, but they do not specify how the weights should be initialized. We study a supervised, data-dependent, geometry-aware initialization for one-hidden-layer sigmoidal MLPs that compiles labeled class geometr...

Yi-Shan Chu · 0 citations
#artificial intelligence Preprint Open access Oct 2026

Flow Annealing Posterior Sampling for Function-Space Regression and Inverse Problems

Principled regression for stochastic processes is a long-standing challenge with deep connections to scientific inverse problems. We introduce Flow Annealing Posterior Sampling (FLAPS), to our knowledge the first function-space posterior sampling framework that unifies stochastic-process regression and PDE inverse prob...

Yaozhong Shi, Zachary E. Ross, Yisong Yue · 0 citations

From tech blogs

See all →
Microsoft Research Blog Oct 6, 2026

What AI gets wrong and what failure teaches us

Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity.  The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.