Skip to content

Category

data science

2,287 papers

#machine learning Preprint Oct 2026

Revisiting Explainable AI through Model-Independent Concept Dictionaries

Modern applications of AI rely on increasingly complex models. Explainable AI (XAI) has emerged as a set of techniques aimed at improving model transparency. However, existing XAI methods typically assume input features to be inherently interpretable, or they rely on intermediate internal abstractions that are difficul...

T. Schnake, Doreen Schöppenthau, Alexander Meyer et al. · 0 citations
#machine learning Preprint Open access Oct 2026

Shared Gaussianization: What Gaussian Regularizers Certify About Contrastive Learning, and What They Miss

What can a distribution-matching regularizer such as SIGReg in LeJEPA certify about contrastive learning? We study shared Gaussianization (SG), a characteristic-function Gaussianity test on the average of two normalized views, scaled by an independent $\chi_d$ radius. Because disagreeing views shorten the average, one...

Ruoyu Zhao, Yuting Chen, Jinheng Zhang et al. · 0 citations
#machine learning Preprint Open access Oct 2026

Finite-Sample Approximation of Hessian-Guided Perturbed Wasserstein Gradient Flows

Wasserstein gradient flow extends gradient descent to probability measures. Its Hessian-guided perturbed variant (PWGF) adds Gaussian perturbations to escape saddle points in nonconvex problems. We investigate when its approximation by finitely many interacting particles remains accurate over growing time horizons. Our...

Ryotaro Kawata, Atsushi Nitanda, Taiji Suzuki · 0 citations
#machine learning Preprint Open access Oct 2026

Pre-training of Bayesian Optimization Algorithm through Bayesian Optimization

Bayesian optimization (BO) is widely used as a standard approach for expensive black-box optimization. However, BO algorithms often involve parameters that must be specified in advance, and their performance can strongly depend on these choices. We propose a framework for optimizing such parameters using sample paths d...

Satoshi Katayama, Shoyo Hunt, Shintaro Masuda et al. · 0 citations
#machine learning Preprint Open access Oct 2026

m-Set Adversarial Bandits with Winner Feedback

We show upper and lower bounds on the regret of $m$-set adversarial bandits for different utilities (winner reward or sum of rewards) and feedback models (winner index, winner reward, sum of rewards, and their combinations). By comparing to standard bounds for combinatorial and MNL bandits, our results reveal how subtl...

Nicol\`o Cesa-Bianchi, Matteo Papini · 0 citations
#machine learning Preprint Open access Oct 2026

Expected Sample Complexity in Multi-Armed Bandits

Sample complexity is a widely used metric in sequential decision-making problems, defined as the number of suboptimal decisions during the interaction between the agent and an environment. We study the sample complexity of stochastic multi-armed bandit problems and introduce the expected sample complexity performance m...

Nadav Sukenik, Nadav Merlis · 0 citations
#machine learning Preprint Open access Oct 2026

Eigenvalues of the Hessian in Deep Learning: The Origin of Symmetry and Its Breaking

Hessian spectra at trained models in deep learning exhibit a persistent pattern: eigenvalues organize into distinct clusters, including a large bulk near zero and a few isolated outliers. This paper shows that a natural account of these spectral phenomena emerges when the original setting is understood as a departure f...

Yossi Arjevani · 0 citations
#machine learning Preprint Open access Oct 2026

Identifiability of a dissipative knowledge-dynamics model: exact recovery under designed excitation, degeneration on observational data

Human learning is a dissipative dynamical process: mastery accumulates through practice, decays through forgetting, and propagates across interdependent concepts. We model it as a nonlinear dissipative system of ordinary differential equations whose parameters are mechanistically meaningful (a concept-transfer matrix e...

Arman Kostanian, Armen Beklaryan · 0 citations
#machine learning Preprint Open access Oct 2026

AdaPS-LiNGAM: Adaptive Predecessor Selection for Linear Non-Gaussian Acyclic Models under Small-Sample Settings

Causal discovery becomes particularly challenging when the available sample size is small relative to the number of variables. This challenge also arises in the linear non-Gaussian acyclic model (LiNGAM), an identifiable framework for causal discovery from observational data. DirectLiNGAM estimates a causal order, whic...

Shun Yanashima, Kentaro Kanamori, Hirofumi Suzuki · 0 citations
#machine learning Preprint Open access Oct 2026

Fluctuations of Nonlinear Observables in Mean Field Neural Network Training

Mean field limits describe the training dynamics of wide neural networks through the evolution of the empirical distribution of their parameters. Although functional central limit theorems characterize the asymptotic fluctuations of this distribution, quantities of practical interest are typically nonlinear observables...

Arnaud Descours (UCBL), Geoffrey Lacour (MaIAGE) · 0 citations
#machine learning Preprint Open access Oct 2026

Leaner Transformers Can Easily Learn to Cluster

Transformers have in-context learning capabilities, where some known learning algorithms can be executed in the forward pass through the model. Recent work shows that transformers can exactly perform Lloyd's algorithm for $k$-means clustering with $n$ points in $d$ dimensions with an embedding size $d_{\textsf{emb}} =...

Charlotte Park, Kenneth L. Clarkson, Lior Horesh et al. · 0 citations
#machine learning Preprint Open access Oct 2026

EntroPrefill: Renyi-Guided Context Pruning with Conditional Stability Guarantees for Retrieval-Augmented Generation

Mid-prefill pruning can reduce the sequence processed by deeper transformer layers, but attention concentration alone does not certify that discarded context is dispensable. We formulate EntroPrefill as a Renyi-guided proposal mechanism coupled to explicit constraints on discarded attention mass. Sink-isolated, regular...

Inbasekaran S · 0 citations

From tech blogs

See all →
Microsoft Research Blog Oct 6, 2026

What AI gets wrong and what failure teaches us

Jennifer Neville did not want to go into computer science—but that’s exactly where she landed. Neville discusses the starts and stops that led to her professional sweet spot and her work identifying “surprising failures” making it hard for AI to handle complexity.  The post What AI gets wrong and what failure teaches us appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.