Skip to content
Preprint

Automatic Model-Order Selection for Nonnegative Matrix Factorization via Column $\ell_{2,0}$ Regularization

Jul 2026 · 0 citations · 35 references
Mathematics

TL;DR

The critical points and local minimizers of the model are characterized, a sufficient-condition result for rank recovery is provided, and whole-sequence convergence of the proposed algorithms under explicit step-size and inertial-parameter conditions are proved using the Kurdyka--\L{}ojasiewicz framework.

Abstract

Nonnegative matrix factorization represents nonnegative signals as additive combinations of latent components, but its factorization rank, and hence the model order, must usually be specified beforehand. An underestimated order discards signal structure, whereas an overestimated order produces redundant components and unstable decompositions. We propose a column $\ell_{2,0}$-regularized formulation that estimates the model order from an initial upper bound by suppressing inactive columns in both factors. A warm-started regularization path progressively removes redundant components without changing the factor dimensions, and a marginal reconstruction-loss criterion selects an order along the path. To solve the resulting nonconvex and discontinuous problem, we develop an inertial proximal alternating linearized minimization method, a scale-balanced variant, and a proximal active-set method based on P-stationarity. The balancing operation equalizes the norms of paired factor columns while preserving their rank-one products. We characterize the critical points and local minimizers of the model, provide a sufficient-condition result for rank recovery, and prove whole-sequence convergence of the proposed algorithms under explicit step-size and inertial-parameter conditions using the Kurdyka--\L{}ojasiewicz framework. Dedicated experiments show that the warm-started $\lambda$-path is more efficient than increasing- and decreasing-order discrete $r$-paths, while scale balancing yields a more stable rank-selection path. Experiments on synthetic data and diverse signal benchmarks show that both iPALM and PASM provide reliable model-order estimates with favorable computational efficiency.

View source

Similar papers

Preprint Aug 2026

Exact Rank-Space KL Projection for Shared-Marginal Low-Rank Factors: Application to Doubly Stochastic Clustering

We study exact Kullback--Leibler (KL) projection for low-rank factorizations whose two nonnegative factors have prescribed row marginals and a shared, learned column marginal. For arbitrary positive row marginals of equal total mass, the joint KL projection reduces exactly to a strictly convex gauge-fixed dual with onl...

En-Liang Hu · 0 citations
#machine learning Preprint Aug 2026

Separable Nonnegative Matrix Factorization Using Powered Ratio-of-Norms Regularization

This work develops efficient algorithms based on the difference-of-convex function algorithm (DCA) and the alternating direction method of multipliers (ADMM) to enhance sparsity and identifiability of the learned factors in separable nonnegative matrix factorization.

Matthew McCarver, Jing Qin · 0 citations
Preprint Sep 2026

Robust low-rank tensor completion via factorized weighted tensor schatten-p norm minimization

Low-rank tensor factorization provides a flexible framework for completing multidimensional data from incomplete and corrupted observations. However, unweighted spectral regularizers impose a common shrinkage profile across singular components, which may excessively attenuate dominant low-rank components, and factorize...

Bing-Hao Wang, Feng Zhang, Wen-Dong Wang et al. · 0 citations
Preprint Aug 2026

Exact Rank and Convex Calibration Dimension Lower Bounds for the Multi-Label F1 Loss

The instance-wise $F_1$ measure is a central performance measure for multi-label classification. For a problem with $s$ labels, it defines a $2^s\times 2^s$ loss matrix. Previous work exhibited $s^2+1$-coordinate affine and shifted low-rank representations and used them to construct quadratic-dimensional convex calibra...

Mingyuan Zhang · 3 citations
#machine learning Preprint Sep 2026

Automatic Rank Allocation for Low-Rank Adaptation in Large Language Models via lp Regularization

Low-rank adaptation (LoRA) has become a popular parameter-efficient fine-tuning method for large language models. A key challenge in LoRA is how to determine the rank of each adaptation matrix, as rank directly controls its capacity and efficiency. Existing adaptive-rank methods typically allocate ranks according to ma...

Ze-Bang Xie, Chuan-Yang Zheng, Yik-Chung Wu et al. · 0 citations
Preprint Aug 2026

Subzero matrix completion for sparse data analysis: large-scale learning of latent low-rank structure

A stochastic, alternating least-squares algorithm that operates on smaller blocks of this dense matrix and scales as a result to much larger problems and is used to analyze the sparse matrix of synaptic weights for the recently published $\textit{Drosphilia}$ connectome.

Lawrence K. Saul, N. Huang, Dennis Bollweg et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.