Skip to content
Preprint

Beyond Modern Asymptotics for Log-Likelihood Ratios in Logistic Regression

Aug 2026 · 0 citations
Mathematics Computer Science

Abstract

We characterize the finite sample behavior of the log-likelihood ratio statistic in binary logistic regression, uniformly over both the design and the target parameter. For $n\geq d\geq 3$, we determine, up to universal constants, its worst case $(1-\delta)$ quantile over all fixed collections of design vectors and all target parameters: \[ d\log\left(\frac{e n}{d}\right)+\log\left(\frac{1}{\delta}\right). \] This is a nonasymptotic analogue of the Wilks $\chi^2_d$ phenomenon and requires no regularity assumptions on the design. The low dimensional cases exhibit unusual behavior. The worst case quantile in dimension $d=2$ is sharply of order \[ \log\log\log n+\log\left(\frac{1}{\delta}\right). \] The worst case quantile in dimension $d=1$ is of order $\log(1/\delta)$, with no dependence on $n$. Finally, i.i.d. Gaussian design vectors recover the classical Wilks scale. In the regime $n\gtrsim d+\log(1/\delta)$, we prove the sharp bound \[ d+\log\left(\frac{1}{\delta}\right). \] Unlike existing asymptotic results, our bounds are uniform over the target parameter, which may depend on $n$, $d$, and $\delta$.

View source

Similar papers

Aug 2026

Limit laws of the iterated logarithm under sub-linear expectations

Let $\{Y_n; n\ge 1\}$ be a sequence of independent and identically distributed random variables with mean zero in Peng's framework of the sub-linear expectation space $(\Omega,\mathscr{H},\widehat{\mathbb E})$, and $S_n=\sum_{i=1}^nY_i$. In this paper, we establish a limit law of \begin{align*}\lim_{n\to \infty}\max_{k\le n}\frac{S_k}{\sqrt{2k \log\log n}}. \end{align*} Different from the result obtained by Chen (2015) in which the limit is a constant, it is shown that under the upper capacity the limit may be prescribed as a given function of $Y_1,Y_2,\ldots$, taking values in the standard deviation interval. As a result, it is also shown that the set of limit points in the compact law of the iterated logarithm can be a symmetric random interval. This paper (Chinese version) has been submitted to Special Issue of Science in China-Mathematics in Celebration of Professor Peng Shige's 80th Birthday. In Theorem 2.2 of the original paper, an additional condition (2.6) is needed.

Li-Xin Zhang, Yongze Song · 0 citations
Preprint Aug 2026

Uniform sine-kernel determinant asymptotics, tail-side quantiles, and prolate eigenvalue bounds

Let $S_c=P_{(0,c)}QP_{(0,c)}$ be the one-dimensional sinc-kernel concentration operator, let $N_a(c)=\#{n:\lambda_n(c)>a}$, and set $\bar L=\log((1-\delta)/\delta)$. We prove, uniformly for each fixed $A>0$, the tail-side quantile formula $N_\delta(c)=c+\pi^{-2}\bar L\log(4\pi^2c/\bar L)+O_A(\log c+\bar L)$ for $6\le\bar L\le A\log c$. It yields corresponding additive formulas for the lower half and full plunge, with main terms respectively $\pi^{-2}\bar L\log(4\pi^2c/\bar L)$ and twice this quantity. An exact one-tail-coordinate selection gives, for fixed $A>0$, $d\ge1$, and $q\in(1/2,1)$, the one-sided tensor-product bound $\Lambda_\delta(c;d)\ge\pi^{-2}d c^{d-1}\bar L\log(4\pi^2c/\bar L)-O_{A,d,q}(c^{d-1}(\log c+\bar L))$ for $L_{d,q}\le\bar L\le A\log c$, where $L_{d,q}=\log(q^{-(d-1)}(e^6+1)-1)$; the tensor content is nontrivial for $d\ge2$. The analytic input is a signed growing-parameter sine-kernel determinant asymptotic: uniformly for $0\le\omega\le A\log s$, $\log\det(I+(e^{2\omega}-1)K_s)=4\omega s/\pi+2\pi^{-2}\omega^2\log(4s)+2\log|G(1+i\omega/\pi)|^2+O_A((1+\omega)^4\log^2s/s)$, where $G$ is the Barnes $G$-function. We prove this negative-coupling counterpart of the Bothner--Deift--Its--Krasovsky theorem by direct IIKS steepest descent. We also retain the uniform head-side results and use a two-way determinant reduction to obtain the moving-depth lower-half bridge bound with constant $1/(32\pi^2)$; extending it to the deeper range uses Kulikov--Dam Larsen and may require a smaller constant. These counting formulas are additive. Their errors become uniformly relative when $\bar L$ tends uniformly to infinity; fixed thresholds are covered separately by Landau--Widom. A Lambert-$W_{-1}$ formula is recorded only for the continuous main term, not for individual eigenvalues.

A. Azimifard · 1 citation
Preprint Aug 2026

Empirical likelihood confidence regions for ordered bivariate means

Let $\boldsymbol{X}_i=(X_{1i},X_{2i})^\top$ be independent and identically distributed observations with mean $\boldsymbol{\mu}=(\mu_1,\mu_2)^\top$ constrained by $\mu_1\leq\mu_2$. We study empirical-likelihood inference for a fixed mean vector and distinguish it from the previously known test of equality against an ordered alternative. At a fixed interior point, the constrained empirical likelihood ratio has the usual $\chi^2_2$ limit. At a fixed boundary point $(m,m)^\top$, its limit is the chi-bar-square distribution $\tfrac12\chi^2_1+\tfrac12\chi^2_2$. By contrast, profiling the unknown common mean in the equality-versus-order test yields $\tfrac12\chi^2_0+\tfrac12\chi^2_1$, the $k=2$ ordered-mean case of El Barmi (1996). We give an exact reduction of the latter statistic to the empirical likelihood of the paired differences, establish the localization step needed for the fixed-boundary expansion, and derive a local-to-boundary limit showing that interior calibration is not uniform over $n^{-1/2}$-neighborhoods of the boundary. Monte Carlo experiments under Gaussian, Student $t_5$, and shifted log-normal sampling examine fixed, boundary, and local regimes with explicit numerical-failure accounting. Illustrative paired-data analyses show the practical distinction between fixed-candidate confidence regions, directional equality tests, and ordinary scalar empirical-likelihood intervals truncated to the nonnegative parameter space.

N. Garg · 0 citations
Preprint Aug 2026

Cubic-Root Gaussian Approximation under Unrestricted Covariance

For Gaussian approximation over high-dimensional rectangles under unrestricted covariance, Chernozhukov et al. (2023b) conjectured that the $n^{-1/4}$ rate, up to logarithmic factors, is near-optimal. We show that, under the coordinatewise subexponential condition with scale $B_n$ and the marginal variance lower bound condition with constant $b$ in Chernozhukov et al. (2023b), the approximation error in dimension $d$ is bounded by \begin{align*} C_b\min\left\{ 1,\, \left(\frac{B_n^2}{n}\right)^{1/3}\{\log(2dn)\}^{7/3} + \frac{B_n}{\sqrt n}\{\log(2dn)\}^{5/2} \right\}. \end{align*} In particular, for bounded $B_n$ and polynomial dimension, the new bound is $n^{-1/3}$ and therefore falsifies the polynomial-dimensional $n^{-1/4}$ near-optimality conjecture. The proof uses a two-stage interpolation and a rank-free matrix-weighted Gaussian surface bound, which may be of independent interest. The initial proof attempt was generated by ChatGPT 5.6 Pro (OpenAI) and subsequently corrected and rewritten by the authors. The machine-checked Lean formalization of the proof can be found at the GitHub repository (https://github.com/WeihanZhang2001/cubic-root-gaussian-approximation-under-unrestricted-covariance).

Zijun Gao, Wei-Han Zhang · 0 citations
Preprint Aug 2026

Sharp Tail Bounds Beyond Twice the Mean

Consider $n$ independent, non-negative, mean at most one random variables, $X_1,X_2,\ldots$. We show the following bound on the probability of their sum exceeding a threshold $t$: \[ \mathbb{P}\left[\sum_{i=1}^n X_i\ge t\right] \leq 1-\left(1-\frac{1}{t}\right)^n \text{ for all } t\ge 2n+1 \,. \] To prove this, we consider a relaxed optimization problem over a set of sequences of ordered, but non-independent random variables. This allows us to reformulate it recursively as dynamic programming problem. The bound becomes an equality for the binary i.i.d.~random variables satisfying $\mathbb{P}\left[X_i=0\right]= 1-\frac{1}{t}$ and $\mathbb{P}\left[X_i=t\right]=\frac{1}{t}$, which remains the maximizer in the relaxed problem.

P. Strack, Jannik M. Westermann · 1 citation
Preprint Aug 2026

The Optimal Rate in the Averaged Random-Marginal Central Limit Theorem for Log-Concave Measures

Let $X$ be a centered isotropic log-concave random vector in $\mathbb{R}^n$. For $\theta\in S^{n-1}$, let $\mu_\theta$ be the law of $\langle X,\theta\rangle$, and let $\Theta$ be uniformly distributed on $S^{n-1}$, independently of $X$. We prove the sharp estimate \[ \textsf{E} W_1(\mu_\Theta,\gamma_1) \le \frac{C}{n}. \] Here the Wasserstein distance is computed after the direction is fixed and is then averaged over the sphere. No symmetry assumption is imposed. A product measure with centered exponential coordinates gives a matching lower bound of order $n^{-1}$. The proof separates the averaged-direction law from the fluctuation among fixed directions. For the first part, a Taylor expansion in the random radius retains a mean-zero cancellation and yields an $O(n^{-1})$ error. For the second, a weighted $L^2$ distance between distribution functions is converted into an exact spherical kernel depending only on $|x|^2$, $|y|^2$, and $\langle x,y\rangle$. Expanding this kernel in $\langle x,y\rangle$, we control its linear, quadratic, and cubic terms using the quadratic variance inequality $\operatorname{Var}(X^\top M X) \le 8\operatorname{Tr}(M^2)$, while fixed-order moment estimates control the remainder.

Xuanang Hu · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.