Statistics Theory

2026-07-10 | | Total: 14

#1 Low-Rank Matrix Recovery via Heavy-Tailed Quadratic Sampling [PDF] [Copy] [Kimi] [REL]

Authors: Gao Huang, Song Li

The problem of recovering an (approximately) low-rank Hermitian matrix $\pmb{M}_0 \in \mathbb{C}^{n \times n}$ of rank $r$ from quadratic sampling matrices of the form $\{\pmb{a}_k \pmb{a}_k^*\}_{k=1}^m$ arises in a variety of applications, including phase retrieval. To obtain rigorous recovery guarantees, the sampling vectors $\{\pmb{a}_k\}_{k=1}^m$ are typically modeled probabilistically. However, most existing theoretical results rely on Gaussian or sub-Gaussian assumptions, which may not accurately capture practical data models. In many applications, sampling vectors exhibit heavier tails, while theoretical understanding in such regimes remains scarce. In this paper, we bridge this gap. We show that two widely used convex approaches, nuclear norm minimization and semidefinite-constrained empirical risk minimization, achieve uniform, stable, and robust recovery under the mild assumption that the entries of the sampling vectors have only finite $4+δ$ moments, with the optimal sample complexity $m = \mathcal{O}(rn)$ up to moment-dependent constants. The two main ingredients of our analysis are moment estimates for quadratic forms established via decoupling, together with recent advances in covariance estimation in heavy-tailed settings. As byproducts, we also establish the optimal sample complexity for low-rank matrix recovery under complex projective $4$-design sampling, thereby improving upon previous results, and obtain stability guarantees for phase retrieval under similarly weak moment assumptions.

Subjects: Statistics Theory , Information Theory

Publish: 2026-07-09 16:36:06 UTC


#2 Exact Permutation Recovery Under Unknown Scalar Affine Transformation [PDF] [Copy] [Kimi] [REL]

Authors: Tigran Galstyan, Avetik Karagulyan, Arshak Minasyan

We study the problem of matching two sets of noisy feature vectors when underlying true features are related by an unknown scalar affine transformation. Our method comprises two primary steps. First, we standardize the feature vectors to estimate the unknown scalar affine transformation. Subsequently, we estimate the permutation by minimizing the Least Sum of Logarithms (LSL) between two sets of observations using the estimated transformation. Our main result shows that the unknown permutation can be perfectly recovered given that the minimal separation distance of true feature vectors scales as $\sqrt{ρ_σ} \vee (d\log n)^{1/4} \vee \sqrt{\log n}$, where $d$ is the ambient dimension, $n$ is the sample size, and $ρ_σ$ is the maximal ratio of noise magnitudes. Interestingly, the obtained rate, under mild heteroscedasticity, coincides with that of the non-affine setting. We additionally demonstrate that there exist configurations requiring a larger minimal separation distance for perfect recovery. The latter makes the matching problem more challenging from minimax perspective compared to the non-affine setting. Consequently, we show that in the problem of feature matching, standardizing the data implicitly estimates the scalar affine parameters. As part of our analysis, we prove non-asymptotic concentration bounds for the affine parameter estimators in the presence of heterogeneous noise magnitudes.

Subject: Statistics Theory

Publish: 2026-07-09 15:59:11 UTC


#3 Functional dependence and synchronous coupling in ergodic autoregressions [PDF] [Copy] [Kimi] [REL]

Authors: Paul Doukhan, Lionel Truquet

Functional dependence measures have become an important tool in the analysis of nonlinear time series and are typically formulated with respect to a given innovation representation of the process. This note points out that the probability space on which such representations yield the expected memory loss properties may not always coincide with the natural dynamical probability space of the model. We exhibit classes of uniformly ergodic autoregressive processes for which the behavior of the natural innovation coupling undergoes a qualitative transition as the model parameter varies. For this family of models, this transition coincides with a change in the sign of an associated Lyapunov exponent. In particular, a positive Lyapunov exponent may prevent the forgetting of initial perturbations along trajectories driven by the same innovations, despite uniform ergodicity of the associated Markov chain. These observations highlight the importance of carefully specifying the underlying probability space when interpreting or applying functional dependence measures.

Subject: Statistics Theory

Publish: 2026-07-09 14:56:31 UTC


#4 A screening approach to nonparametric inference from the M/G/1 workload [PDF] [Copy] [Kimi] [REL]

Authors: Royi Jacobovic, Binyamin Kobzantsev

We address a long-standing open problem posed by Hansen and Pitts (2006) on nonparametric inference for the service-time distribution in an M/G/1 workload model. We consider an M/G/1 queue with unknown arrival rate $λ>0$ and service-time distribution $B(\cdot)$, without assuming stability or stationarity. A statistician observes the workload process at discrete times $t=0,1,\ldots,n$ and aims to estimate $B(w)$ at a fixed point $w>0$. We propose an estimator $B_n(w)$ based solely on the observed workload trajectory. The construction relies on a screening mechanism that extracts conditionally i.i.d. compound Poisson increments from the workload process, thereby reducing the dependent-data problem to a Laplace-transform inversion framework. Under mild regularity assumptions on $B(\cdot)$, i.e., continuous differentiability on $[0,\infty)$, twice differentiability at $w$, and a finite second moment, we establish the bound \[ \mathbb{E}\bigl|B_n(w)-B(w)\bigr| =\mathcal{O}\!\left(\frac{\log n}{\sqrt{n}}\right), \qquad n\to\infty. \]This provides the first solution to the Hansen-Pitts problem achieving a parametric $L^1$-risk rate (up to a logarithmic factor), without requiring stationarity, stability, or knowledge of the arrival rate.

Subjects: Statistics Theory , Probability

Publish: 2026-07-09 13:30:03 UTC


#5 Building confidence regions for Reeb graphs using the interleaving distance [PDF] [Copy] [Kimi] [REL]

Authors: Matteo Pegoraro, Alberto Conforti, Mathieu Carrière

We develop confidence regions for Reeb graphs from finite samples using the interleaving distance. Given a point cloud equipped with a filter function, we construct a finite proximity graph, extend the filter linearly, and use the Reeb cosheaf of the resulting filtered graph as the primary estimator. Mapper graphs are then treated as controlled cover-based coarsenings of this estimator, separating the statistical approximation problem from the visualization problem. We prove stability bounds for the Reeb estimators obtained both using intrinsic and extrinsic metrics, the latter under positive-reach assumptions, and derive interleaving-distance confidence regions from either \((a,b)\)-standard sampling assumptions or subsampling-based Hausdorff scale estimates. We also compare this object-level metric viewpoint with persistence-based guarantees by showing that the extended-persistence pseudometric is bounded by twice the interleaving distance, with sharp constant \(1\) for the \(H_0\)-related components. Numerical experiments illustrate how statistically significant features can be identified and then projected to Mapper graphs for interpretation.

Subjects: Statistics Theory , Algebraic Topology

Publish: 2026-07-09 13:18:35 UTC


#6 An Exact Distribution-Free Test for Means of Nonnegative Random Variables [PDF] [Copy] [Kimi] [REL]

Authors: Nikos Vlassis, Philip S. Thomas

Let $X=(X_1,\ldots,X_n)$ be independent nonnegative random variables, not necessarily identically distributed. Let $D=(D_0,D_1,\ldots,D_n)\sim\operatorname{Dir}(1,\ldots,1)$ be independent of $X$, and define $K(x)=\mathbb{P}\{\sum_{i=1}^n x_iD_i\le1\}$. We prove that, for every $n\ge1$, whenever $\mathbb{E} X_i\le1$ for every $i$, $\mathbb{P}\{K(X)\leα\}\leα$ for all $0\leα\le1$. Thus $K(X)$ is a finite-sample, distribution-free $p$-value for testing the null hypothesis $\mathbb{E}X_i \le 1$ for all $i$. This proves a conjecture of Gaffke (2005).

Subjects: Statistics Theory , Probability

Publish: 2026-07-09 12:39:29 UTC


#7 Parameter inference for partially observed branching processes [PDF] [Copy] [Kimi] [REL]

Authors: Simone Baldassarri, Michel Mandjes, Jiesen Wang

In this paper, we study an age-dependent branching process. In the simplest setting, the population is divided into two age groups, namely juveniles and adults. Our objective is to estimate the model parameters using observations of the total population size only (i.e., juveniles plus adults). Focusing on the ergodic regime of the model, we introduce a method-of-moments estimator and establish its asymptotic normality. Several extensions are discussed, including models with more than two age groups.

Subjects: Statistics Theory , Probability

Publish: 2026-07-09 06:51:12 UTC


#8 From Bayes' Rule to Bayes Rules: Optimal Information Processing and Axiomatic Foundations Beyond Probability [PDF] [Copy] [Kimi] [REL]

Authors: Jeremie Houssineau, Badr-Eddine Chérief-Abdellatif

This paper develops principled updating rules for possibilistic inference, where uncertainty about a fixed parameter is represented by a possibility function, the maxitive analogue of a probability distribution, and comparisons are made pointwise via a partial order. From two complementary foundations, an information-conservation viewpoint and an axiomatic viewpoint, we derive the same canonical update: the posterior is the prior-likelihood product followed by supremum normalisation. The two derivations agree for an arbitrary loss, differing only in where the learning-rate parameter enters. This parameter controls epistemic strength and is not identifiable from the normalising evidence alone, clarifying the role of analogous learning-rate parameters in generalised Bayesian updating.

Subjects: Statistics Theory , Information Theory

Publish: 2026-07-09 00:53:41 UTC


#9 The logistic-normal integral and the moments of the logistic-normal distribution [PDF] [Copy] [Kimi] [REL]

Author: Dan Pirjol

The logistic-normal integral appears in problems of statistical estimation for logistic models with Gaussian random effects, and generalized linear mixed models. We study the numerical evaluation of this integral and of its derivatives, and give closed form evaluations at certain points and series expansions. There is a continuum of possible series expansions, and we single out one series expansion which is optimal for numerical evaluation. We propose an algorithm for a precise numerical evaluation, based on the optimal series, with good approximation error control in the tails. As an application we give explicit results for the first two moments of a logistic-normal random variable.

Subjects: Statistics Theory , Classical Analysis and ODEs

Publish: 2026-07-08 19:53:03 UTC


#10 A Design-Based Approach to Testing and Inference in (Quasi-)Experiments with Spillovers [PDF] [Copy] [Kimi] [REL]

Author: Yechan Park

Economic policies rarely affect only their direct targets. To study these spillovers, researchers summarize who else was treated with a simple exposure measure, such as the share of treated neighbors within a radius. But for many settings, economic theory provides little guidance on choosing the functional form (e.g., ring) of that measure or its parameters (e.g., radius). We show that the data can inform both choices. Correctly specified exposure measures imply orthogonality conditions that can be used for both estimation and testing. We establish consistency and asymptotic normality of the resulting estimator under spatial and network dependence in a design-based framework, with all randomness arising from treatment assignment. We then characterize the efficient moment conditions. Applied to two large-scale anti-poverty programs, the framework supports some prior radius estimates but rejects others. In the latter case, the revised radius yields substantively different policy-effect estimates.

Subjects: Econometrics , Statistics Theory , Methodology

Publish: 2026-07-09 16:16:17 UTC


#11 High-Dimensional Procrustes Matching via Tree Counts [PDF] [Copy] [Kimi] [REL]

Authors: Xiaochun Niu, Tselil Schramm, Jiaming Xu

Suppose we observe two sets of $n$ Gaussian vectors in $\mathbb{R}^d$, with the promise that, after applying a permutation of $[n]$ and a rotation of $\mathbb{R}^d$, the two sets are $ρ$-correlated. The Procrustes matching problem asks us to recover the unknown permutation of $[n]$ that aligns the two sets. The problem is well-studied in the low-dimensional regime $d=O(\log n)$, but the high-dimensional regime $d\gg \log n$ has remained largely uncharted: prior matching guarantees require nearly perfect correlation $ρ=1-o(1)$, even for information-theoretic recovery. Our main result is a polynomial-time algorithm for exact recovery at constant correlation. The algorithm works by computing and comparing weighted counts of a specially chosen family of ``wide'' trees. So long as $d\ge \mathrm{polylog}(n)$, the algorithm succeeds with high probability for any $ρ^2>\sqrtα$, where $α\approx 0.338$ is Otter's tree-counting constant. We complement this algorithmic result with an improved information-theoretic guarantee, showing that exact recovery is possible when $ρ^2 \gtrsim \max\{\log n/d,\sqrt{\log n/n}\}$. We also carry out a low-degree advantage calculation, which suggests that the condition $ρ^2 > \sqrtα$ is necessary for any tree-counting algorithm.

Subjects: Machine Learning , Information Theory , Machine Learning , Statistics Theory

Publish: 2026-07-09 14:33:47 UTC


#12 Testing Covariance Separability in High Dimensions [PDF] [Copy] [Kimi] [REL]

Authors: Tomas Masak, Marcus Mayrhofer, Una Radojičić

Separability is an important structural assumption often placed on the covariance when working with matrix-variate data, because it greatly simplifies both interpretation and computation of subsequent covariance-based statistical tasks. Yet testing the separability assumption is difficult in the high-dimensional regime. We propose to test separability by recasting the problem as a sphericity test after whitening the data using the separable maximum likelihood estimate of the covariance. The test is calibrated by Monte Carlo simulation, yielding finite-sample level control. Furthermore, we prove the test's high-dimensional consistency under dense alternatives. To reduce its reliance on distributional assumptions, we introduce an angular version of the test based on radial normalization after whitening. We demonstrate the practical utility, empirical power, and computational efficiency of the prop

Subjects: Methodology , Statistics Theory

Publish: 2026-07-09 12:11:37 UTC


#13 Why Constants Matter in Distribution Testing: From Uniformity to Calibration [PDF] [Copy] [Kimi] [REL]

Author: Alon Kipnis

Distribution goodness-of-fit testing has developed a powerful rate-level theory: we often know how the required sample size scales with the alphabet size, the separation from the null, and the target error probability. Uniformity testing is the canonical example. One can distinguish the uniform distribution on $N$ categories from alternatives at total-variation distance at least $ε$ with far fewer than $N$ samples, and the optimal scaling is now well understood. But rate-level theory leaves an important question unresolved: among several tests with the same sample-complexity order, which one actually gives the best risk or power? This is a constant-level question. It is especially relevant in modern applications where distribution testing is used not merely as an asymptotic abstraction, but as a practical design tool. This note argues that sharp constants in distribution testing play a role analogous to Fisher information in parametric estimation and Pinsker's constant in nonparametric estimation. First, they distinguish between tests that are all rate-optimal but not equally powerful. Second, they reveal the effective signal-to-noise ratio governing the testing problem. Third, they can guide tuning-parameter choices in downstream applications. We illustrate this perspective through large-alphabet uniformity testing and then explain why the same logic matters for choosing the number of bins in calibration testing.

Subjects: Information Theory , Statistics Theory

Publish: 2026-07-09 11:55:50 UTC


#14 Bias-Corrected Multiplier Bootstrap Inference for Spectral Edges of Large Covariance Matrices [PDF] [Copy] [Kimi] [REL]

Authors: Xiucai Ding, Yichen Hu, Jiahui Xie

Inference for spectral edges of large covariance matrices is a fundamental problem in high-dimensional statistics. A major difficulty is that the largest non-spiked sample eigenvalues, which serve as natural estimators of the edge, fluctuate on the Tracy--Widom scale. Consequently, valid inference requires accurate centering by the deterministic spectral edge together with a precise scaling constant, both of which are often difficult to estimate in practice under general unknown population covariance structures. In this paper, we propose a bias-corrected multiplier bootstrap procedure for inference on the deterministic edge of the bulk spectrum. The key idea is to introduce a carefully calibrated multiplier perturbation that regularizes the edge fluctuation to a slightly larger scale at which Gaussian approximation becomes tractable. The resulting confidence interval is constructed directly from bootstrap eigenvalues, together with a data-driven recentering step that corrects the bootstrap-induced shift of the deterministic edge. On the theoretical side, we show that, after bias correction and rescaling, the largest few non-spiked bootstrap eigenvalues are asymptotically Gaussian conditionally on the data. Building on this result, we establish the asymptotic validity of the proposed confidence interval, whose length is only slightly larger than the Tracy--Widom scale, and prove vanishing coverage under alternatives in which additional spikes separate from the bulk at a local scale larger than $n^{-1/6}$. As a consequence, the same confidence interval yields a threshold-free estimator for the number of spikes, without requiring the spikes to be distinct or very large. Equivalently, the procedure yields a data-driven and theoretically justified cutoff for the scree plot.

Subjects: Methodology , Statistics Theory

Publish: 2026-07-09 03:50:10 UTC