stat.TH ↗ arXiv
6 papers in this category
Concentration of Regularized Sparse Random Matrices: Spectral Edge Bounds via Nonbacktracking Operators
In sparse random matrices, spectral outliers (eigenvalues and singular values located away from the bulk) emerge due to degree fluctuations: high degrees inflate the operator norm, while low column degrees reduce the least singular value. As proved by Feige and Ofek (2005) and Le, Levina, and Vershynin (2017), degree regularization enforces concentration at the expected norm scale. However, precise bounds incorporating the cutoffs remain unexplored and challenging since regularization introduces dependencies among entries.
For the first time in the literature, we provide variance- and cutoff-dependent bounds for extreme singular values and eigenvalues of regularized inhomogeneous random matrices. In the absence of regularization, our lower bound for the least singular value matches the same leading constant obtained by Brailovskaya and van Handel (2024). Moreover, our error term vanishes under the milder condition $d/\log N\to\infty$, as opposed to their stronger requirement $d/(\log N)^4\to\infty$. A key ingredient is to extend spectral radius bounds for nonbacktracking matrices to the dependent setting. We build on approaches for independent cases established by Benaych-Georges, Bordenave, and Knowles (2020), as well as Dumitriu and Zhu (2024), and carefully handle edges traversed only once. Our proof framework separates deterministic spectral comparisons from probabilistic estimates: once Loewner inequalities and columnwise variance controls are established, the remaining probabilistic analysis boils down to verifying the graph moment conditions formulated in this paper. We hope this framework can be extended to handle general random matrices with more complex dependencies.
New matrix perturbation bounds with relative strength: Perturbation of eigenspaces
Matrix perturbation bounds (such as Weyl and Davis--Kahan) are used abundantly in many areas of mathematics and data science. Many bounds (such as the above two) involve the spectral norm of the noise matrix and are sharp in worst-case analysis. In order to refine these classical bounds, we introduce a new parameter, which we refer to as the relative strength. This parameter measures the strength of the action of the noise matrix on the relevant eigenvectors of the ground matrix. It has turned out that in a number of situations, we can use the relative strength as a replacement for the spectral norm (which can be seen as the absolute strength). This has led to a number of notable improvements under certain sets of assumptions, which are frequently met in practice. A representative example is the case when the noise matrix is random.
For the purpose of our study, we introduce a new method of analysis, which combines the classical contour integral argument with new (combinatorial) ideas. This method is robust and of independent interest. In the current paper, we focus on the perturbation of eigenspaces (Davis--Kahan type results). Perturbation bounds for eigenspaces are essential in statistics and theoretical computer science, and thus deserve a special treatment. Furthermore, this will lay the ground for the more technical treatment of general matrix functionals, which appears in a future paper.
Sharp spectral norm concentration of sparse random tensors
We prove a sharp concentration inequality for the spectral norm of sparse random tensors with independent Bernoulli entries. Let $T$ be an order-$k$ tensor of dimension $n\times\cdots\times n$ with independent Bernoulli$(p)$ entries, where $k$ is fixed. For any $c,r>0$, we show that $\|T-\mathbb E T\|\le C_{k,r,c}\sqrt{np}$ with probability at least $1-n^{-r}$ whenever $np\ge c\log n$. We extend this bound to inhomogeneous Bernoulli sampling with deterministic entrywise weights. This removes the logarithmic factor in the work of Zhou and Zhu (2021). The proof follows the Kahn--Szemerédi light--heavy decomposition with a refined estimate on the heavy tuple part. We also obtain a log-free second eigenvalue bound for the random hypergraph model of Friedman and Wigderson (1995).
An Alon-Boppana Bound for the Non-Backtracking Operator
For any fixed $k$, we prove a lower bound on the $k$th largest modulus of an eigenvalue of the non-backtracking matrix $B$. Specifically, consider any deterministic or random family of graphs that converges locally to the unimodular Galton-Watson tree with root degree distribution $D$, and set $κ:=\mathbb E[D(D-1)]/\mathbb E[D]$. Given $κ>1$ and an exponential-moment bound on the empirical degree distributions, we show that $|λ_k(B)|\geq\sqrtκ-o_N(1)$, where $N$ is the number of vertices. When restricted to locally tree-like regular graphs, this recovers a well-known consequence of the Ihara-Bass formula. In the specific case where the graph is generated through the Erdős-Rényi model with expected degree $d>1$, this proves a conjecture of Bordenave, Lelarge, and Massoulié.
To do this, we show that the normalized log-determinant of the Bethe-Hessian of the graph is bounded by that of the Bethe-Hessian of its local limit. This bound is violated if the eigenvalues of the non-backtracking matrix are too small. We establish this using an effective-conductance interpretation of the tree Green's function recursion.
Identifiability of Nonnegative Tensor Decompositions via Positive Scattering
Identifiability of tensor decompositions is often established through linear-algebraic conditions on the factor families. For nonnegative decompositions, however, positivity provides additional information that is not captured by dimension and independence alone: nonnegative terms cannot cancel, and their supports constrain competing decompositions. We introduce a positive scattering term that quantifies this additional source of identifiability and combine it with the dimension budget underlying the Lovitz--Petrov generalization of Kruskal's theorem. For every subset of components, we obtain two sufficient conditions: a threshold of $2|S|-2$ guarantees minimality and nonnegative rank, while the stronger threshold $2|S|-1$ guarantees uniqueness among nonnegative decompositions of the same length. The key result is a positive splitting inequality for irreducible exchanges of nonnegative rank-one tensors, which combines the dimension constraint with support-induced geometric rigidity. Although the scattering term is defined through an optimization over intermediate factor spaces, we show that its mode costs are exactly $0$, $1$, or $+\infty$, yielding an exact activation characterization in terms of graph connectivity. The resulting criterion can strictly certify sparse nonnegative tensor decompositions beyond the reach of Kruskal and Lovitz--Petrov conditions, including examples for which those conditions fail even after reshaping. In the matrix case, the two criteria reduce respectively to full-rank factorization and two-sided separability.
Random Width and Brightness: Polyhedral Density Theory, Reconstruction, and Gaussian Identifiability
Published
• View Publication
• BIB
Let U be uniformly distributed on the unit sphere. We develop a self-contained forward and inverse theory for the random width w_K(U) and brightness b_K(U) of three-dimensional convex bodies. For every full-dimensional polytope, a global spherical co-area formula expresses the width density as a finite sum of angular apertures determined by the normal fan of its difference body; in particular, the density is piecewise real analytic with a finite geometrically determined critical set. This theory yields exact densities for the width of the regular tetrahedron, resolving a question of Finch, and for the regular truncated octahedron, together with the tetrahedral brightness law and the equivalent rhombic-dodecahedral width law. On the inverse side, second- and third-order polarized cosine-transform moments reconstruct finite labelled direction systems whenever the observed triangles span the cycle space of the correlation graph; signed-graph switching describes the unavoidable ambiguity. In contrast, equal three-dimensional intrinsic volumes do not determine either the width law or the brightness law, even for centrally symmetric bodies. Removing the spatial rank constraint gives a dimension-free identifiability theorem for centered multivariate folded-normal vectors: pairwise absolute moments and an anchored family of triple absolute moments, comprising |m - 1|^2 labelled observations for a complete correlation graph, determine the correlation matrix up to diagonal sign conjugacy without fourth-order moments. A harmonic decomposition further identifies the degree-two variance contribution as a constant multiple of the squared Frobenius norm of the traceless part of the weighted frame operator and explains why this contribution vanishes under irreducible symmetry.