arXiv++ Combinatorics

Browse math.CO papers from arXiv

math.PR ↗ arXiv

440 papers in this category
Combinatorial Hopf algebras in noncommutative probabilility
We prove that the generalized moment-cumulant relations introduced in [arXiv:1711.00219] are given by the action of the Eulerian idempotents on the Solomon-Tits algebras, whose direct sum builds up the Hopf algebra of Word Quasi-Symmetric Functions $\WQSym$. We prove $t$-analogues of these identities (in which the coefficient of $t$ gives back the original version), and a similar $t$-analogue of Goldberg's formula for the coefficients of the Hausdorff series. This amounts to the determination of the action of all the Eulerian idempotents on a product of exponentials.
2026-09-23 v4
Upper tail bounds for irregular graphs
We consider the upper tail large deviations of subgraph counts for irregular graphs $\mathrm{H}$ in $\mathbb{G}(n,p)$, the sparse Erdős-Rényi graph on $n$ vertices with edge connectivity probability $p \in (0,1)$. For $n^{-1/Δ} \ll p \ll 1$, where $Δ$ is the maximum degree of $\mathrm{H}$, we derive the upper tail large deviations for any irregular graph $\mathrm{H}$. On the other hand, we show that for $p$ such that $1 \ll n^{v_{\mathrm{H}}} p^{e_{\mathrm{H}}} \ll (\log n)^{α^{*}_{\mathrm{H}}/\left(α^{*}_{\mathrm{H}}-1\right)}$, where $v_{\mathrm{H}}$ and $e_{\mathrm{H}}$ denote the number of vertices and edges of $\mathrm{H}$, and $α^*_{\mathrm{H}}$ denotes the fractional independence number, the upper tail large deviations of the number of unlabelled copies of $\mathrm{H}$ in $\mathbb{G}(n,p)$ is given by that of a sequence of Poisson random variables with diverging mean, for any strictly balanced graph $\mathrm{H}$. Restricting to the $r$-armed star graph we further prove a localized behavior in the intermediate range of $p$ (left open by the above two results) and show that the mean-field approximation is asymptotically tight for the logarithm of the upper tail probability. This work further identifies the typical structures of $\mathbb{G}(n,p)$ conditioned on upper tail rare events in the localized regime.
2026-09-22
Down-the-middle isometries for the Airy and KPZ sheets
We find a new coupling between the Airy sheet $\mathcal{S}$ and the Airy line ensemble $\{ \mathcal{A}_n \}_{n \in \mathbb{N}}$, which we call the down-the-middle isometry. In this coupling, the Airy sheet is encoded by last passage values that start at $\mathcal{A}_k(0)$ for some $k \in \mathbb{N}$ and end on the top line $\mathcal{A}_1$. Using this coupling, we give a new construction of the (extended) Airy sheet as a scaling limit of Brownian last passage percolation and study coalescence of geodesics in the directed landscape. This coupling extends to discrete polymers and last passage percolation models, as well as to the KPZ line ensemble and KPZ sheet.
2026-09-22
Counterexamples to the Ramos conjecture for two hyperplanes
For every $n\ge2$, we construct $4n-2$ nondegenerate Gaussian measures on $\mathbb{R}^{6n-3}$ that cannot be simultaneously equipartitioned by two affine hyperplanes. This disproves the Ramos conjecture for two hyperplanes. Combined with known upper bounds, the construction shows that $3\cdot2^{s-1}-2$ is the least dimension guaranteeing a common two-hyperplane equipartition of $2^s-2$ absolutely continuous probability measures, for every $s\ge3$. We characterize the Gaussian equipartition threshold in terms of the least number of positive definite quadratic measurements needed for phase retrieval. Modified complex polynomial multiplication gives $2r-2$ positive definite measurements in every even dimension $r\ge4$. This number is optimal when $r=2^k+2$, $k\ge1$.
2026-09-22
Concentration of Regularized Sparse Random Matrices: Spectral Edge Bounds via Nonbacktracking Operators
In sparse random matrices, spectral outliers (eigenvalues and singular values located away from the bulk) emerge due to degree fluctuations: high degrees inflate the operator norm, while low column degrees reduce the least singular value. As proved by Feige and Ofek (2005) and Le, Levina, and Vershynin (2017), degree regularization enforces concentration at the expected norm scale. However, precise bounds incorporating the cutoffs remain unexplored and challenging since regularization introduces dependencies among entries. For the first time in the literature, we provide variance- and cutoff-dependent bounds for extreme singular values and eigenvalues of regularized inhomogeneous random matrices. In the absence of regularization, our lower bound for the least singular value matches the same leading constant obtained by Brailovskaya and van Handel (2024). Moreover, our error term vanishes under the milder condition $d/\log N\to\infty$, as opposed to their stronger requirement $d/(\log N)^4\to\infty$. A key ingredient is to extend spectral radius bounds for nonbacktracking matrices to the dependent setting. We build on approaches for independent cases established by Benaych-Georges, Bordenave, and Knowles (2020), as well as Dumitriu and Zhu (2024), and carefully handle edges traversed only once. Our proof framework separates deterministic spectral comparisons from probabilistic estimates: once Loewner inequalities and columnwise variance controls are established, the remaining probabilistic analysis boils down to verifying the graph moment conditions formulated in this paper. We hope this framework can be extended to handle general random matrices with more complex dependencies.
Extremal subtrees of critical beta-splitting trees
We determine the most and least likely shapes for an instance of the critical beta-splitting tree via a connection to data compression and Huffman's minimum redundancy codes. This allows us to answer combinatorial questions about the distribution of clades posed by Aldous and Janson, stated as problem 7 in arXiv:2303.02529.
2026-09-21
Dictators are most informative
We prove the Courtade-Kumar conjecture: among all Boolean functions $f\colon \{-1,1\}^n\to\{-1,1\}$, a dictator retains the most information about a uniformly random input observed through independent binary noise.
2026-09-21 v2
Graph Decompositions at the Expectation Threshold
For \(n\ge3\) and a graph \(H\) on at most \(n\) vertices, let \(q(H)\) be the expectation threshold for its containment in \(G(n,p)\). We prove that there are absolute constants \(a,L>0\) such that, for every \(C>0\), every graph of degeneracy at most \(C\log n/\log\log n\) has a deterministic edge decomposition into at most \(\lceil a(C+1)\rceil\) pieces, each with ordinary containment threshold at most \(Lq(H)\). This removes the maximum-degree hypothesis from a theorem of Ascoli, He, Park, and Talagrand in their approach to Talagrand's discrete convexity problem. With constants depending on the fixed parameters, a structural extension allows the addition of \(O(\log n)\) vertices with arbitrary incident edges. The uniform main bound also gives an \(O(1+\log\log n)\)-piece decomposition for every target graph, with an absolute threshold multiplier. The main ingredient converts Li's two-set coupling for spread measures into a partition of a prescribed neighborhood list, fixed before the random host is sampled. It allows arbitrary overlaps and repetitions and controls all Hall matching conditions simultaneously by bounding target demand and host supply through common two-sided approximations under the biased host measure.
2026-09-21 v3
Pal's permanent conjecture: proof for block uniform matrices
Consider a symmetric function $\mathcal{C}(x,y)$ on $[0,1]\times[0,1]$ which is twice continuously differentiable up to the boundary, and which satisfies $ \mathcal{C}(x,y)=\mathcal{C}(1-x,1-y)$. Let $A^{(n)} = \big(a^{(n)}_{i,j}\, :\, i,j \in [n]\big)$ be the matrix with entries $a^{(n)}_{i,j}\, =\, \exp(-\mathcal{C}(i/n,j/n))$. Soumik Pal conjectured the asymptotics $$\operatorname{perm}\big(A^{(n)}\big)/n!\sim \exp\big(n Λ[\mathcal{C}]\big)/ \sqrt{\mathcal{D}[\mathcal{C}]}$$ as $n \to \infty$ for known functionals that arise naturally in the context of entropy regularized optimal transport. The functional $Λ[\mathcal{C}]$ is the known large deviation rate function, already proved rigorously by Sumit Mukherjee. It is $\int_{0}^1 \int_0^{1} (α(x)+β(y))\, dx\, dy$ where $α(x)+β(y)$ is chosen such that $ρ(x,y) := \exp(-\mathcal{C}(x,y)-α(x)-β(y))$ has uniform marginals. The algebraic term $\mathcal{D}[c]$ is given by Peter McCullagh's formula for doubly stochastic matrices: $\operatorname{det}_F(I+J-T^*T)$, the Fredholm determinant, where $I$ is the identity on $L^2([0,1])$, $Jf(x) \equiv \int_{0}^1 f(z)\, dz$ (for all $x$) and $Tf(x) = \int_0^1 ρ(x,y) f(y)\, dy$. We prove the conjecture for functions $\mathcal C$ that are constant on blocks, exploiting a well-known Ross Pinsky's combinatorial decomposition of permutations in blocks.
2026-09-21 v2
Permutations from Random Walk
Xavier and Yushi run a "random race" as follows. An atomless probability distribution $μ$ on the real line is chosen. The runners begin at zero. At time $i$ Xavier draws $\mathbf{X}_i$ from $μ$ and advances that distance, while Yushi advances by an independent drawing $\mathbf{Y}_i$. After $n$ such moves, what is the probability that Yushi led all the way? That the answer (namely, $4^{-n}\binom{2n}{n}$) is independent of $μ$ follows from a classical theorem of Darling, stating that for symmetric atomless increments, the distribution of each individual rank in the permutation obtained by ranking the partial sums is independent of the step law. We give a self-contained proof and extend the result to the permutations generated by partial sums of uniformly random signed permutations of any fixed, finite, generic set of reals. For atomless increments with mean zero and finite variance, without assuming symmetry, we show that random-walk permutations approach a random object that we call the "Wiener permuton," whose expected pattern densities equal the probabilities of the corresponding permutations generated by finite random walks with centered Laplace increments. Finally, we exhibit an infinite family of constructions whose limiting permutons interpolate between the Wiener permuton and the recursive separable permuton; each has the same intensity permuton, providing a single two-dimensional extension of the classical arcsine law for all of them.
2026-09-21 v3
Asymptotics for the harmonic descent chain and applications to critical beta-splitting trees
Motivated by the connection to a probabilistic model of phylogenetic trees introduced by Aldous, we study the recursive sequence governed by the rule $x_n = \sum_{i=1}^{n-1} \frac{1}{h_{n-1}(n-i)} x_i$ where $h_{n-1} = \sum_{j=1}^{n-1} 1/j$, known as the harmonic descent chain. While it is known that this sequence converges to an explicit limit $x$, not much is known about the rate of convergence. We first show that a class of recursive sequences including the above are decreasing and use this to bound the rate of convergence. Moreover, for the harmonic descent chain we prove the asymptotic $x_n - x = n^{-γ_* + o(1)}$ for an implicit exponent $γ_*$. As a consequence, we deduce central limit theorems for various statistics of the critical beta-splitting random tree. This answers a number of questions of Aldous, Janson, and Pittel.
2026-09-20 v7
New matrix perturbation bounds with relative strength: Perturbation of eigenspaces
Matrix perturbation bounds (such as Weyl and Davis--Kahan) are used abundantly in many areas of mathematics and data science. Many bounds (such as the above two) involve the spectral norm of the noise matrix and are sharp in worst-case analysis. In order to refine these classical bounds, we introduce a new parameter, which we refer to as the relative strength. This parameter measures the strength of the action of the noise matrix on the relevant eigenvectors of the ground matrix. It has turned out that in a number of situations, we can use the relative strength as a replacement for the spectral norm (which can be seen as the absolute strength). This has led to a number of notable improvements under certain sets of assumptions, which are frequently met in practice. A representative example is the case when the noise matrix is random. For the purpose of our study, we introduce a new method of analysis, which combines the classical contour integral argument with new (combinatorial) ideas. This method is robust and of independent interest. In the current paper, we focus on the perturbation of eigenspaces (Davis--Kahan type results). Perturbation bounds for eigenspaces are essential in statistics and theoretical computer science, and thus deserve a special treatment. Furthermore, this will lay the ground for the more technical treatment of general matrix functionals, which appears in a future paper.
2026-09-19 v2
The speed of convergence in the Cooper-Dutle dueling game
In 2013 Cooper and Dutle invented a dueling scenario where Alice and Bob shoot at each other until one is hit. Each shot is successful with some fixed probability $p$, $0 < p < 1$. The shooting order is given by a greedy algorithm, where at each step a shot is assigned to the player whose current probability of success is smaller. Cooper and Dutle observed that as $p \rightarrow 0$, the resulting sequence of shots (by Alice or Bob) converges to the infinite Thue-Morse sequence $\mathbf{t}$, but left the speed of convergence as an open problem. In this note we determine the speed of this convergence.
2026-09-19 v2
Multiplicative comparisons of Rényi entropies for weighted Bernoulli sums
We establish multiplicative comparisons between Rényi entropies of different orders for weighted sums of independent Bernoulli random variables. In particular, we prove a logarithmic comparison between the zeroth- and infinity-order Rényi entropies, which yields a polynomial improvement over the square-root bound of Jain, Sah, and Sawhney. As an application, this leads to an improved parameterized running time for the randomized bin-packing algorithm of Nederlof, Pawlewicz, Swennenhuis, and Wȩgrzycki. We also obtain explicit dimension-free, constant-factor comparisons between Rényi entropies of positive orders.
2026-09-18
Directed distances in spanning-tree-decorated planar maps: exact exponent, scaling limit and universality
We define a natural orientation on a spanning-tree-decorated planar map whereby, roughly speaking, each directed edge in the map is oriented to match the direction of the contour exploration of the spanning tree. We study directed distances (lengths of shortest directed paths) with respect to this orientation. We construct the Busemann function which measures directed distances to $\infty$ along a natural interface in the uniform infinite spanning-tree-decorated map. We show that this Busemann function, re-scaled appropriately, converges in law to a $3/2$-stable Lévy process. We also show that in a uniform spanning-tree-decorated map with $n$ edges, directed distances are typically of order $n^{1/3}$. Using a strong coupling argument, we deduce analogous statements for directed distances in other random planar maps in the $\sqrt 2$-Liouville quantum gravity (LQG) universality class, including uniform meandric systems and mated-CRT maps for $γ=\sqrt 2$. These results give the scaling dimension for a hypothetical directed version of the $\sqrt 2$-LQG metric. Our proof strategy is inspired by work of Borga and Gwynne (2025) on directed distances in bipolar-oriented triangulations.
2026-09-18
Gaussian Vertex-Face Balance in Random Convex Polyhedra with Fixed Edge Count
Choose uniformly among the combinatorial types of convex three-dimensional polyhedra with a fixed admissible number $e$ of edges, and let $V_e$ be the number of vertices. This resolves a fixed-edge limit-distribution question posed by Rüdinger: if $β_e=V_e/(e+2)$, then $\sqrt e(β_e-1/2)\Rightarrow N(0,1/32)$, equivalently $\operatorname{Var}(V_e)\sim e/32$. Using the classical rooted enumeration and asymmetry results of Bender and Wormald, we derive a relative lattice local limit theorem on every $o(e^{3/4})$ window, a quartic correction from the rate function on every $o(e^{5/6})$ window, precise moderate-tail constants, a quadratic moderate-deviation principle for every $a_e\to\infty$ with $a_e=o(\sqrt e)$, and a full speed-$e$ large-deviation principle with an explicit good rate function. All fixed standardised moments converge. When $e=3m$, the two extremal vertex counts have equal probability asymptotic to $\frac{6561}{32\sqrt2}(4/27)^m$. Rooted and unrooted fixed-edge laws are also uniformly exponentially close, with relative discrepancy $O(ρ^e)$.
2026-09-18
Local large deviations for triangles in sparse random graphs
We revisit a classic topic in probabilistic combinatorics, the lower-tail large-deviation problem for triangles in the random graph $G(n,p)$. Here we aim for first-order asymptotics for the quantity $\mathbb{P}[X=k]$ with $X$ the number of triangles in $G(n,p)$ and $0 \le k \le (1-ε) \mathbb{E}X$, in the sparse regime in which the logarithmic asymptotics are Poissonian. When $k=0$ (the case of triangle-freeness) and $p = p(n)$ is sufficiently small, first-order asymptotics are known via Janson's inequality and results of Stark and Wormald; when $k$ is sufficiently close to $\mathbb{E}X$, first-order asymptotics are known via local central limit theorems. Our main result gives first-order asymptotics for all $k$ in the above range when $p = o(n^{-2/3})$, improving upon the result of Frieze that required $p = o(n^{-4/5})$. We also characterize, up to vanishing total variation distance, the distribution of the $k$ triangles in the corresponding conditional distribution and give an efficient algorithm to approximately sample from this conditional distribution. Notably, when $k = Θ(μ)$ there is a transition at $p = Θ(n^{-4/5})$ from an asymptotically uniform triangle distribution to a distribution asymptotically singular to uniform.
A short proof of the Bernoulli decomposition theorem
We give a construction of admissible partitions that yields the Bernoulli decomposition theorem
2026-09-18 v3
A Resolution of the McCarty Conjecture
The McCarty Conjecture states that any McCarty Matrix (an $n\times n$ matrix $A$ with positive integer entries and each of the $2n$ row and column sums equal to $n$), can be additively decomposed into two other matrices, $B$ and $C$, such that $B$ has row and column sumsets both equal to $\{1, 2,... n\}$, and $C$ has row and column sumsets both equal to $\{0, 1,... n-1\}$. The problem can also be formulated in terms of bipartite graphs. In this paper we use probabilistic methods to resolve this conjecture.
Unimodality of Independence Polynomials for Sufficiently Large Forests
We prove that the independence sequence of every sufficiently large forest is unimodal. The result follows from establishing log-concavity on a central interval of the sequence, along with monotonicity of the initial and final segments. The main analytic step is a central limit theorem for the size of a random independent set sampled with the hard-core model, uniform over all forests and over an interval of positive fugacities.