random
7100 papers tagged with this keyword
Heating Up Quasi-Monte Carlo Graph Random Features: A Diffusion Kernel Perspective
We build upon a recently introduced class of quasi-graph random features (q-GRFs), which have demonstrated the ability to yield lower variance estimators of the 2-regularized Laplacian kernel (Choromanski 2023). Our research investigates whether similar results can be achieved with alternative kernel functions, specifically the Diffusion (or Heat), Matérn, and Inverse Cosine kernels. We find that the Diffusion kernel performs most similarly to the 2-regularized Laplacian, and we further explore graph types that benefit from the previously established antithetic termination procedure. Specifically, we explore Erdős-Rényi and Barabási-Albert random graph models, Binary Trees, and Ladder graphs, with the goal of identifying combinations of specific kernel and graph type that benefit from antithetic termination. We assert that q-GRFs achieve lower variance estimators of the Diffusion (or Heat) kernel on Ladder graphs. However, the number of rungs on the Ladder graphs impacts the algorithm's performance; further theoretical results supporting our experimentation are forthcoming. This work builds upon some of the earliest Quasi-Monte Carlo methods for kernels defined on combinatorial objects, paving the way for kernel-based learning algorithms and future real-world applications in various domains.
A Random-Walk Concentration Principle for Occupancy Processes on Finite Graphs
This paper concerns discrete-time occupancy processes on a finite graph. Our results can be formulated in two theorems, which are stated for vertex processes, but also applied to edge process (e.g., dynamic random graphs). The first theorem shows that concentration of local state averages is controlled by a random walk on the graph. The second theorem concerns concentration of polynomials of the vertex states. For dynamic random graphs, this allows to estimate deviations of edge density, triangle density, and more general subgraph densities. Our results only require Lipschitz continuity and hold for both dense and sparse graphs.
Gaussian to log-normal transition for independent sets in a percolated hypercube
Independent sets in graphs, i.e., subsets of vertices where no two are adjacent, have long been studied, for instance as a model of hard-core gas. The $d$-dimensional hypercube, $\{0,1\}^d$, with the nearest neighbor structure, has been a particularly appealing choice for the base graph, owing in part to its many symmetries. Results go back to the work of Korshunov and Sapozhenko who proved sharp results on the count of such sets as well as structure theorems for random samples drawn uniformly. Of much interest is the behavior of such Gibbs measures in the presence of disorder. In this direction, Kronenberg and Spinka [KS] initiated the study of independent sets in a random subgraph of the hypercube obtained by considering an instance of bond percolation with probability $p$. Relying on tools from statistical mechanics they obtained a detailed understanding of the moments of the partition function, say $\mathcal{Z}$, of the hard-core model on such random graphs and consequently deduced certain fluctuation information, as well as posed a series of interesting questions. In particular, they showed in the uniform case that there is a natural phase transition at $p=2/3$ where $\mathcal{Z}$ transitions from being concentrated for $p>2/3$ to not concentrated at $p=2/3$.
In this article, developing a probabilistic framework, as well as relying on certain cluster expansion inputs from [KS], we present a detailed picture of both the fluctuations of $\mathcal{Z}$ as well as the geometry of a randomly sampled independent set. In particular, we establish that $\mathcal{Z}$, properly centered and scaled, converges to a standard Gaussian for $p>2/3$, and to a sum of two i.i.d. log-normals at $p=2/3$. A particular step in the proof which could be of independent interest involves a non-uniform birthday problem for which collisions emerge at $p=2/3$.
Smoothed analysis for graph isomorphism
Published
• View Publication
• BIB
There is no known polynomial-time algorithm for graph isomorphism testing, but elementary combinatorial "refinement" algorithms seem to be very efficient in practice. Some philosophical justification is provided by a classical theorem of Babai, Erdős and Selkow: an extremely simple polynomial-time combinatorial algorithm (variously known as "naïve refinement", "naïve vertex classification", "colour refinement" or the "1-dimensional Weisfeiler-Leman algorithm") yields a so-called canonical labelling scheme for "almost all graphs". More precisely, for a typical outcome of a random graph $G(n,1/2)$, this simple combinatorial algorithm assigns labels to vertices in a way that easily permits isomorphism-testing against any other graph.
We improve the Babai-Erdős-Selkow theorem in two directions. First, we consider randomly perturbed graphs, in accordance with the smoothed analysis philosophy of Spielman and Teng: for any graph $G$, naïve refinement becomes effective after a tiny random perturbation to $G$ (specifically, the addition and removal of $O(n\log n)$ random edges). Actually, with a twist on naïve refinement, we show that $O(n)$ random additions and removals suffice. These results significantly improve on previous work of Gaudio-Rácz-Sridhar, and are in certain senses best-possible.
Second, we complete a long line of research on canonical labelling of random graphs: for any $p$ (possibly depending on $n$), we prove that a random graph $G(n,p)$ can typically be canonically labelled in polynomial time. This is most interesting in the extremely sparse regime where $p$ has order of magnitude $c/n$; denser regimes were previously handled by Bollobás, Czajka-Pandurangan, and Linial-Mosheiff. Our proof also provides a description of the automorphism group of a typical outcome of $G(n,p_n)$ (slightly correcting a prediction of Linial-Mosheiff).
Spread blow-up lemma with an application to perturbed random graphs
Combining ideas of Pham, Sah, Sawhney, and Simkin on spread perfect matchings in super-regular bipartite graphs with an algorithmic blow-up lemma, we prove a spread version of the blow-up lemma. Intuitively, this means that there exists a probability measure over copies of a desired spanning graph $H$ in a given system of super-regular pairs which does not heavily pin down any subset of vertices. This allows one to complement the use of the blow-up lemma with the recently resolved Kahn-Kalai conjecture. As an application, we prove an approximate version of a conjecture of Böttcher, Parczyk, Sgueglia, and Skokan on the threshold for appearance of powers of Hamilton cycles in perturbed random graphs.
On the $H$-space of a random graph
The edge space $\mathcal{E}(G)$ of a graph $G$ is the vector space $\mathbb{F}_2^{E(G)}$ with members naturally identified with subgraphs of $G$, and the $H$-space is the subspace $\mathcal{C}_H(G)$ of $ \mathcal{E}(G)$ spanned by copies of the graph $H$. We are interested in when the random graph $G = G_{n,p}$ is likely to satisfy \[\mathcal{C}_H(G) = \mathcal{W}_H(G),\] where $\mathcal{W}_H(G)$ takes one of four natural values, depending on the value of $\mathcal{C}_H(K_n)$. We show that for strictly $2$-balanced $H$, w.h.p. the above equality holds whenever every edge of $G$ is in a copy of $H$.
On the local convergence of integer-valued Lipschitz functions on regular trees
Published
• View Publication
• BIB
We study random integer-valued Lipschitz functions on regular trees. It was shown by Peled, Samotij and Yehudayoff that such functions are localized, however, finer questions about the structure of Gibbs measures remain unanswered. Our main result is that the weak limit of a uniformly chosen 1-Lipschitz function with 0 boundary condition on a $d$-ary tree of height $n$ exists as $n \to \infty$ if $2 \le d \le 7$, but not if $d \ge 8$, thereby partially answering a question posed by Peled, Samotij and Yehudayoff. For large $d$, the value at the root alternates between being almost entirely concentrated on 0 for even $n$ and being roughly uniform on $\{-1,0,1\}$ for odd $n$, leading to different limits as $n$ approaches infinity along evens or odds. For $d \ge 8$, the essence of this phenomenon is preserved, which obstructs the convergence. For $d \le 7$, this phenomenon ceases to exist, and the law of the value at the root loses its connection with the parity of $n$. Along the way, we also obtain an alternative proof of localization. The key idea is a fixed point convergence result for a related operator on $\ell^\infty$, and a procedure to show that the iterations get into a `basin of attraction' of the fixed point. We also prove some accompanying analogous `even-odd phenomenon' type results about $M$-lipschitz functions on general non-amenable graphs with high enough expansion (this includes for example the large $d$ case for regular trees). We also prove a convergence result for 1-Lipschitz functions with $\{0,1\}$ boundary condition. This last result relies on an absolute value FKG for uniform 1-Lipschitz functions when shifted by $1/2$.
Hypergeometric Functions of Random Matrices and Quasimodular Forms
Hypergeometric functions of complex matrices were introduced by James in multivariate statistics. These special functions play many roles in random matrix theory. The main goal of this paper is to suggest a new use for them as holomorphic observables of the Circular Unitary Ensemble. We analyze the high-dimensional behavior of the expected derivatives of these random analytic functions, and show that they admit asymptotic expansions which can be described in terms of quasimodular forms, giving an apparently new connection between the CUE and number theory.
Polyhedral volume ratios, Izmestiev's Colin de Verdiere matrices and Spectral Gaps
We present a relation between volumes of certain lower dimensional simplices associated to a full-dimensional primal and polar dual polytope in R^k. We then discuss an application of this relation to a geometric construction of a Colin de Verdiere matrix by Ivan Izmestiev. In the second part of the paper, we introduce a variation of vertex transitive polytopes, translate their associated Colin de Verdiere matrices into random walk matrices, and investigate extremality properties of the spectral gaps of these random walk matrices in two concrete examples - permutahedra of Coxeter groups and polytopes associated to the pure rotational tetrahedral group - where maximal spectral gaps correspond to equilateral polytopes.
Simulating Simple Random Walks With a Deck of Cards
Published
• View Publication
• BIB
When we want to simulate the realization of a symmetric simple random walk on $\mathbb Z^d$, we use $(2d)$-side fair dice to decide to which neighbor it jumps at each step if $d\geq 2$ or we simply use a fair coin when $d=1$. Assume that instead of using a dice or a coin we want to do a simulation using a well shuffled deck with $K$ cards of each of the $2d$ suits. In the first step the probability of jumping to each neighbor is $(2d)^{-1}$, but from the second step it becomes biased. Of course if we continue performing this simulation, the total variation distance between its law and the law of the random walk will increase until all cards are used. In this paper we investigate the minimum number of cards $N=2d K$ that a deck must contain so that the total variation distance between the law of a $n$-step simulation and the law of a $n$-step realization of the random walk is smaller than a chosen threshold $\varepsilon \in (0,1)$. More generally, we prove that when $N=cn$ this distance converges, as $n \to \infty$, to a Gaussian profile which depends on $c\geq 2d$. Furthermore, our analysis shows that this Gaussian profile vanishes as $c \to \infty$, proving the convergence of a multivariate hypergeometric distribution to a multinomial distribution in total variation.
Free cumulants and freeness for unitarily invariant random tensors
We address the question of the asymptotic description of random tensors that are local-unitary invariant, that is, invariant by conjugation by tensor products of independent unitary matrices. We consider both the mixed case of a tensor with $D$ inputs and $D$ outputs, and the case where there is a factorization between the inputs and outputs, called pure, which includes the random tensor models extensively studied in the physics literature.
The finite size and asymptotic moments are defined using correlations of certain invariant polynomials encoded by $D$-tuples of permutations, up to relabeling equivalence. Finite size free cumulants associated to the expectations of these invariants are defined through invertible finite size moment-cumulants formulas.
Two important cases are considered asymptotically: pure random tensors that scale like a complex Gaussian, and mixed random tensors that scale like a Wishart tensor. In both cases, we derive a notion of tensorial free cumulants associated to first order invariants, through moment-cumulant formulas involving summations over non-crossing permutations. The pure and mixed cases involve the same combinatorics, but differ by the invariants that define the distribution at first order. In both cases, the tensorial free-cumulants of a sum of two independent tensors are shown to be additive. A preliminary discussion of higher orders is provided.
Tensor freeness is then defined as the vanishing of mixed first order tensorial free cumulants. The equivalent formulation at the level of asymptotic moments is derived in the pure and mixed cases, and we provide an algebraic construction of tensorial probability spaces, which generalize non-commutative probability spaces: random tensors converge in distribution to elements of these spaces, and tensor freeness of random variables corresponds to tensor freeness of the subspaces they generate.
A combinatorial approach to phase transitions in random graph isomorphism problems
We consider two independent Erdős-Rényi random graphs, with possibly different parameters, and study two isomorphism problems, a graph embedding problem and a common subgraph problem. Under certain conditions on the graph parameters we show a sharp asymptotic phase transition as the graph sizes tend to infinity. This extends known results for the case of uniform Erdős-Rényi random graphs. Our approach is primarily combinatorial, naturally leading to several related problems for further exploration.
Combinatorics of a dissimilarity measure for pairs of draws from discrete probability vectors on finite sets of objects
Motivated by a problem in population genetics, we examine the combinatorics of dissimilarity for pairs of random unordered draws of multiple objects, with replacement, from a collection of distinct objects. Consider two draws of size $K$ taken with replacement from a set of $I$ objects, where the two draws represent samples from potentially distinct probability distributions over the set of $I$ objects. We define the set of \emph{identity states} for pairs of draws via a series of actions by permutation groups, describing the enumeration of all such states for a given $K \geq 2$ and $I \geq 2$. Given two probability vectors for the $I$ objects, we compute the probability of each identity state. From the set of all such probabilities, we obtain the expectation for a dissimilarity measure, finding that it has a simple form that generalizes a result previously obtained for the case of $K=2$. We determine when the expected dissimilarity between two draws from the same probability distribution exceeds that of two draws taken from different probability distributions. We interpret the results in the setting of the genetics of polyploid organisms, those whose genetic material contains many copies of the genome ($K > 2$).
Creating Subgraphs in Semi-Random Hypergraph Games
Published
• View Publication
• BIB
The semi-random hypergraph process is a natural generalisation of the semi-random graph process, which can be thought of as a one player game. For fixed $r < s$, starting with an empty hypergraph on $n$ vertices, in each round a set of $r$ vertices $U$ is presented to the player independently and uniformly at random. The player then selects a set of $s-r$ vertices $V$ and adds the hyperedge $U \cup V$ to the $s$-uniform hypergraph. For a fixed (monotone) increasing graph property, the player's objective is to force the graph to satisfy this property with high probability in as few rounds as possible.
We focus on the case where the player's objective is to construct a subgraph isomorphic to an arbitrary, fixed hypergraph $H$. In the case $r=1$ the threshold for the number of rounds required was already known in terms of the degeneracy of $H$. In the case $2 \le r < s$, we give upper and lower bounds on this threshold for general $H$, and find further improved upper bounds for cliques in particular. We identify cases where the upper and lower bounds match. We also demonstrate that the lower bounds are not always tight by finding exact thresholds for various paths and cycles.
Bivariate exponential integrals and edge-bicolored graphs
Published in Le Matematiche, 80 (1), 167-187 (2025)
• View Publication
• BIB
We show that specific exponential bivariate integrals serve as generating functions of labeled edge-bicolored graphs. Based on this, we prove an asymptotic formula for the number of regular edge-bicolored graphs with arbitrary weights assigned to different vertex structures. The asymptotic behavior is governed by the critical points of a polynomial. As an application, we discuss the Ising model on a random 4-regular graph and show how its phase transitions arise from our formula.
Tree height and the asymptotic mean of the Colijn-Plazzotta rank of unlabeled binary rooted trees
Published
• View Publication
• BIB
The Colijn--Plazzotta ranking is a bijective encoding of the unlabeled binary rooted trees with positive integers. We show that the rank $f(t)$ of a tree $t$ is closely related to its height $h$, the length of the longest path from a leaf to the root. We consider the rank $f(τ_n)$ of a random $n$-leaf tree $τ_n$ under each of three models: (i) uniformly random unlabeled unordered binary rooted trees, or unlabeled topologies; (ii) uniformly random leaf-labeled binary trees, or labeled topologies under the uniform model; and (iii) random binary search trees, or labeled topologies under the Yule--Harding model. Relying on the close relationship between tree rank and tree height, we obtain results concerning the asymptotic properties of $\log \log f(τ_n)$. In particular, we find $\mathbb{E} \{\log_2 \log f(τ_n)\} \sim 2 \sqrt{πn}$ for uniformly random unlabeled ordered binary rooted trees and uniformly random leaf-labeled binary trees, and for a constant $α\approx 4.31107$, $\mathbb{E}\{\log_2 \log f(τ_n)\} \sim α\log n $ for leaf-labeled binary trees under the Yule--Harding model. We show that the mean of $f(τ_n)$ itself under the three models is largely determined by the rank $c_{n-1}$ of the highest-ranked tree -- the caterpillar -- obtaining an asymptotic relationship with $π_n c_{n-1}$, where $π_n$ is a model-specific function of $n$. The results resolve open problems, providing a new class of results on an encoding useful in mathematical phylogenetics.
The difference between the chromatic and the cochromatic number of a random graph
The cochromatic number $ζ(G)$ of a graph $G$ is the minimum number of colours needed for a vertex colouring where every colour class is either an independent set or a clique. Let $χ(G)$ denote the usual chromatic number. Around 1991 Erdős and Gimbel asked: For the random graph $G \sim G_{n, 1/2}$, does $χ(G)-ζ(G) \rightarrow \infty$ whp? Erdős offered \$100 for a positive and \$1,000 for a negative answer.
We give a positive answer to this question for roughly 95% of all values $n$.
The hitting time of nice factors
Consider the random $u$-uniform hypergraph (or $u$-graph) process on $n$ vertices, where $n$ is divisible by $r>u\ge 2$. It was recently shown that with high probability, as soon as every vertex is covered by a copy of the complete $u$-graph $K_r$, it also contains a $K_r$-factor (RSA, Vol. 65 II, Sept. 2024). The hitting time result is obtained using a process coupling, which is based on the proof of the corresponding sharp threshold result (RSA, Vol. 61 IV, Dec. 2022). The latter, however, was not only derived for complete $u$-graphs, but for a broader class of so-called nice $u$-graphs.
The purpose of this article is to extend the process coupling for complete $u$-graphs to the full scope of the sharp threshold result: nice $u$-graphs. As a byproduct, we obtain the extension of the hitting time result to nice $u$-graphs. Since the relevant combinatorial bounds in the proof for the $K_r$-case cannot be generalized, we introduce new arguments that do not only apply to nice u-graphs, but will be relevant for the broader class of strictly 1-balanced u-graphs. Further, we show how the remainder of the process coupling for the $K_r$-case can be utilized in a black-box manner for any u-graph. These advances pave the way for future generalizations.
Canonical labelling of sparse random graphs
We show that if $p=O(1/n)$, then the Erdős-Rényi random graph $G(n,p)$ with high probability admits a canonical labeling computable in time $O(n\log n)$. Combined with the previous results on the canonization of random graphs, this implies that $G(n,p)$ with high probability admits a polynomial-time canonical labeling whatever the edge probability function $p$. Our algorithm combines the standard color refinement routine with simple post-processing based on the classical linear-time tree canonization. Noteworthy, our analysis of how well color refinement performs in this setting allows us to complete the description of the automorphism group of the 2-core of $G(n,p)$.
Concentration of information on discrete groups
Published
• View Publication
• BIB
Motivated by the Asymptotic Equipartition Property and its recently discovered role in the cutoff phenomenon, we initiate the systematic study of varentropy on discrete groups. Our main result is an approximate tensorization inequality which asserts that the varentropy of any conjugacy-invariant random walk is, up to a universal multiplicative constant, at most that of the free Abelian random walk with the same jump rates. In particular, it is always bounded by the number d of generators, uniformly in time and in the size of the group. This universal estimate is sharp and can be seen as a discrete analogue of a celebrated result of Bobkov and Madiman concerning random d-dimensional vectors with a log-concave density (AOP 2011). A key ingredient in our proof is the fact that conjugacy-invariant random walks have non-negative Bakry-Émery curvature, a result which seems new and of independent interest.