random
7100 papers tagged with this keyword
Exact nuclear norm, completion and decomposition for random overcomplete tensors via degree-4 SOS
In this paper we show that simple semidefinite programs inspired by degree $4$ SOS can exactly solve the tensor nuclear norm, tensor decomposition, and tensor completion problems on tensors with random asymmetric components. More precisely, for tensor nuclear norm and tensor decomposition, we show that w.h.p. these semidefinite programs can exactly find the nuclear norm and components of an $(n\times n\times n)$-tensor $\mathcal{T}$ with $m\leq n^{3/2}/polylog(n)$ random asymmetric components. Unlike most of the previous algorithms, our algorithm provides a certificate for the decomposition, does not require knowledge about the number of components in the decomposition and does not make any assumptions on the sizes of the coefficients in the decomposition. As a byproduct, we show that w.h.p. the nuclear norm decomposition exactly coincides with the minimum rank decomposition for tensors with $m\leq n^{3/2}/polylog(n)$ random asymmetric components.
For tensor completion, we show that w.h.p. the semidefinite program, introduced by Potechin & Steurer (2017) for tensors with orthogonal components, can exactly recover an $(n\times n\times n)$-tensor $\mathcal{T}$ with $m$ random asymmetric components from only $n^{3/2}m polylog(n)$ randomly observed entries. For non-orthogonal tensors, this improves the dependence on $m$ of the number of entries needed for exact recovery over all previously known algorithms and provides the first theoretical guarantees for exact tensor completion in the overcomplete regime.
Kim--Vu's sandwich conjecture is true for $d \gg \log^4 n$
Kim and Vu made the following conjecture (\textit{Advances in Mathematics}, 2004): if $d\gg \log n$, then the random $d$-regular graph $G(n,d)$ can be ``sandwiched'' between $G(n,p_*)$ and $G(n,p^*)$ where $p_*$ and $p^*$ are both asymptotically equal to $d/n$.
This famous conjecture was previously proved for all $d\gg (n\log n)^{3/4}$.
In this paper, we confirm the conjecture when $d \gg \log^4 n$. We also extend this result to near-regular degree sequences.
Prague dimension of random graphs
Published in Combinatorica 43 (2023), 853-884
• View Publication
• BIB
The Prague dimension of graphs was introduced by Nesetril, Pultr and Rodl in the 1970s. Proving a conjecture of Furedi and Kantor, we show that the Prague dimension of the binomial random graph is typically of order n/log n for constant edge-probabilities. The main new proof ingredient is a Pippenger-Spencer type edge-coloring result for random hypergraphs with large uniformities, i.e., edges of size O(log n).
Super-clustering of consecutive numbers in $p$-shifted random permutations
Let $A^{(n)}_{l;k}\subset S_n$ denote the event that the set of $l$ consecutive numbers $\{k,k+1,\cdots, k+l-1\}$ appear in a set of $l$ consecutive positions. Let $p=\{p_j\}_{j=1}^\infty$ be a distribution on $\mathbb{N}$ with $p_j>0$. Let $P_n$ denote the probability measure on $S_n$ corresponding to the $p$-shifted random permutation. Our main result, under the additional assumption that $\{p_j\}_{j=1}^\infty$ is non-increasing, is that $$ \begin{aligned} &\lim_{l\to\infty}\lim_{n\to\infty}P_n(A^{(n )}_{l,k})=\big(\prod_{j=1}^{k-1}\sum_{i=1}^jp_i\big) \big(\prod_{j=1}^\infty\sum_{i=1}^jp_i\big), \end{aligned} $$ and that if $\lim_{n\to\infty}\min(k_n,n-k_n)=\infty$, then $$ \begin{aligned} &\lim_{l\to\infty}\lim_{n\to\infty}P_n(A^{(n )}_{l,k_n})= \big(\prod_{j=1}^\infty\sum_{i=1}^jp_i\big)^2. \end{aligned} $$ In particular these limits are positive if and only if $\sum_{j=1}^\infty jp_j<\infty$. We say that super-clustering occurs when the limits are positive. We also give a new characterization of the class of $p$-shifted probability distributions on $S_\infty$.
Finding the Second-Best Candidate under the Mallows Model
Published
• View Publication
• BIB
The well-known secretary problem in sequential analysis and optimal stopping theory asks one to maximize the probability of finding the optimal candidate in a sequentially examined list under the constraint that accept/reject decisions are made in real-time. A version of the problem is the so-called postdoc problem, for which the question of interest is to devise a strategy that identifies the second-best candidate with highest possible probability of success.
We study the postdoc problem in its combinatorial form. In this setting, a permutation $π$ of length $N$ is sampled according to some distribution on the symmetric group $S_N$ and the elements of $π$ are revealed one-by-one from left to right so that at each step, one can only observe the relative orders of the elements. At each step, one must decide to either accept or reject the currently presented element and cannot recall the decision in the future. The question of interest is to find the optimal strategy for selecting the position of the second-largest value. We solve the postdoc problem for the untraditional setting where the candidates are not presented uniformly at random but rather according to permutations drawn from the Mallows distribution. The Mallows distribution assigns to each permutation $π\in S_N$ a weight $θ^{c(π)}$, where the function c counts the number of inversions in $π$. To identify the optimal stopping criteria for the significantly more challenging postdoc problem, we adopt a combinatorial methodology that includes new proof techniques and novel methodological extensions compared to the analysis first introduced in the setting of the secretary problem. The optimal strategies depend on the parameter $θ$ of the Mallows distribution and can be determined exactly by solving well-defined recurrence relations.
Muttalib--Borodin plane partitions and the hard edge of random matrix ensembles
We study probabilistic and combinatorial aspects of natural volume-and-trace weighted plane partitions and their continuous analogues. We prove asymptotic limit laws for the largest parts of these ensembles in terms of new and known hard- and soft-edge distributions of random matrix theory. As a corollary we obtain an asymptotic transition between Gumbel and Tracy--Widom GUE fluctuations for the largest part of such plane partitions, with the continuous Bessel kernel providing the interpolation. We interpret our results in terms of two natural models of directed last passage percolation (LPP): a discrete $(\max, +)$ infinite-geometry model with rapidly decaying geometric weights, and a continuous $(\min, \cdot)$ model with power weights.
Multiple Random Walks on Graphs: Mixing Few to Cover Many
Published in Combinatorics, Probability and Computing, 32(4):594 - 637, 2023
• View Publication
• BIB
Random walks on graphs are an essential primitive for many randomised algorithms and stochastic processes. It is natural to ask how much can be gained by running $k$ multiple random walks independently and in parallel. Although the cover time of multiple walks has been investigated for many natural networks, the problem of finding a general characterisation of multiple cover times for worst-case start vertices (posed by Alon, Avin, Koucký, Kozma, Lotker, and Tuttle~in 2008) remains an open problem.
First, we improve and tighten various bounds on the stationary cover time when $k$ random walks start from vertices sampled from the stationary distribution. For example, we prove an unconditional lower bound of $Ω((n/k) \log n)$ on the stationary cover time, holding for any $n$-vertex graph $G$ and any $1 \leq k =o(n\log n )$. Secondly, we establish the stationary cover times of multiple walks on several fundamental networks up to constant factors. Thirdly, we present a framework characterising worst-case cover times in terms of stationary cover times and a novel, relaxed notion of mixing time for multiple walks called the partial mixing time. Roughly speaking, the partial mixing time only requires a specific portion of all random walks to be mixed. Using these new concepts, we can establish (or recover) the worst-case cover times for many networks including expanders, preferential attachment graphs, grids, binary trees and hypercubes.
Triangles in randomly perturbed graphs
Published
• View Publication
• BIB
We study the problem of finding pairwise vertex-disjoint triangles in the randomly perturbed graph model, which is the union of any $n$-vertex graph $G$ satisfying a given minimum degree condition and the binomial random graph $G(n,p)$. We prove that asymptotically almost surely $G \cup G(n,p)$ contains at least $\min\{δ(G), \lfloor n/3 \rfloor\}$ pairwise vertex-disjoint triangles, provided $p \ge C \log n/n$, where $C$ is a large enough constant. This is a perturbed version of an old result of Dirac.
Our result is asymptotically optimal and answers a question of Han, Morris, and Treglown [RSA, 2021, no. 3, 480--516] in a strong form. We also prove a stability version of our result, which in the case of pairwise vertex-disjoint triangles extends a result of Han, Morris, and Treglown [RSA, 2021, no. 3, 480--516]. Together with a result of Balogh, Treglown, and Wagner [CPC, 2019, no. 2, 159--176] this fully resolves the existence of triangle factors in randomly perturbed graphs.
We believe that the methods introduced in this paper are useful for a variety of related problems: we discuss possible generalisations to clique factors, cycle factors, and $2$-universality.
Bounds for the multilevel construction
One of the main problems in random network coding is to compute good lower and upper bounds on the achievable cardinality of the so-called subspace codes in the projective space $\mathcal{P}_q(n)$ for a given minimum distance. The determination of the exact maximum cardinality is a very tough discrete optimization problem involving a huge number of symmetries. Besides some explicit constructions for \textit{good} subspace codes several of the most success full constructions involve the solution of discrete optimization subproblems itself, which mostly have not been not been solved systematically. Here we consider the multilevel a.k.a.\ Echelon--Ferrers construction and given lower and upper bounds for the achievable cardinalities. From a more general point of view, we solve maximum clique problems in weighted graphs, where the weights can be polynomials in the field size $q$.
Rigid structures in the universal enveloping traffic space
For any tracial non-commutative probability space $(\mathcal{A}, \varphi)$, Cébron, Dahlqvist, and Male showed that one can always construct an enveloping traffic space $(\mathcal{G}(\mathcal{A}), τ_\varphi)$ that extends the trace. This construction provides a universal object that allows one to appeal to the traffic probability framework in generic situations, prioritizing an understanding of its structure. In this article, we prove that $(\mathcal{G}(\mathcal{A}), τ_\varphi)$ admits a canonical free product decomposition $\mathcal{A} * \mathcal{A}^\intercal * Θ(\mathcal{G}(\mathcal{A}))$. In particular, $\mathcal{A}^\intercal$ is an anti-isomorphic copy of $\mathcal{A}$, and $Θ(\mathcal{G}(\mathcal{A}))$ is, up to degeneracy, a commutative algebra generated by Gaussian random variables with a covariance structure diagonalized by the graph operations. If $(\mathcal{A}, \varphi)$ itself is a free product, then we describe how this additional structure lifts into $(\mathcal{G}(\mathcal{A}), τ_\varphi)$. Here, we find a connection between free independence and classical independence opposite the usual direction. Up to degeneracy, we further show that $(\mathcal{G}(\mathcal{A}), τ_\varphi)$ is spanned by tree-like graph operations. Finally, we apply our results to the study of large (possibly dependent) random matrices. Our analysis relies on the combinatorics of cactus graphs and the resulting cactus-cumulant correspondence.
Heisenberg XX chain, non-homogeneously parameterised generating exponential, and diagonally restricted plane partitions
The mean values of non-homogeneously parameterized generating exponential are obtained and investigated for the periodic Heisenberg XX model. The norm-trace generating function of boxed plane partitions with fixed volume of their diagonal parts is obtained as N-particles average of the generating exponential. The generating function of self-avoiding walks of random turns vicious walkers is obtained in terms of the circulant matrices that leads to generalizations of the Ramus's identity. Under various specifications of the generating exponential, the N-particles averages arise for a set of inconsecutive flipped spins and for powers of the first moment of flipped spins distribution at large length of the chain. These averages are expressed through the numbers of closed trajectories with constrained initial/final positions. The estimates at large temporal parameter are expressed through the numbers of diagonally restricted plane partitions characterized by fixed values of the main diagonal trace or by fixed heights of the diagonal columns in one-to-one correspondence with the flipped spins positions.
Decompositions of quasirandom hypergraphs into hypergraphs of bounded degree
Published
• View Publication
• BIB
We prove that any quasirandom uniform hypergraph $H$ can be approximately decomposed into any collection of bounded degree hypergraphs with almost as many edges. In fact, our results also apply to multipartite hypergraphs and even to the sparse setting when the density of $H$ quickly tends to $0$ in terms of the number of vertices of $H$. Our results answer and address questions of Kim, Kühn, Osthus and Tyomkyn; and Glock, Kühn and Osthus as well as Keevash.
The provided approximate decompositions exhibit strong quasirandom properties which is very useful for forthcoming applications. Our results also imply approximate solutions to natural hypergraph versions of long-standing graph decomposition problems, as well as several decomposition results for (quasi)random simplicial complexes into various more elementary simplicial complexes such as triangulations of spheres and other manifolds.
Large deviations of the greedy independent set algorithm on sparse random graphs
Published in Random Struct. Algorithms 61, No. 2, 353-363 (2022)
• View Publication
• BIB
We study the greedy independent set algorithm on sparse Erdős-Rényi random graphs ${\mathcal G}(n,c/n)$. This range of $p$ is of interest due to the threshold at $c=e$, beyond which it appears that greedy algorithms are affected by a sudden change in the independent set landscape. A large deviation principle was recently established by Bermolen et al. (2020), however, the proof and rate function are somewhat involved. Upper bounds for the rate function were obtained earlier by Pittel (1982). By discrete calculus, we identify the optimal trajectory realizing a given large deviation and obtain the rate function in a simple closed form. In particular, we show that Pittel's bounds are sharp. The proof is brief and elementary. We think the methods presented here will be useful in analyzing the tail behavior of other random growth and exploration processes.
High Dimensional Expanders: Eigenstripping, Pseudorandomness, and Unique Games
Published
• View Publication
• BIB
Higher order random walks (HD-walks) on high dimensional expanders (HDX) have seen an incredible amount of study and application since their introduction by Kaufman and Mass [KM16], yet their broader combinatorial and spectral properties remain poorly understood. We develop a combinatorial characterization of the spectral structure of HD-walks on two-sided local-spectral expanders [DK17], which offer a broad generalization of the well-studied Johnson and Grassmann graphs. Our characterization, which shows that the spectra of HD-walks lie tightly concentrated in a few combinatorially structured strips, leads to novel structural theorems such as a tight $\ell_2$-characterization of edge-expansion, as well as to a new understanding of local-to-global algorithms on HDX.
Towards the latter, we introduce a spectral complexity measure called Stripped Threshold Rank, and show how it can replace the (much larger) threshold rank in controlling the performance of algorithms on structured objects. Combined with a sum-of-squares proof of the former $\ell_2$-characterization, we give a concrete application of this framework to algorithms for unique games on HD-walks, in many cases improving the state of the art [RBS11, ABS15] from nearly-exponential to polynomial time (e.g. for sparsifications of Johnson graphs or of slices of the $q$-ary hypercube). Our characterization of expansion also holds an interesting connection to hardness of approximation, where an $\ell_\infty$-variant for the Grassmann graphs was recently used to resolve the 2-2 Games Conjecture [KMS18]. We give a reduction from a related $\ell_\infty$-variant to our $\ell_2$-characterization, but it loses factors in the regime of interest for hardness where the gap between $\ell_2$ and $\ell_\infty$ structure is large. Nevertheless, we open the door for further work on the use of HDX in hardness of approximation and unique games.
On the 2-colorability of random hypergraphs
Published in Proc. 6th Intl. Workshop on Randomization and Approximation Techniques in Computer Science (RANDOM '02) 78-90 (2002)
• View Publication
• BIB
A 2-coloring of a hypergraph is a mapping from its vertices to a set of two colors such that no edge is monochromatic. Let $H_k(n,m)$ be a random $k$-uniform hypergraph on $n$ vertices formed by picking $m$ edges uniformly, independently and with replacement. It is easy to show that if $r \geq r_c = 2^{k-1} \ln 2 - (\ln 2) /2$, then with high probability $H_k(n,m=rn)$ is not 2-colorable. We complement this observation by proving that if $r \leq r_c - 1$ then with high probability $H_k(n,m=rn)$ is 2-colorable.
Matchings on trees and the adjacency matrix: A determinantal viewpoint
Let $G$ be a finite tree. For any matching $M$ of $G$, let $U(M)$ be the set of vertices uncovered by $M$. Let $\mathcal{M}_G$ be a uniform random maximum size matching of $G$. In this paper, we analyze the structure of $U(\mathcal{M}_G)$. We first show that $U(\mathcal{M}_G)$ is a determinantal process. We also show that for most vertices of $G$, the process $U(\mathcal{M}_G)$ in a small neighborhood of that vertex can be well approximated based on a somewhat larger neighborhood of the same vertex. Then we show that the normalized Shannon entropy of $U(\mathcal{M}_G)$ can be also well approximated using the local structure of $G$. In other words, in the realm of trees, the normalized Shannon entropy of $U(\mathcal{M}_G)$ -- that is, the normalized logarithm of the number of maximum size matchings of $G$ -- is a Benjamini-Schramm continuous parameter.
We show that $U(\mathcal{M}_G)$ is a determinantal process through establishing a new connection between $U(\mathcal{M}_G)$ and the adjacency matrix of $G$. This result sheds a new light on the well-known fact that on a tree, the number of vertices uncovered by a maximum size matching is equal to the nullity of the adjacency matrix.
Some of the proofs are based on the well established method of introducing a new perturbative parameter, which we call temperature, and then define the positive temperature analogue of $\mathcal{M}_G$, the so called monomer-dimer model, and let the temperature go to zero.
Combinatorial Bernoulli Factories
Published in Bernoulli, 29(2), pp.1246-1274 (2023)
• View Publication
• BIB
A Bernoulli factory is an algorithmic procedure for exact sampling of certain random variables having only Bernoulli access to their parameters. Bernoulli access to a parameter $p \in [0,1]$ means the algorithm does not know $p$, but has sample access to independent draws of a Bernoulli random variable with mean equal to $p$. In this paper, we study the problem of Bernoulli factories for polytopes: given Bernoulli access to a vector $x\in P$ for a given polytope $P\subset [0,1]^n$, output a randomized vertex such that the expected value of the $i$-th coordinate is \emph{exactly} equal to $x_i$. For example, for the special case of the perfect matching polytope, one is given Bernoulli access to the entries of a doubly stochastic matrix $[x_{ij}]$ and asked to sample a matching such that the probability of each edge $(i,j)$ be present in the matching is exactly equal to $x_{ij}$. We show that a polytope $P$ admits a Bernoulli factory if and and only if $P$ is the intersection of $[0,1]^n$ with an affine subspace. Our construction is based on an algebraic formulation of the problem, involving identifying a family of Bernstein polynomials (one per vertex) that satisfy a certain algebraic identity on $P$. The main technical tool behind our construction is a connection between these polynomials and the geometry of zonotope tilings. We apply these results to construct an explicit factory for the perfect matching polytope. The resulting factory is deeply connected to the combinatorial enumeration of arborescences and may be of independent interest. For the $k$-uniform matroid polytope, we recover a sampling procedure known in statistics as Sampford sampling.
Exact Phase Transitions of Model RB with Slower-Growing Domains
The second moment method has always been an effective tool to lower bound the satisfiability threshold of many random constraint satisfaction problems. However, the calculation is usually hard to carry out and as a result, only some loose results can be obtained. In this paper, based on a delicate analysis which fully exploit the power of the second moment method, we prove that random RB instances can exhibit exact phase transition under more relaxed conditions, especially slower-growing domain size. These results are the best by using the second moment method, and new tools should be introduced for any better results.
Singularity of random symmetric matrices revisited
Published
• View Publication
• BIB
Let $M_n$ be drawn uniformly from all $\pm 1$ symmetric $n \times n$ matrices. We show that the probability that $M_n$ is singular is at most $\exp(-c(n\log n)^{1/2})$, which represents a natural barrier in recent approaches to this problem. In addition to improving on the best-known previous bound of Campos, Mattos, Morris and Morrison of $\exp(-c n^{1/2})$ on the singularity probability, our method is different and considerably simpler.
Motif Estimation via Subgraph Sampling: The Fourth Moment Phenomenon
Published in Ann. Statist. 50(2): 987-1011 (April 2022)
• View Publication
• BIB
Network sampling is an indispensable tool for understanding features of large complex networks where it is practically impossible to search over the entire graph. In this paper, we develop a framework for statistical inference for counting network motifs, such as edges, triangles, and wedges, in the widely used subgraph sampling model, where each vertex is sampled independently, and the subgraph induced by the sampled vertices is observed. We derive necessary and sufficient conditions for the consistency and the asymptotic normality of the natural Horvitz-Thompson (HT) estimator, which can be used for constructing confidence intervals and hypothesis testing for the motif counts based on the sampled graph. In particular, we show that the asymptotic normality of the HT estimator exhibits an interesting fourth-moment phenomenon, which asserts that the HT estimator (appropriately centered and rescaled) converges in distribution to the standard normal whenever its fourth-moment converges to 3 (the fourth-moment of the standard normal distribution). As a consequence, we derive the exact thresholds for consistency and asymptotic normality of the HT estimator in various natural graph ensembles, such as sparse graphs with bounded degree, Erdos-Renyi random graphs, random regular graphs, and dense graphons.