arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

7100 papers tagged with this keyword
2020-10-21
Real roots near the unit circle of random polynomials
Published • View Publication • BIB
Let $f_n(z) = \sum_{k = 0}^n \varepsilon_k z^k$ be a random polynomial where $\varepsilon_0,\ldots,\varepsilon_n$ are i.i.d. random variables with $\mathbb{E} \varepsilon_1 = 0$ and $\mathbb{E} \varepsilon_1^2 = 1$. Letting $r_1, r_2,\ldots, r_k$ denote the real roots of $f_n$, we show that the point process defined by $\{|r_1| - 1,\ldots, |r_k| - 1 \}$ converges to a non-Poissonian limit on the scale of $n^{-1}$ as $n \to \infty$. Further, we show that for each $δ> 0$, $f_n$ has a real root within $Θ_δ(1/n)$ of the unit circle with probability at least $1 - δ$. This resolves a conjecture of Shepp and Vanderbei from 1995 by confirming its weakest form and refuting its strongest form.
2020-10-21 v2
On the robustness of the metric dimension of grid graphs to adding a single edge
Published • View Publication • BIB
The metric dimension (MD) of a graph is a combinatorial notion capturing the minimum number of landmark nodes needed to distinguish every pair of nodes in the graph based on graph distance. We study how much the MD can increase if we add a single edge to the graph. The extra edge can either be selected adversarially, in which case we are interested in the largest possible value that the MD can take, or uniformly at random, in which case we are interested in the distribution of the MD. The adversarial setting has already been studied by [Eroh et. al., 2015] for general graphs, who found an example where the MD doubles on adding a single edge. By constructing a different example, we show that this increase can be as large as exponential. However, we believe that such a large increase can occur only in specially constructed graphs, and that in most interesting graph families, the MD at most doubles on adding a single edge. We prove this for $d$-dimensional grid graphs, by showing that $2d$ appropriately chosen corners and the endpoints of the extra edge can distinguish every pair of nodes, no matter where the edge is added. For the special case of $d=2$, we show that it suffices to choose the four corners as landmarks. Finally, when the extra edge is sampled uniformly at random, we conjecture that the MD of 2-dimensional grids converges in probability to $3+\mathrm{Ber}(8/27)$, and we give an almost complete proof.
2020-10-20
Area Statistics for Large Oscillating Tableaux
In this note we show that the area of the partitions making up an oscillating tableaux is described by a random walk on the first quadrant of $\mathbb{Z}^2$ with certain position dependent weights. We are able to recursively calculate the moments of the walk. As the length of the oscillating tableaux becomes large we show that this random walk converges to a Gaussian stochastic process.
2020-10-20 v2
Sparse reconstruction in spin systems I: iid spins
Published • View Publication • BIB
For a sequence of Boolean functions $f_n : \{-1,1\}^{V_n} \longrightarrow \{-1,1\}$, defined on increasing configuration spaces of random inputs, we say that there is sparse reconstruction if there is a sequence of subsets $U_n \subseteq V_n$ of the coordinates satisfying $|U_n| = o(|V_n|)$ such that knowing the coordinates in $U_n$ gives us a non-vanishing amount of information about the value of $f_n$. We first show that, if the underlying measure is a product measure, then no sparse reconstruction is possible for any sequence of transitive functions. We discuss the question in different frameworks, measuring information content in $L^2$ and with entropy. We also highlight some interesting connections with cooperative game theory. Beyond transitive functions, we show that the left-right crossing event for critical planar percolation on the square lattice does not admit sparse reconstruction either. Some of these results answer questions posed by Itai Benjamini.
2020-10-19 v2
independence: Fast Rank Tests
In 1948 Hoeffding devised a nonparametric test that detects dependence between two continuous random variables X and Y, based on the ranking of n paired samples (Xi,Yi). The computation of this commonly-used test statistic takes O(n log n) time. Hoeffding's test is consistent against any dependent probability density f(x,y), but can be fooled by other bivariate distributions with continuous margins. Variants of this test with full consistency have been considered by Blum, Kiefer, and Rosenblatt (1961), Yanagimoto (1970), Bergsma and Dassios (2010). The so far best known algorithms to compute these stronger independence tests have required quadratic time. Here we improve their run time to O(n log n), by elaborating on new methods for counting ranking patterns, from a recent paper by the author and Leng (SODA'21). Therefore, in all circumstances under which the classical Hoeffding independence test is applicable, we provide novel competitive algorithms for consistent testing against all alternatives. Our R package, independence, offers a highly optimized implementation of these rank-based tests. We demonstrate its capabilities on large-scale datasets.
2020-10-18 v2
On the permanent of a random symmetric matrix
Published • View Publication • BIB
Let $M_{n}$ denote a random symmetric $n\times n$ matrix, whose entries on and above the diagonal are i.i.d. Rademacher random variables (taking values $\pm 1$ with probability $1/2$ each). Resolving a conjecture of Vu, we prove that the permanent of $M_{n}$ has magnitude $n^{n/2+o(n)}$ with probability $1-o(1)$. Our result can also be extended to more general models of random matrices.
Revisiting Shao and Sokal's $B_2$ index of phylogenetic balance
Published in Journal of Mathematical Biology 83:52 (2021) • View Publication • BIB
Measures of phylogenetic balance, such as the Colless and Sackin indices, play an important role in phylogenetics. Unfortunately, these indices are specifically designed for phylogenetic trees, and do not extend naturally to phylogenetic networks (which are increasingly used to describe reticulate evolution). This led us to consider a lesser-known balance index, whose definition is based on a probabilistic interpretation that is equally applicable to trees and to networks. This index, known as the $B_2$ index, was first proposed by Shao and Sokal in 1990. Surprisingly, it does not seem to have been studied mathematically since. Likewise, it is used only sporadically in the biological literature, where it tends to be viewed as arcane. In this paper, we study mathematical properties of $B_2$ such as its expectation and variance under the most common models of random trees and its extremal values over various classes of phylogenetic networks. We also assess its relevance in biological applications, and find it to be comparable to that of the Colless and Sackin indices. Altogether, our results call for a reevaluation of the status of this somewhat forgotten measure of phylogenetic balance.
2020-10-16
Central Limit Theorem for Majority Dynamics: Bribing Three Voters Suffices
Published • View Publication • BIB
Given a graph $G$ and some initial labelling $σ: V(G) \to \{Red, Blue\}$ of its vertices, the \textit{majority dynamics model} is the deterministic process where at each stage, every vertex simultaneously replaces its label with the majority label among its neighbors (remaining unchanged in the case of a tie). We prove---for a wide range of parameters---that if an initial assignment is fixed and we independently sample an Erdős--Rényi random graph, $G_{n,p}$, then after one step of majority dynamics, the number of vertices of each label follows a central limit law. As a corollary, we provide a strengthening of a theorem of Benjamini, Chan, O'Donnell, Tamuz, and Tan about the number of steps required for the process to reach unanimity when the initial assignment is also chosen randomly. Moreover, suppose there are initially three more red vertices than blue. In this setting, we prove that if we independently sample the graph $G_{n,1/2}$, then with probability at least $51\%$, the majority dynamics process will converge to every vertex being red. This improves a result of Tran and Vu who addressed the case that the initial lead is at least 10.
2020-10-16
The threshold for the square of a Hamilton cycle
Published • View Publication • BIB
Resolving a conjecture of Kühn and Osthus from 2012, we show that $p= 1/\sqrt{n}$ is the threshold for the random graph $G_{n,p}$ to contain the square of a Hamilton cycle.
2020-10-15
On sparse random combinatorial matrices
Published • View Publication • BIB
Let $Q_{n,d}$ denote the random combinatorial matrix whose rows are independent of one another and such that each row is sampled uniformly at random from the subset of vectors in $\{0,1\}^n$ having precisely $d$ entries equal to $1$. We present a short proof of the fact that $\Pr[\det(Q_{n,d})=0] = O\left(\frac{n^{1/2}\log^{3/2} n}{d}\right)=o(1)$, whenever $d=ω(n^{1/2}\log^{3/2} n)$. In particular, our proof accommodates sparse random combinatorial matrices in the sense that $d = o(n)$ is allowed. We also consider the singularity of deterministic integer matrices $A$ randomly perturbed by a sparse combinatorial matrix. In particular, we prove that $\Pr[\det(A+Q_{n,d})=0]=O\left(\frac{n^{1/2}\log^{3/2} n}{d}\right)$, again, whenever $d=ω(n^{1/2}\log^{3/2} n)$ and $A$ has the property that $(1,-d)$ is not an eigenpair of $A$.
2020-10-14 v3
Minimum stationary values of sparse random directed graphs
Published • View Publication • BIB
We consider the stationary distribution of the simple random walk on the directed configuration model with bounded degrees. Provided that the minimum out-degree is at least $2$, with high probability (whp) there is a unique stationary distribution. We show that the minimum positive stationary value is whp $n^{-(1+C+o(1))}$ for some constant $C \ge 0$ determined by the degree distribution. In particular, $C$ is the competing combination of two factors: (1) the contribution of atypically "thin" in-neighbourhoods, controlled by subcritical branching processes; and (2) the contribution of atypically "light" trajectories, controlled by large deviation rate functions. Additionally, our proof implies that whp the hitting and the cover time are both $n^{1+C+o(1)}$. Our results complement those of Caputo and Quattropani who showed that if the minimum in-degree is at least 2, stationary values have logarithmic fluctuations around $n^{-1}$.
UV mission planning under uncertainty in vehicles' availability
Published • View Publication • BIB
Heterogeneous unmanned vehicles (UVs) are used in various defense and civil applications. Some of the civil applications of UVs for gathering data and monitoring include civil infrastructure management, agriculture, public safety, law enforcement, disaster relief, and transportation. This paper presents a two-stage stochastic model for a fuel-constrained UV mission planning problem with multiple refueling stations under uncertainty in availability of UVs. Given a set of points of interests (POI), a set of refueling stations for UVs, and a base station where the UVs are stationed and their availability is random, the objective is to determine route for each UV starting and terminating at the base station such that overall incentives collected by visiting POIs is maximized. We present an outer approximation based decomposition algorithm to solve large instances, and perform extensive computational experiments using random instances. Additionally, a data driven simulation study is performed using robot operating system (ROS) framework to corroborate the use of the stochastic programming approach.
2020-10-13 v2
Sharp invertibility of random Bernoulli matrices
Let $p \in (0,1/2)$ be fixed, and let $B_n(p)$ be an $n\times n$ random matrix with i.i.d. Bernoulli random variables with mean $p$. We show that for all $t \ge 0$, \[\mathbb{P}[s_n(B_n(p)) \le tn^{-1/2}] \le C_p t + 2n(1-p)^{n} + C_p (1-p-ε_p)^{n},\] where $s_n(B_n(p))$ denotes the least singular value of $B_n(p)$ and $C_p, ε_p > 0$ are constants depending only on $p$. In particular, \[\mathbb{P}[B_{n}(p) \text{ is singular}] = 2n(1-p)^{n} + C_{p}(1-p-ε_p)^{n},\] which confirms a conjecture of Litvak and Tikhomirov. We also confirm a conjecture of Nguyen by showing that if $Q_{n}$ is an $n\times n$ random matrix with independent rows that are uniformly distributed on the central slice of $\{0,1\}^{n}$, then \[\mathbb{P}[Q_{n} \text{ is singular}] = (1/2 + o_n(1))^{n}.\] This provides, for the first time, a sharp determination of the logarithm of the probability of singularity in any natural model of random discrete matrices with dependent entries.
2020-10-13 v2
Singularity of discrete random matrices
Published • View Publication • BIB
Let $ξ$ be a non-constant real-valued random variable with finite support, and let $M_{n}(ξ)$ denote an $n\times n$ random matrix with entries that are independent copies of $ξ$. For $ξ$ which is not uniform on its support, we show that \begin{align*} \mathbb{P}[M_{n}(ξ)\text{ is singular}] &= \mathbb{P}[\text{zero row or column}] + (1+o_n(1))\mathbb{P}[\text{two equal (up to sign) rows or columns}], \end{align*} thereby confirming a folklore conjecture. As special cases, we obtain: (1) For $ξ= \text{Bernoulli}(p)$ with fixed $p \in (0,1/2)$, \[\mathbb{P}[M_{n}(ξ)\text{ is singular}] = 2n(1-p)^{n} + (1+o_n(1))n(n-1)(p^2 + (1-p)^2)^{n},\] which determines the singularity probability to two asymptotic terms. Previously, no result of such precision was available in the study of the singularity of random matrices. (2) For $ξ= \text{Bernoulli}(p)$ with fixed $p \in (1/2,1)$, \[\mathbb{P}[M_{n}(ξ)\text{ is singular}] = (1+o_n(1))n(n-1)(p^2 + (1-p)^2)^{n}.\] Previously, only the much weaker upper bound of $(\sqrt{p} + o_n(1))^{n}$ was known due to the work of Bourgain-Vu-Wood. For $ξ$ which is uniform on its support: (1) We show that \begin{align*} \mathbb{P}[M_{n}(ξ)\text{ is singular}] &= (1+o_n(1))^{n}\mathbb{P}[\text{two rows or columns are equal}]. \end{align*} (2) Perhaps more importantly, we provide a sharp analysis of the contribution of the `compressible' part of the unit sphere to the lower tail of the smallest singular value of $M_{n}(ξ)$.
2020-10-13 v2
Covariance within Random Integer Compositions
Fix a positive integer $N$. Select an additive composition $ξ$ of $N$ uniformly out of $2^{N-1}$ possibilities. The interplay between the number of parts in $ξ$ and the maximum part in $ξ$ is our focus. It is not surprising that correlations $ρ(N)$ between these quantities are negative; we earlier gave inconclusive evidence that $\lim_{N \to \infty} ρ(N)$ is strictly less than zero. A proof of this result would imply asymptotic dependence. We now retract our presumption in such an unforeseen outcome. Similar experimental findings apply when $ξ$ is a 1-free composition, i.e., possessing only parts $\geq 2$.
Counting Subgraphs in Degenerate Graphs
Published • View Publication • BIB
We consider the problem of counting the number of copies of a fixed graph $H$ within an input graph $G$. This is one of the most well-studied algorithmic graph problems, with many theoretical and practical applications. We focus on solving this problem when the input $G$ has bounded degeneracy. This is a rich family of graphs, containing all graphs without a fixed minor (e.g. planar graphs), as well as graphs generated by various random processes (e.g. preferential attachment graphs). We say that $H$ is easy if there is a linear-time algorithm for counting the number of copies of $H$ in an input $G$ of bounded degeneracy. A seminal result of Chiba and Nishizeki from '85 states that every $H$ on at most 4 vertices is easy. Bera, Pashanasangi, and Seshadhri recently extended this to all $H$ on 5 vertices, and further proved that for every $k > 5$ there is a $k$-vertex $H$ which is not easy. They left open the natural problem of characterizing all easy graphs $H$. Bressan has recently introduced a framework for counting subgraphs in degenerate graphs, from which one can extract a sufficient condition for a graph $H$ to be easy. Here we show that this sufficient condition is also necessary, thus fully answering the Bera--Pashanasangi--Seshadhri problem. We further resolve two closely related problems; namely characterizing the graphs that are easy with respect to counting induced copies, and with respect to counting homomorphisms.
2020-10-12 v2
Cayley graphs with few automorphisms: the case of infinite groups
Published in Annales Henri Lebesgue, Volume 5 (2022), pp. 73-92 • View Publication • BIB
We characterize the finitely generated groups that admit a Cayley graph whose only automorphisms are the translations, confirming a conjecture by Watkins from 1976. The proof relies on random walk techniques. As a consequence, every finitely generated group admits a Cayley graph with countable automorphism group. We also treat the case of directed graphs.
2020-10-09 v2
Effective resistance is more than distance: Laplacians, Simplices and the Schur complement
Published • View Publication • BIB
This article discusses a geometric perspective on the well-known fact in graph theory that the effective resistance is a metric on the nodes of a graph. The classical proofs of this fact make use of ideas from electrical circuits or random walks; here we describe an alternative approach which combines geometric (using simplices) and algebraic (using the Schur complement) ideas. These perspectives are unified in a matrix identity of Miroslav Fiedler, which beautifully summarizes a number of related ideas at the intersection of graphs, Laplacian matrices and simplices, with the metric property of the effective resistance as a prominent consequence.
2020-10-06 v2
Strongly separable matrices for nonadaptive combinatorial group testing
Published • View Publication • BIB
In nonadaptive combinatorial group testing (CGT), it is desirable to identify a small set of up to $d$ defectives from a large population of $n$ items with as few tests (i.e. large rate) and efficient identifying algorithm as possible. In the literature, $d$-disjunct matrices ($d$-DM) and $\bar{d}$-separable matrices ($\bar{d}$-SM) are two classical combinatorial structures having been studied for several decades. It is well-known that a $d$-DM provides a more efficient identifying algorithm than a $\bar{d}$-SM, while a $\bar{d}$-SM could have a larger rate than a $d$-DM. In order to combine the advantages of these two structures, in this paper, we introduce a new notion of \emph{strongly $d$-separable matrix} ($d$-SSM) for nonadaptive CGT and show that a $d$-SSM has the same identifying ability as a $d$-DM, but much weaker requirements than a $d$-DM. Accordingly, the general bounds on the largest rate of a $d$-SSM are established. Moreover, by the random coding method with expurgation, we derive an improved lower bound on the largest rate of a $2$-SSM which is much higher than the best known result of a $2$-DM.
2020-10-05 v2
The Limit Shape of the Leaky Abelian Sandpile Model
Published • View Publication • BIB
The leaky abelian sandpile model (Leaky-ASM) is a growth model in which $n$ grains of sand start at the origin in $\mathbb{Z}^2$ and diffuse along the vertices according to a toppling rule. A site can topple if its amount of sand is above a threshold. In each topple a site sends some sand to each neighbor and leaks a portion $1-1/d$ of its sand. We compute the limit shape as a function of $d$ in the symmetric case where each topple sends an equal amount of sand to each neighbor. The limit shape converges to a circle as $d\to 1$ and a diamond as $d\to\infty$. We compute the limit shape by comparing the odometer function at a site to the probability that a killed random walk dies at that site. When $d\to 1$ the Leaky-ASM converges to the abelian sandpile model (ASM) with a modified initial configuration. We also prove the limit shape is a circle when simultaneously with $n\to\infty$ we have that $d=d_n$ converges to $1$ slower than any power of $n$. To gain information about the ASM faster convergence is necessary.