uniform distribution
178 papers tagged with this keyword
Asymptotic Properties of Random Restricted Partitions
Published
• View Publication
• BIB
We study two types of probability measures on the set of integer partitions of $n$ with at most $m$ parts. The first one chooses the random partition with a chance related to its largest part only. We then obtain the limiting distributions of all of the parts together and that of the largest part as $n$ tends to infinity while $m$ is fixed or tends to infinity. In particular, if $m$ goes to infinity not too fast, the largest part satisfies the central limit theorem. The second measure is very general. It includes the Dirichlet distribution and the uniform distribution as special cases. We derive the asymptotic distributions of the parts jointly by taking limits of $n$ and $m$ in the same manner as that in the first probability measure.
The combinatorics of the colliding bullets problem
Published
• View Publication
• BIB
The finite colliding bullets problem is the following simple problem: consider a gun, whose barrel remains in a fixed direction; let $(V_i)_{1\le i\le n}$ be an i.i.d.\ family of random variables with uniform distribution on $[0,1]$; shoot $n$ bullets one after another at times $1,2,\dots, n$, where the $i$th bullet has speed $V_i$. When two bullets collide, they both annihilate. We give the distribution of the number of surviving bullets, and in some generalisation of this model. While the distribution is relatively simple (and we found a number of bold claims online), our proof is surprisingly intricate and mixes combinatorial and geometric arguments; we argue that any rigorous argument must very likely be rather elaborate.
Bernoulli Correlations and Cut Polytopes
Published
• View Publication
• BIB
Given $n$ symmetric Bernoulli variables, what can be said about their correlation matrix viewed as a vector? We show that the set of those vectors $R(\mathcal{B}_n)$ is a polytope and identify its vertices. Those extreme points correspond to correlation vectors associated to the discrete uniform distributions on diagonals of the cube $[0,1]^n$. We also show that the polytope is affinely isomorphic to a well-known cut polytope ${\rm CUT}(n)$ which is defined as a convex hull of the cut vectors in a complete graph with vertex set $\{1,\ldots,n\}$. The isomorphism is obtained explicitly as $R(\mathcal{B}_n)= {\mathbf{1}}-2~{\rm CUT}(n)$. As a corollary of this work, it is straightforward using linear programming to determine if a particular correlation matrix is realizable or not. Furthermore, a sampling method for multivariate symmetric Bernoullis with given correlation is obtained. In some cases the method can also be used for general, not exclusively Bernoulli, marginals.
On the Maximum Size of Block Codes Subject to a Distance Criterion
Published
• View Publication
• BIB
We establish a general formula for the maximum size of finite length block codes with minimum pairwise distance no less than $d$. The achievability argument involves an iterative construction of a set of radius-$d$ balls, each centered at a codeword. We demonstrate that the number of such balls that cover the entire code alphabet cannot exceed this maximum size. Our approach can be applied to codes $i)$ with elements over arbitrary code alphabets, and $ii)$ under a broad class of distance measures, thereby ensuring the generality of our formula. Our formula indicates that the maximum code size can be fully characterized by the cumulative distribution function of the distance measure evaluated at two independent and identically distributed random codewords. When the two random codewords assume a uniform distribution over the entire code alphabet, our formula recovers and obtains a natural generalization of the Gilbert-Varshamov (GV) lower bound. We also establish a general formula for the zero-error capacity of any sequence of channels. Finally, we extend our study to the asymptotic setting, where we establish first- and second-order bounds on the asymptotic code rate subject to a normalized minimum distance constraint.
Probabilistic and Geometrical Applications to Graph Theory
This paper consists of two halves.
In the first half of the paper, we consider real-valued functions $f$ whose domain is the vertex set of a graph $G$ and that are Lipschitz with respect to the graph distance. By placing a uniform distribution on the vertex set, we treat $f$ as a random variable. We investigate the link between the isoperimetric function of $G$ and the functions $f$ that have maximum variance or meet the bound established by the subgaussian inequality. We present several results describing the extremal functions, and use those results to resolve: (A) a conjecture by Bobkov, Houdré, and Tetali characterizing the extremal functions of the subgaussian inequality of the odd cycle, and (B) a conjecture by Alon, Boppana, and Spencer on the relationship between maximum variance functions and the isoperimetric function of product graphs.
While establishing a discrete analogue of the curved Brunn-Minkowski inequality for the discrete hypercube, Ollivier and Villani suggested several avenues for research. We resolve them in second half of the paper as follows.
(1) They propose that a bound on $t$-midpoints can be obtained by repeated application of the bound on midpoints, if the original sets are convex. We construct a specific example where this reasoning fails, and then prove our construction is general by characterizing the convex sets in the discrete hypercube.
(2) A second proposed technique to bound $t$-midpoints involves new results in concentration of measure. We follow through on this proposal, with heavy use on results from the first half of the paper.
(3) We show that the curvature of the discrete hypercube is not positive or zero.
A generalization of Kátai's orthogonality criterion with applications
Published in Discrete and Continuous Dynamical Systems, Volume 39 (2019), Number 5, pp. 2581-2612
• View Publication
• BIB
We study properties of arithmetic sets coming from multiplicative number theory and obtain applications in the theory of uniform distribution and ergodic theory. Our main theorem is a generalization of Kátai's orthogonality criterion. Here is a special case of this theorem:
Let $a\colon\mathbb{N}\to\mathbb{C}$ be a bounded sequence satisfying $$ \sum_{n\leq x} a(pn)\overline{a(qn)} = {\rm o}(x),~\text{for all distinct primes $p$ and $q$.} $$ Then for any multiplicative function $f$ and any $z\in\mathbb{C}$ the indicator function of the level set $E=\{n\in\mathbb{N}:f(n)=z\}$ satisfies $$ \sum_{n\leq x} \mathbb{1}_E(n)a(n)={\rm o}(x). $$
With the help of this theorem one can show that if $E=\{n_1<n_2<\ldots\}$ is a level set of a multiplicative function having positive upper density, then for a large class of sufficiently smooth functions $h\colon(0,\infty)\to\mathbb{R}$ the sequence $(h(n_j))_{j\in\mathbb{N}}$ is uniformly distributed $\bmod~1$. This class of functions $h(t)$ includes: all polynomials $p(t)=a_kt^k+\ldots+a_1t+a_0$ such that at least one of the coefficients $a_1,a_2,\ldots,a_k$ is irrational, $t^c$ for any $c>0$ with $c\notin \mathbb{N}$, $\log^r(t)$ for any $r>2$, $\log(Γ(t))$, $t\log(t)$, and $\frac{t}{\log t}$. The uniform distribution results, in turn, allow us to obtain new examples of ergodic sequences, i.e. sequences along which the ergodic theorem holds.
The flip Markov chain for connected regular graphs
Published
• View Publication
• BIB
Mahlmann and Schindelhauer (2005) defined a Markov chain which they called $k$-Flipper, and showed that it is irreducible on the set of all connected regular graphs of a given degree (at least 3). We study the 1-Flipper chain, which we call the flip chain, and prove that the flip chain converges rapidly to the uniform distribution over connected $2r$-regular graphs with $n$ vertices, where $n\geq 8$ and $r = r(n)\geq 2$. Formally, we prove that the distribution of the flip chain will be within $\varepsilon$ of uniform in total variation distance after $\text{poly}(n,r,\log(\varepsilon^{-1}))$ steps. This polynomial upper bound on the mixing time is given explicitly, and improves markedly on a previous bound given by Feder et al.(2006). We achieve this improvement by using a direct two-stage canonical path construction, which we define in a general setting.
This work has applications to decentralised networks based on random regular connected graphs of even degree, as a self-stabilising protocol in which nodes spontaneously perform random flips in order to repair the network.
Tournament limits: Degree distributions, score functions and self-converseness
Motivated by known results for finite tournaments, we define and study the score functions of tournament kernels and the degree distributions of tournament limits. Our main theorem completely characterises those distributions that appear as the degree distribution of some tournament limit and those functions that appear as the score function of some tournament kernel. We also show that only the uniform distribution can be realised as the outdegree distribution of a unique tournament limit. Finally we define self-converse tournament limits and kernels and characterise their degree distributions and score functions.
The probability of avoiding consecutive patterns in the Mallows distribution
Published
• View Publication
• BIB
We use various combinatorial and probabilistic techniques to study growth rates for the probability that a random permutation from the Mallows distribution avoids consecutive patterns. The Mallows distribution behaves like a $q$-analogue of the uniform distribution by weighting each permutation $π$ by $q^{inv(π)}$, where $inv(π)$ is the number of inversions in $π$ and $q$ is a positive, real-valued parameter. We prove that the growth rate exists for all patterns and all $q>0$, and we generalize Goulden and Jackson's cluster method to keep track of the number of inversions in permutations avoiding a given consecutive pattern. Using singularity analysis, we approximate the growth rates for length-3 patterns, monotone patterns, and non-overlapping patterns starting with 1, and we compare growth rates between different patterns. We also use Stein's method to show that, under certain assumptions on $q$, the length of $σ$, and $inv(σ)$, the number of occurrences of a given pattern $σ$ is well approximated by the normal distribution.
Longest monotone subsequences and rare regions of pattern-avoiding permutations
Published in Electronic Journal of Combinatorics, Volume 24 (2017), Issue 4, Paper #P4.13
• View Publication
• BIB
We consider the distributions of the lengths of the longest monotone and alternating subsequences in classes of permutations of size $n$ that avoid a specific pattern or set of patterns, with respect to the uniform distribution on each such class. We obtain exact results for any class that avoids two patterns of length 3, as well as results for some classes that avoid one pattern of length 4 or more. In our results, the longest monotone subsequences have expected length proportional to $n$ for pattern-avoiding classes, in contrast with the $\sqrt n$ behaviour that holds for unrestricted permutations.
In addition, for a pattern $τ$ of length $k$, we scale the plot of a random $τ$-avoiding permutation down to the unit square and study the "rare region," which is the part of the square that is exponentially unlikely to contain any points. We prove that when $τ_1>τ_k$, the complement of the rare region is a closed set that contains the main diagonal of the unit square. For the case $τ_1=k,$ we also show that the lower boundary of the part of the rare region above the main diagonal is a curve that is Lipschitz continuous and strictly increasing on $[0,1]$.
Uniform measures on braid monoids and dual braid monoids
Published in Journal of Algebra, Elsevier, Volume 473, March 2017, pages 627-666
• View Publication
• BIB
We aim at studying the asymptotic properties of typical positive braids, respectively positive dual braids. Denoting by $μ_k$ the uniform distribution on positive (dual) braids of length $k$, we prove that the sequence $(μ_k)_k$ converges to a unique probability measure $μ_{\infty}$ on infinite positive (dual) braids. The key point is that the limiting measure $μ_{\infty}$ has a Markovian structure which can be described explicitly using the combinatorial properties of braids encapsulated in the Möbius polynomial. As a by-product, we settle a conjecture by Gebhardt and Tawn (J. Algebra, 2014) on the shape of the Garside normal form of large uniform braids.
The distribution of minimum-weight cliques and other subgraphs in graphs with random edge weights
Published
• View Publication
• BIB
We determine, asymptotically in $n$, the distribution and mean of the weight of a minimum-weight $k$-clique (or any strictly balanced graph $H$) in a complete graph $K_n$ whose edge weights are independent random values drawn from the uniform distribution or other continuous distributions. For the clique, we also provide explicit (non-asymptotic) bounds on the distribution's CDF in a form obtained directly from the Stein-Chen method, and in a looser but simpler form. The direct form extends to other subgraphs and other edge-weight distributions. We illustrate the clique results for various values of $k$ and $n$. The results may be applied to evaluate whether an observed minimum-weight copy of a graph $H$ in a network provides statistical evidence that the network's edge weights are not independently distributed but have some structure.
Walking on the Edge and Cosystolic Expansion
Random walks on regular bounded degree expander graphs have numerous applications. A key property of these walks is that they converge rapidly to the uniform distribution on the vertices. The recent study of expansion of high dimensional simplicial complexes, which are the high dimensional analogues of graphs, calls for the natural generalization of random walks to higher dimensions. In particular, a high order random walk on a $2$-dimensional simplicial complex moves at random between neighboring edges of the complex, where two edges are considered neighbors if they share a common triangle. We show that if a regular $2$-dimensional simplicial complex is a cosystolic expander and the underlying graph of the complex has a spectral gap larger than $1/2$, then the random walk on the edges of the complex converges rapidly to the uniform distribution on the edges.
Poset edge densities, nearly reduced words, and barely set-valued tableaux
Published in J. Combin. Theory Ser. A 158 (2018), 66-125
• View Publication
• BIB
In certain finite posets, the expected down-degree of their elements is the same whether computed with respect to either the uniform distribution or the distribution weighting an element by the number of maximal chains passing through it. We show that this coincidence of expectations holds for Cartesian products of chains, connected minuscule posets, weak Bruhat orders on finite Coxeter groups, certain lower intervals in Young's lattice, and certain lower intervals in the weak Bruhat order below dominant permutations. Our tools involve formulas for counting nearly reduced factorizations in 0-Hecke algebras; that is, factorizations that are one letter longer than the Coxeter group length.
Size biased couplings and the spectral gap for random regular graphs
Published in Ann. Probab., 46(1):72-125, 2018
• View Publication
• BIB
Let $λ$ be the second largest eigenvalue in absolute value of a uniform random $d$-regular graph on $n$ vertices. It was famously conjectured by Alon and proved by Friedman that if $d$ is fixed independent of $n$, then $λ=2\sqrt{d-1} +o(1)$ with high probability. In the present work we show that $λ=O(\sqrt{d})$ continues to hold with high probability as long as $d=O(n^{2/3})$, making progress towards a conjecture of Vu that the bound holds for all $1\le d\le n/2$. Prior to this work the best result was obtained by Broder, Frieze, Suen and Upfal (1999) using the configuration model, which hits a barrier at $d=o(n^{1/2})$. We are able to go beyond this barrier by proving concentration of measure results directly for the uniform distribution on $d$-regular graphs. These come as consequences of advances we make in the theory of concentration by size biased couplings. Specifically, we obtain Bennett-type tail estimates for random variables admitting certain unbounded size biased couplings.
Symmetric Graphs with respect to Graph Entropy
Published
• View Publication
• BIB
Let $F_G(P)$ be a functional defined on the set of all the probability distributions on the vertex set of a graph $G$. We say that $G$ is \emph{symmetric with respect to $F_G(P)$} if the uniform distribution on $V(G)$ maximizes $F_G(P)$. Using the combinatorial definition of the entropy of a graph in terms of its vertex packing polytope and the relationship between the graph entropy and fractional chromatic number, we characterize all graphs which are symmetric with respect to graph entropy. We show that a graph is symmetric with respect to graph entropy if and only if its vertex set can be uniformly covered by its maximum size independent sets. Furthermore, given any strictly positive probability distribution $P$ on the vertex set of a graph $G$, we show that $P$ is a maximizer of the entropy of graph $G$ if and only if its vertex set can be uniformly covered by its maximum weighted independent sets. We also show that the problem of deciding if a graph is symmetric with respect to graph entropy, where the weight of the vertices is given by probability distribution $P$, is co-NP-hard.
Random Interval Graphs
Published in J. Discrete Math. Sci. Cryptography 20 (8): 1697-1720, 2017
• View Publication
• BIB
In this thesis, which is supervised by Dr. David Penman, we examine random interval graphs. Recall that such a graph is defined by letting $X_{1},\ldots X_{n},Y_{1},\ldots Y_{n}$ be $2n$ independent random variables, with uniform distribution on $[0,1]$. We then say that the $i$th of the $n$ vertices is the interval $[X_{i},Y_{i}]$ if $X_{i}<Y_{i}$ and the interval $[Y_{i},X_{i}]$ if $Y_{i}<X_{i}$. We then say that two vertices are adjacent if and only if the corresponding intervals intersect.
We recall from our MA902 essay that fact that in such a graph, each edge arises with probability $2/3$, and use this fact to obtain estimates of the number of edges. Next, we turn to how these edges are spread out, seeing that (for example) the range of degrees for the vertices is much larger than classically, by use of an interesting geometrical lemma. We further investigate the maximum degree, showing it is always very close to the maximum possible value $(n-1)$, and the striking result that it is equal to $(n-1)$ with probability exactly $2/3$. We also recall a result on the minimum degree, and contrast all these results with the much narrower range of values obtained in the alternative \lq comparable\rq\, model $G(n,2/3)$ (defined later).
We then study clique numbers, chromatic numbers and independence numbers in the Random Interval Graphs, presenting (for example) a result on independence numbers which is proved by considering the largest chain in the associated interval order.
Last, we make some brief remarks about other ways to define random interval graphs, and extensions of random interval graphs, including random dot product graphs and other ways to define random interval graphs. We also discuss some areas these ideas should be usable in. We close with a summary and some comments.
Cutoff on all Ramanujan graphs
Published
• View Publication
• BIB
We show that on every Ramanujan graph $G$, the simple random walk exhibits cutoff: when $G$ has $n$ vertices and degree $d$, the total-variation distance of the walk from the uniform distribution at time $t=\frac{d}{d-2}\log_{d-1} n + s\sqrt{\log n}$ is asymptotically $\mathbb{P}(Z > c\, s)$ where $Z$ is a standard normal variable and $c=c(d)$ is an explicit constant. Furthermore, for all $1 \leq p \leq \infty$, $d$-regular Ramanujan graphs minimize the asymptotic $L^p$-mixing time for SRW among all $d$-regular graphs. Our proof also shows that, for every vertex $x$ in $G$ as above, its distance from $n-o(n)$ of the vertices is asymptotically $\log_{d-1} n$.
Generic properties of subgroups of free groups and finite presentations
Published in Contemporary Mathematics 677 (2016) 1-44
• View Publication
• BIB
Asymptotic properties of finitely generated subgroups of free groups, and of finite group presentations, can be considered in several fashions, depending on the way these objects are represented and on the distribution assumed on these representations: here we assume that they are represented by tuples of reduced words (generators of a subgroup) or of cyclically reduced words (relators). Classical models consider fixed size tuples of words (e.g. the few-generator model) or exponential size tuples (e.g. Gromov's density model), and they usually consider that equal length words are equally likely. We generalize both the few-generator and the density models with probabilistic schemes that also allow variability in the size of tuples and non-uniform distributions on words of a given length.Our first results rely on a relatively mild prefix-heaviness hypothesis on the distributions, which states essentially that the probability of a word decreases exponentially fast as its length grows. Under this hypothesis, we generalize several classical results: exponentially generically a randomly chosen tuple is a basis of the subgroup it generates, this subgroup is malnormal and the tuple satisfies a small cancellation property, even for exponential size tuples. In the special case of the uniform distribution on words of a given length, we give a phase transition theorem for the central tree property, a combinatorial property closely linked to the fact that a tuple freely generates a subgroup. We then further refine our results when the distribution is specified by a Markovian scheme, and in particular we give a phase transition theorem which generalizes the classical results on the densities up to which a tuple of cyclically reduced words chosen uniformly at random exponentially generically satisfies a small cancellation property, and beyond which it presents a trivial group.
Bias vs structure of polynomials in large fields, and applications in information theory
Published
• View Publication
• BIB
Let $f$ be a polynomial of degree $d$ in $n$ variables over a finite field $\mathbb{F}$. The polynomial is said to be unbiased if the distribution of $f(x)$ for a uniform input $x \in \mathbb{F}^n$ is close to the uniform distribution over $\mathbb{F}$, and is called biased otherwise. The polynomial is said to have low rank if it can be expressed as a composition of a few lower degree polynomials. Green and Tao [Contrib. Discrete Math 2009] and Kaufman and Lovett [FOCS 2008] showed that bias implies low rank for fixed degree polynomials over fixed prime fields. This lies at the heart of many tools in higher order Fourier analysis. In this work, we extend this result to all prime fields (of size possibly growing with $n$). We also provide a generalization to nonprime fields in the large characteristic case. However, we state all our applications in the prime field setting for the sake of simplicity of presentation.
Using the above generalization to large fields as a starting point, we are also able to settle the list decoding radius of fixed degree Reed-Muller codes over growing fields. The case of fixed size fields was solved by Bhowmick and Lovett [STOC 2015], which resolved a conjecture of Gopalan-Klivans-Zuckerman [STOC 2008]. Here, we show that the list decoding radius is equal the minimum distance of the code for all fixed degrees, even when the field size is possibly growing with $n$.
Additionally, we effectively resolve the weight distribution problem for Reed-Muller codes of fixed degree over all fields, first raised in 1977 in the classic textbook by MacWilliams and Sloane [Research Problem 15.1 in Theory of Error Correcting Codes].