uniform distribution
178 papers tagged with this keyword
Limiting spectral laws for sparse random circulant matrices
Fix a positive integer $d$ and let $(G_n)_{n\geq1}$ be a sequence of finite abelian groups with orders tending to infinity. For each $n \geq 1$, let $C_n$ be a uniformly random $G_n$-circulant matrix with entries in $\{0,1\}$ and exactly $d$ ones in each row/column. We show that the empirical spectral distribution of $C_n$ converges weakly in expectation to a probability measure $μ$ on $\mathbb{C}$ if and only if the distribution of the order of a uniform random element of $G_n$ converges weakly to a probability measure $ρ$ on $\mathbb{N}^*$, the one-point compactification of the natural numbers. Furthermore, we show that convergence in expectation can be strengthened to convergence in probability if and only if $ρ$ is a Dirac mass $δ_m$. In this case, $μ$ is the $d$-fold convolution of the uniform distribution on the $m$-th roots of unity if $m\in\mathbb{N}$ or the unit circle if $m = \infty$. We also establish that, under further natural assumptions, the determinant of $C_n$ is $\pm\exp((c_{m,d}+o(1))|G_n|)$ with high probability, where $c_{m,d}$ is a constant depending only on $m$ and $d$.
Partial results for union-closed conjectures on the weighted cube
The celebrated union-closed conjecture is concerned with the cardinalities of various subsets of the Boolean $d$-cube. The cardinality of such a set is equivalent, up to a constant, to its measure under the uniform distribution, so we can pose more general conjectures by choosing a different probability distribution on the cube. In particular, for any sequence of probabilities $(p_i)_{i=1}^d$ we can consider the product of $d$ independent Bernoulli random variables, with success probabilities $p_i$. In this short note, we find a generalised form of Karpas' special case of the union-closed conjecture for families $\mathcal{F}$ with density at least half. We also generalise Knill's logarithmic lower bound.
Random sampling of contingency tables and partitions: Two practical examples of the Burnside process
Published in Stat Comput 35, 181 (2025)
• View Publication
• BIB
This paper gives new, efficient algorithms for approximate uniform sampling of contingency tables and integer partitions. The algorithms use the Burnside process, a general algorithm for sampling a uniform orbit of a finite group acting on a finite set. We show that a technique called `lumping' can be used to derive efficient implementations of the Burnside process. For both contingency tables and partitions, the lumped processes have far lower per step complexity than the original Markov chains. We also define a second Markov chain for partitions called the reflected Burnside process. The reflected Burnside process maintains the computational advantages of the lumped process but empirically converges to the uniform distribution much more rapidly. By using the reflected Burnside process we can easily sample uniform partitions of size $10^{10}$.
CayleyPy RL: Pathfinding and Reinforcement Learning on Cayley Graphs
Published
• View Publication
• BIB
This paper is the second in a series of studies on developing efficient artificial intelligence-based approaches to pathfinding on extremely large graphs (e.g. $10^{70}$ nodes) with a focus on Cayley graphs and mathematical applications. The open-source CayleyPy project is a central component of our research. The present paper proposes a novel combination of a reinforcement learning approach with a more direct diffusion distance approach from the first paper. Our analysis includes benchmarking various choices for the key building blocks of the approach: architectures of the neural network, generators for the random walks and beam search pathfinding. We compared these methods against the classical computer algebra system GAP, demonstrating that they "overcome the GAP" for the considered examples. As a particular mathematical application we examine the Cayley graph of the symmetric group with cyclic shift and transposition generators. We provide strong support for the OEIS-A186783 conjecture that the diameter is equal to n(n-1)/2 by machine learning and mathematical methods. We identify the conjectured longest element and generate its decomposition of the desired length. We prove a diameter lower bound of n(n-1)/2-n/2 and an upper bound of n(n-1)/2+ 3n by presenting the algorithm with given complexity. We also present several conjectures motivated by numerical experiments, including observations on the central limit phenomenon (with growth approximated by a Gumbel distribution), the uniform distribution for the spectrum of the graph, and a numerical study of sorting networks. To stimulate crowdsourcing activity, we create challenges on the Kaggle platform and invite contributions to improve and benchmark approaches on Cayley graph pathfinding and other tasks.
Probabilistic $(m,n)$-Parking Functions
Published
• View Publication
• BIB
In this article, we establish new results on the probabilistic parking model (introduced by Durmíc, Han, Harris, Ribeiro, and Yin) with $m$ cars and $n$ parking spots and probability parameter $p\in[0,1]$. For any $ m \leq n$ and $p \in [0,1]$, we study the parking preference of the last car, denoted $a_m$, and determine the conditional distribution of $a_m$ and compute its expected value. We show that both formulas depict explicit dependence on the probability parameter $p$. We study the case where $m = cn $ for some $ 0 < c < 1 $ and investigate the asymptotic behavior and show that the presence of ``extra spots'' on the street significantly affects the rate at which the conditional distribution of $ a_m $ converges to the uniform distribution on $[n]$. Even for small $ \varepsilon = 1 - c $, an $ \varepsilon $-proportion of extra spots reduces the convergence rate from $ 1/\sqrt{n} $ to $ 1/n $ when $ p \neq 1/2 $. Additionally, we examine how the convergence rate depends on $c$, while keeping $n$ and $p$ fixed. We establish that as $c$ approaches zero, the total variation distance between the conditional distribution of $a_m$ and the uniform distribution on $[n]$ decreases at least linearly in $c$.
The hard-core model in graph theory
An independent set may not contain both a vertex and one of its neighbours. This basic fact makes the uniform distribution over independent sets rather special. We consider the hard-core model, an essential generalization of the uniform distribution over independent sets. We show how its local analysis yields remarkable insights into the global structure of independent sets in the host graph, in connection with, for instance, Ramsey numbers, graph colourings, and sphere packings.
Longest subsequence for certain repeated up/down patterns in random permutations avoiding a pattern of length three
Published
• View Publication
• BIB
Let $S_n$ denote the set of permutations of $[n]$ and let $σ=σ_1\cdotsσ_n\in S_n$. For a subsequence $\{σ_{i_j}\}_{j=1}^k$ of $\{σ_i\}_{i=1}^n$ of length $k\ge2$, construct
the ``up/down'' sequence $V_1\cdots V_{k-1}$ defined by $$ V_j=\begin{cases} U,\ \text{if}\ σ_{i_j+1}-σ_{i_j}>0;\\ D,\ \text{if}\ σ_{i_j+1}-σ_{i_j}<0.\end{cases} $$ Consider now a fixed up/down pattern: $V_1\cdots V_l$, where $l\in\mathbb{N}$ and $V_j\in\{U, D\},\ j\in[l]$. Given a permutation $σ\in S_n$, consider the length of the longest subsequence of $σ$ that repeats this pattern.
For example, consider $l=3$ and $V_1V_2V_3=UUD$. Then for the permutation $342617985\in S_9$, the length of the longest subsequence that repeats the pattern $UUD$ is 7; it is obtained by 3461798 and 3461785.
The above framework includes two well-known cases. The pattern $U$ is the celebrated case of the longest increasing subsequence. The pattern $UD$ (or $DU$) is the case of the longest alternating subsequence. These have been studied both under the uniform distribution on $S_n$ as well as under the uniform distribution on those permutations in $S_n$ which avoid a particular pattern of length three.
In this paper, we consider the patterns $UUD$ and $UUUD$ under the uniform distribution on those permutations in $S_n$ which avoid the pattern $132$. We prove that the expected value of the longest increasing subsequence following the pattern $UUD$ is asymptotic to $\frac37n$ and the expected value of the longest increasing subsequence following the pattern $UUUD$ is asymptotic to $\frac4{11}n$. (For $UD$ (alternating subsequences) it is known to be $\frac12n$.) This leads directly to appropriate corresponding results for permutations avoiding any particular pattern of length three.
On the satisfiability of random $3$-SAT formulas with $k$-wise independent clauses
The problem of identifying the satisfiability threshold of random $3$-SAT formulas has received a lot of attention during the last decades and has inspired the study of other threshold phenomena in random combinatorial structures. The classical assumption in this line of research is that, for a given set of $n$ Boolean variables, each clause is drawn uniformly at random among all sets of three literals from these variables, independently from other clauses. Here, we keep the uniform distribution of each clause, but deviate significantly from the independence assumption and consider richer families of probability distributions. For integer parameters $n$, $m$, and $k$, we denote by $\DistFamily_k(n,m)$ the family of probability distributions that produce formulas with $m$ clauses, each selected uniformly at random from all sets of three literals from the $n$ variables, so that the clauses are $k$-wise independent. Our aim is to make general statements about the satisfiability or unsatisfiability of formulas produced by distributions in $\DistFamily_k(n,m)$ for different values of the parameters $n$, $m$, and $k$.
Convergence of distributions on paths
Published in In: Fernau, H., Jansen, K. (eds) Fundamentals of Computation Theory. FCT 2023. Lecture Notes in Computer Science, vol 14292. Springer
• View Publication
• BIB
We study the convergence of distributions on finite paths of weighted digraphs, namely the family of Boltzmann distributions and the sequence of uniform distributions. Targeting applications to the convergence of distributions on paths, we revisit some known results from reducible nonnegative matrix theory and obtain new ones, with a systematic use of tools from analytic combinatorics. In several fields of mathematics, computer science and system theory, including concurreny theory, one frequently faces non strongly connected weighted digraphs encoding the elements of combinatorial structures of interest; this motivates our study.
The central limit theorem for entries of random matrices with specific rank over finite fields
Published
• View Publication
• BIB
Let $\mathbb{F}_q$ be the finite field of order $q$, and $\mathcal{A}$ a non-empty proper subset of $\mathbb{F}_q$. Let $\mathbf{M}$ be a random $m \times n$ matrix of rank $r$ over $\mathbb{F}_q$ taken with uniform distribution. It was proved recently by Sanna that as $m,n \to \infty$ and $r,q,\mathcal{A}$ are fixed, the number of entries of $\mathbf{M}$ in $\mathcal{A}$ approaches a normal distribution. The question was raised as to whether or not one can still obtain a central limit theorem of some sort when $r$ goes to infinity in a way controlled by $m$ and $n$. In this paper we answer this question affirmatively.
Boosting uniformity in quasirandom groups: fast and simple
Published
• View Publication
• BIB
We study the communication complexity of multiplying $k\times t$ elements from the group $H=\text{SL}(2,q)$ in the number-on-forehead model with $k$ parties. We prove a lower bound of $(t\log H)/c^{k}$. This is an exponential improvement over previous work, and matches the state-of-the-art in the area.
Relatedly, we show that the convolution of $k^{c}$ independent copies of a 3-uniform distribution over $H^{m}$ is close to a $k$-uniform distribution. This is again an exponential improvement over previous work which needed $c^{k}$ copies. The proofs are remarkably simple; the results extend to other quasirandom groups.
We also show that for any group $H$, any distribution over $H^{m}$ whose weight-$k$ Fourier coefficients are small is close to a $k$-uniform distribution. This generalizes previous work in the abelian setting, and the proof is simpler.
Fast computation of permanents over $\mathbb{F}_3$ via $\mathbb{F}_2$ arithmetic
We present a method of representing an element of $\mathbb{F}_3^n$ as an element of $\mathbb{F}_n^2 \times \mathbb{F}_n^2$ which in practice will be a pair of unsigned integers. We show how to do addition, subtraction and pointwise multiplication and division of such vectors quickly using primitive binary operations (and, or, xor). We use this machinery to develop a fast algorithm for computing the permanent of a matrix in $\mathbb{F}_3^{n\times n}$. We present Julia code for a natural implementation of the permanent and show that our improved implementation gives, roughly, a factor of 80 speedup for problems of practical size. Using this improved code, we perform Monte Carlo simulations that suggest that the distribution of $\mbox{perm}(A)$ tends to the uniform distribution as $n \to \infty$.
Models of random spanning trees
Published
• View Publication
• BIB
There are numerous randomized algorithms to generate spanning trees in a given ambient graph; several target the uniform distribution on trees (UST), while in practice the fastest and most frequently used draw random weights on the edges and then employ a greedy algorithm to choose the minimum-weight spanning tree (MST). Though MST is a workhorse in applications, the mathematical properties of random MST are far less explored than those of UST. In this paper we develop tools for the quantitative study of random MST. We consider the standard case that the weights are drawn i.i.d. from a single distribution on the real numbers, as well as successive generalizations that lead to \emph{product measures}, where the weights are independently drawn from arbitrary distributions.
On entropy Marton-type inequalities and small symmetric differences with cosets of abelian groups
We recognise that an entropy inequality akin to the main intermediate goal of recent works (Gowers, Green, Manners, Tao [3],[2]) regarding a conjecture of Marton provides a black box from which we can also through a short deduction recover another description: if a finite subset $A$ of an abelian group $G$ is such that the distribution of the sums $a+b$ with $(a,b) \in A \times A$ is only slightly more spread out than the uniform distribution on $A$, then $A$ has small symmetric difference with some finite coset of $G$. The resulting bounds are necessarily sharp up to a logarithmic factor.
Regular bipartite multigraphs have many (but not too many) symmetries
Let $k$ and $l$ be integers, both at least 2. A $(k,l)$-bipartite graph is an $l$-regular bipartite multigraph with coloured bipartite sets of size $k$. Define $χ(k,l)$ and $μ(k,l)$ to be the minimum and maximum order of automorphism groups of $(k,l)$-bipartite graphs, respectively. We determine $χ(k,l)$ and $μ(k,l)$ for $k\geq 8$, and analyse the generic situation when $k$ is fixed and $l$ is large. In particular, we show that almost all such graphs have automorphism groups which fix the vertices pointwise and have order far less than $μ(k,l)$. These graphs are intimately connected with both contingency tables with uniform margins and uniform set partitions; we examine the uniform distribution on the set of $k\times k$ contingency tables with uniform margin $l$, showing that with high probability all entries stray far from the mean. We also show that the symmetric group acting on uniform set partitions is non-synchronizing.
The Erdős-Rényi Random Graph Conditioned on Every Component Being a Clique
Motivated by an application in community detection, we consider an \ER random graph conditioned on the rare event that all connected components are fully connected. Such graphs can be considered as partitions of vertices into cliques. Hence, this conditional distribution defines a distribution over partitions. We show that a popular community detection method is equivalent to Bayesian inference with this distribution as prior over the community partitions. Using tools from analytic combinatorics, we prove limit theorems for several graph observables in this conditional distribution: the number of cliques; the number of edges; and the degree distribution. We consider several regimes of the connection probability $p$ as the number of vertices $n$ diverges. For $p=\tfrac{1}{2}$, the conditioning yields the uniform distribution over set partitions, which is well-studied, but has not been studied as a graph distribution before. For $p<\tfrac{1}{2}$, we show that the number of cliques is of the order $n/\sqrt{\log n}$, while for $p>\tfrac{1}{2}$, we prove that the graph consists of a single clique with high probability. This shows that there is a phase transition at $p=\tfrac{1}{2}$. We additionally study the near-critical regime $p_n\downarrow\tfrac{1}{2}$, as well as the sparse regime $p_n\downarrow0$. Finally, we discuss the implications of these results for community detection.
The distribution on permutations induced by a random parking function
Published
• View Publication
• BIB
A parking function on $[n]$ creates a permutation in $S_n$ via the order in which the $n$ cars appear in the $n$ parking spaces. Placing the uniform probability measure on the set of parking functions on $[n]$ induces a probability measure on $S_n$. We initiate a study of some properties of this distribution. Let $P_n^{\text{park}}$ denote this distribution on $S_n$ and let $P_n$ denote the uniform distribution on $S_n$. In particular, we obtain an explicit formula for $P_n^{\text{park}}(σ)$ for all $σ\in S_n$. Then we show that for all but an asymptotically $P_n$-negligible set of permutations, one has $P_n^{\text{park}}(σ)\in\left(\frac{(2-ε)^n}{(n+1)^{n-1}},\frac{(2+ε)^n}{(n+1)^{n-1}}\right)$. However, this accounts for only an exponentially small part of the $P_n^{\text{park}}$-probability. We also obtain an explicit formula for $P_n^{\text{park}}(σ^{-1}_{n-j+1}=i_1,σ^{-1}_{n-j+2}=i_2,\cdots, σ^{-1}_n=i_j)$, the probability that the last $j$ cars park in positions $i_1,\cdots, i_j$ respectively, and show that the $j$-dimensional random vector $(n+1-σ^{-1}_{n-j+l}, n+1-σ^{-1}_{n-j+2},\cdots, n+1-σ^{-1}_{n})$ under $P_n^{\text{park}}$ converges in distribution to a random vector $(\sum_{r=1}^jX_r,\sum_{r=2}^j X_r,\cdots, X_{j-1}+X_j,X_j)$, where $\{X_r\}_{r=1}^j$ are IID with the Borel distribution. We then show that in fact for $j_n=o(n^\frac16)$, the final $j_n$ cars will park in increasing order with probability approaching 1 as $n\to\infty$. We also obtain an explicit formula for the expected value of the left-to-right maximum statistic $X_n^{\text{LR-max}}$, which counts the total number of left-to-right maxima in a permutation, and show that $E_n^{\text{park}}X_n^{\text{LR-max}}$ grows approximately on the order $n^\frac12$.
Better-than-average uniform random variables and Eulerian numbers, or: How many candidates should a voter approve?
Consider $n$ independent random numbers with a uniform distribution on $[0,1]$. The number of them that exceed their mean is shown to have an Eulerian distribution, i.e., it is described by the Eulerian numbers. This is related to, but distinct from, the well known fact that the integer part of the sum of independent random numbers uniform on $[0,1]$ has an Eulerian distribution. One motivation for this problem comes from voting theory.
Limit theorems for fixed point biased permutations avoiding a pattern of length three
Published
• View Publication
• BIB
We prove limit theorems for the number of fixed points occurring in a random pattern-avoiding permutation distributed according to a one-parameter family of biased distributions. The bias parameter exponentially tilts the distribution towards favoring permutations with more or fewer fixed points than is typical under the uniform distribution. One case we study features a phase transition where the limiting distribution changes abruptly from negative binomial to Rayleigh to normal depending on the bias parameter.
Solving a Random Asymmetric TSP Exactly in Quasi-Polynomial Time w.h.p
Published
• View Publication
• BIB
Let the costs $C(i,j)$ for an instance of the Asymmetric Traveling Salesperson Problem (ATSP) be independent copies of a non-negative random variable $C$ from a class of distributions that include the uniform $[0,1]$ distribution and the exponential mean 1 distribution with mean 1. We describe an algorithm that solves ATSP exactly in time $e^{\log^{2+o(1)}n}$, w.h.p.