random
7100 papers tagged with this keyword
The rate of the convergence of the mean score in random sequence comparison
Published
• View Publication
• BIB
We consider a general class of super-additive scores measuring the similarity of two independent sequences of $n$ i.i.d. letters from a finite alphabet. Our object of interest is the mean score by letter $l_n$. By the subadditivity $l_n$ is nondecreasing and converges to a limit $l$. We give a simple method of bounding the difference $l-l_n$ and obtaining the rate of convergence. Our result generalizes a previous result of Alexander, where only the special case of the longest common subsequence is considered.
L-cumulants, L-cumulant embeddings and algebraic statistics
Published
• View Publication
• BIB
Focusing on the discrete probabilistic setting we generalize the combinatorial definition of cumulants to L-cumulants. This generalization keeps all the desired properties of the classical cumulants like semi-invariance and vanishing for independent blocks of random variables. These properties make L-cumulants useful for the algebraic analysis of statistical models. We illustrate this for general Markov models and hidden Markov processes in the case when the hidden process is binary. The main motivation of this work is to understand cumulant-like coordinates in algebraic statistics and to give a more insightful explanation why tree cumulants give such an elegant description of binary hidden tree models. Moreover, we argue that L-cumulants can be used in the analysis of certain classical algebraic varieties.
A first order phase transition in the threshold-$θ\ge 2$ contact process on random $r$-regular graphs and $r$-trees
Published in Stochastic Processes and Their Applications 123 (2013), no. 2, 561-578
• View Publication
• BIB
We consider the discrete-time threshold-$θ\ge 2$ contact process on a random r-regular graph on n vertices. In this process, a vertex with at least θoccupied neighbors at time t will be occupied at time t+1 with probability p, and vacant otherwise. We show that if $θ\ge 2$ and $r \ge θ+2$, $ε_1$ is small and p is at least $p_1(ε_1)$, then starting from all vertices occupied the fraction of occupied vertices stays above $1-2ε_1$ up to time $\exp(γ_1(r)n)$ with probability at least $1 - \exp(-γ_1(r)n)$. In the other direction, we show that for $p_2 < 1$ there is an $ε_2(p_2)>0$ so that if $p \le p_2$ and the number of occupied vertices in the initial configuration is at most $ε_2(p_2)n$, then with high probability all vertices are vacant at time $C_2(p_2) \log(n)$. These two conclusions imply that on the random r-regular graph there cannot be a quasi-stationary distribution with density of occupied vertices between 0 and $ε_2(p_1)$, and allow us to conclude that the process on the r-tree has a first order phase transition.
Turánnical hypergraphs
Published in Random Structures Algorithms 42 (2013), no. 1, 29-58
• View Publication
• BIB
This paper is motivated by the question of how global and dense restriction sets in results from extremal combinatorics can be replaced by less global and sparser ones. The result we consider here as an example is Turan's theorem, which deals with graphs G=([n],E) such that no member of the restriction set consisting of all r-tuples on [n] induces a copy of K_r.
Firstly, we examine what happens when this restriction set is replaced just by all r-tuples touching a given m-element set. That is, we determine the maximal number of edges in an n-vertex such that no K_r hits a given vertex set.
Secondly, we consider sparse random restriction sets. An r-uniform hypergraph R on vertex set [n] is called Turannical (respectively epsilon-Turannical), if for any graph G on [n] with more edges than the Turan number ex(n,K_r) (respectively (1+\eps)ex(n,K_r), no hyperedge of R induces a copy of K_r in G. We determine the thresholds for random r-uniform hypergraphs to be Turannical and to epsilon-Turannical.
Thirdly, we transfer this result to sparse random graphs, using techniques recently developed by Schacht [Extremal results for random discrete structures] to prove the Kohayakawa-Luczak-Rodl Conjecture on Turan's theorem in random graphs.
Couplings for irregular combinatorial assemblies
Published
• View Publication
• BIB
When approximating the joint distribution of the component counts of a decomposable combinatorial structure that is `almost' in the logarithmic class, but nonetheless has irregular structure, it is useful to be able first to establish that the distribution of a certain sum of non-negative integer valued random variables is smooth. This distribution is not like the normal, and individual summands can contribute a non-trivial amount to the whole, so its smoothness is somewhat surprising. In this paper, we consider two coupling approaches to establishing the smoothness, and contrast the results that are obtained.
Independent sets in random graphs from the weighted second moment method
Published
• View Publication
• BIB
We prove new lower bounds on the likely size of a maximum independent set in a random graph with a given average degree. Our method is a weighted version of the second moment method, where we give each independent set a weight based on the total degree of its vertices.
Properties of Uniform Doubly Stochastic Matrices
We investigate the properties of uniform doubly stochastic random matrices, that is non-negative matrices conditioned to have their rows and columns sum to 1. The rescaled marginal distributions are shown to converge to exponential distributions and indeed even large sub-matrices of side-length $o(n^{1/2-ε})$ behave like independent exponentials. We determine the limiting empirical distribution of the singular values the the matrix. Finally the mixing time of the associated Markov chains is shown to be exactly 2 with high probability.
Gowers norms, regularization and limits of functions on abelian groups
For every natural number k we prove a decomposition theorem for bounded measurable functions on compact abelian groups into a structured part, a quasi random part and a small error term. In this theorem quasi randomness is measured with the Gowers norm U(k+1) and the structured part is a bounded complexity ``nilspace-polynomial'' of degree k. This statement implies a general inverse theorem for the U(k+1) norm. (We discuss some consequences in special families of groups such as bounded exponent groups, zero characteristic groups and the circle group.) Along these lines we introduce a convergence notion and corresponding limit objects for functions on abelian groups. This subject is closely related to the recently developed graph and hypergraph limit theory. An important goal of this paper is to put forward a new algebraic aspect of the notion ``higher order Fourier analysis''. According to this, k-th order Fourier analysis is regarded as the study of continuous morphisms between structures called compact k-step nilspaces. All our proofs are based on an underlying theory of topological nilspace factors of ultra product groups.
Random graphs with few disjoint cycles
Published in Combinatorics, Probability and Computing 20 (2011) 763 -- 775
• View Publication
• BIB
The classical Erdős-Pósa theorem states that for each positive integer k there is an f(k) such that, in each graph G which does not have k+1 disjoint cycles, there is a blocker of size at most f(k); that is, a set B of at most f(k) vertices such that G-B has no cycles. We show that, amongst all such graphs on vertex set {1,..,n}, all but an exponentially small proportion have a blocker of size k. We also give further properties of a random graph sampled uniformly from this class; concerning uniqueness of the blocker, connectivity, chromatic number and clique number. A key step in the proof of the main theorem is to show that there must be a blocker as in the Erdős-Pósa theorem with the extra `redundancy' property that B-v is still a blocker for all but at most k vertices v in B.
Matrices with prescribed row and column sums
Published
• View Publication
• BIB
This is a survey of the recent progress and open questions on the structure of the sets of 0-1 and non-negative integer matrices with prescribed row and column sums. We discuss cardinality estimates, the structure of a random matrix from the set, discrete versions of the Brunn-Minkowski inequality and the statistical dependence between row and column sums.
Diamond-free Families
Published
• View Publication
• BIB
Given a finite poset P, we consider the largest size La(n,P) of a family of subsets of $[n]:=\{1,...,n\}$ that contains no subposet P. This problem has been studied intensively in recent years, and it is conjectured that $π(P):= \lim_{n\rightarrow\infty} La(n,P)/{n choose n/2}$ exists for general posets P, and, moreover, it is an integer. For $k\ge2$ let $\D_k$ denote the $k$-diamond poset $\{A< B_1,...,B_k < C\}$. We study the average number of times a random full chain meets a $P$-free family, called the Lubell function, and use it for $P=\D_k$ to determine $π(\D_k)$ for infinitely many values $k$. A stubborn open problem is to show that $π(\D_2)=2$; here we make progress by proving $π(\D_2)\le 2 3/11$ (if it exists).
The diamond-free process
Published
• View Publication
• BIB
Let K_4^- denote the diamond graph, formed by removing an edge from the complete graph K_4. We consider the following random graph process: starting with n isolated vertices, add edges uniformly at random provided no such edge creates a copy of K_4^-. We show that, with probability tending to 1 as $n \to \infty$, the final size of the graph produced is $Θ(\sqrt{\log(n)} \cdot n^{3/2})$. Our analysis also suggests that the graph produced after i edges are added resembles the random graph, with the additional condition that the edges which do not lie on triangles form a random-looking subgraph.
The final size of the C_4-free process
Published
• View Publication
• BIB
We consider the following random graph process: starting with n isolated vertices, add edges uniformly at random provided no such edge creates a copy of C_4. We show that, with probability tending to 1 as $n \to \infty$, the final graph produced by this process has maximum degree O((n \log n)^{1/3}) and consequently size O(n^{4/3}\log(n)^{1/3}), which are sharp up to constants. This confirms conjectures of Bohman and Keevash and of Osthus and Taraz, and improves upon previous bounds due to Bollobás and Riordan and Osthus and Taraz.
A Comparison of Two Proximity Catch Digraph Families in Testing Spatial Clustering
Published
• View Publication
• BIB
We consider two parametrized random digraph families, namely, proportional-edge and central similarity proximity catch digraphs (PCDs) and compare the performance of these two PCD families in testing spatial point patterns. These PCD families are based on relative positions of data points from two classes and the relative density of the PCDs is used as a statistic for testing segregation and association against complete spatial randomness. When scaled properly, the relative density of a PCD is a U-statistic. We extend the distribution of the relative density of central similarity PCDs for expansion parameter being larger than one. We compare the asymptotic distribution of the statistic for the two PCD families, using the standard central limit theory of U-statistics. We compare finite sample performance of the tests by Monte Carlo simulations and prove the consistency of the tests under the alternatives. The asymptotic performance of the tests under the alternatives is assessed by Pitman's asymptotic efficiency. We find the optimal expansion parameters of the PCDs for testing each of the segregation and association alternatives in finite samples and in the limit. We demonstrate that in terms of empirical power (i.e., for finite samples) relative density of central similarity PCD has better performance (which occurs for expansion parameter values larger than one) under segregation alternative, while relative density of proportional-edge PCD has better performance under association alternative. The methods are illustrated in a real-life example from plant ecology.
Asymptotic normality of the size of the giant component via a random walk
Published in J. Combinatorial Theory B 102 (2012), 53--61
• View Publication
• BIB
In this paper we give a simple new proof of a result of Pittel and Wormald concerning the asymptotic value and (suitably rescaled) limiting distribution of the number of vertices in the giant component of $G(n,p)$ above the scaling window of the phase transition. Nachmias and Peres used martingale arguments to study Karp's exploration process, obtaining a simple proof of a weak form of this result. We use slightly different martingale arguments to obtain a much sharper result with little extra work.
The rigidity transition in random graphs
Published
• View Publication
• BIB
As we add rigid bars between points in the plane, at what point is there a giant (linear-sized) rigid component, which can be rotated and translated, but which has no internal flexibility? If the points are generic, this depends only on the combinatorics of the graph formed by the bars. We show that if this graph is an Erdos-Renyi random graph G(n,c/n), then there exists a sharp threshold for a giant rigid component to emerge. For c < c_2, w.h.p. all rigid components span one, two, or three vertices, and when c > c_2, w.h.p. there is a giant rigid component. The constant c_2 \approx 3.588 is the threshold for 2-orientability, discovered independently by Fernholz and Ramachandran and Cain, Sanders, and Wormald in SODA'07. We also give quantitative bounds on the size of the giant rigid component when it emerges, proving that it spans a (1-o(1))-fraction of the vertices in the (3+2)-core. Informally, the (3+2)-core is maximal induced subgraph obtained by starting from the 3-core and then inductively adding vertices with 2 neighbors in the graph obtained so far.
Fractional colorings of cubic graphs with large girth
We show that every (sub)cubic n-vertex graph with sufficiently large girth has fractional chromatic number at most 2.2978 which implies that it contains an independent set of size at least 0.4352n. Our bound on the independence number is valid to random cubic graphs as well as it improves existing lower bounds on the maximum cut in cubic graphs with large girth.
The sharp threshold for bootstrap percolation in all dimensions
Published
• View Publication
• BIB
In r-neighbour bootstrap percolation on a graph G, a (typically random) set A of initially 'infected' vertices spreads by infecting (at each time step) vertices with at least r already-infected neighbours. This process may be viewed as a monotone version of the Glauber dynamics of the Ising model, and has been extensively studied on the d-dimensional grid $[n]^d$. The elements of the set A are usually chosen independently, with some density p, and the main question is to determine $p_c([n]^d,r)$, the density at which percolation (infection of the entire vertex set) becomes likely.
In this paper we prove, for every pair $d \ge r \ge 2$, that there is a constant L(d,r) such that $p_c([n]^d,r) = [(L(d,r) + o(1)) / log_(r-1) (n)]^{d-r+1}$ as $n \to \infty$, where $log_r$ denotes an r-times iterated logarithm. We thus prove the existence of a sharp threshold for percolation in any (fixed) number of dimensions. Moreover, we determine L(d,r) for every pair (d,r).
Finding Hidden Cliques in Linear Time with High Probability
Published
• View Publication
• BIB
We are given a graph $G$ with $n$ vertices, where a random subset of $k$ vertices has been made into a clique, and the remaining edges are chosen independently with probability $\tfrac12$. This random graph model is denoted $G(n,\tfrac12,k)$. The hidden clique problem is to design an algorithm that finds the $k$-clique in polynomial time with high probability. An algorithm due to Alon, Krivelevich and Sudakov uses spectral techniques to find the hidden clique with high probability when $k = c \sqrt{n}$ for a sufficiently large constant $c > 0$. Recently, an algorithm that solves the same problem was proposed by Feige and Ron. It has the advantages of being simpler and more intuitive, and of an improved running time of $O(n^2)$. However, the analysis in the paper gives success probability of only $2/3$. In this paper we present a new algorithm for finding hidden cliques that both runs in time $O(n^2)$, and has a failure probability that is less than polynomially small.
On the Mixing of Diffusing Particles
Published in Phys. Rev. E 82, 061103 (2010)
• View Publication
• BIB
We study how the order of N independent random walks in one dimension evolves with time. Our focus is statistical properties of the inversion number m, defined as the number of pairs that are out of sort with respect to the initial configuration. In the steady-state, the distribution of the inversion number is Gaussian with the average <m>~N^2/4 and the standard deviation sigma N^{3/2}/6. The survival probability, S_m(t), which measures the likelihood that the inversion number remains below m until time t, decays algebraically in the long-time limit, S_m t^{-beta_m}. Interestingly, there is a spectrum of N(N-1)/2 distinct exponents beta_m(N). We also find that the kinetics of first-passage in a circular cone provides a good approximation for these exponents. When N is large, the first-passage exponents are a universal function of a single scaling variable, beta_m(N)--> beta(z) with z=(m-<m>)/sigma. In the cone approximation, the scaling function is a root of a transcendental equation involving the parabolic cylinder equation, D_{2 beta}(-z)=0, and surprisingly, numerical simulations show this prediction to be exact.