arXiv++ Combinatorics

Browse math.CO papers from arXiv

random

7100 papers tagged with this keyword
Tropical Support Vector Machines: Evaluations and Extension to Function Spaces
Published • View Publication • BIB
Support Vector Machines (SVMs) are one of the most popular supervised learning models to classify using a hyperplane in an Euclidean space. Similar to SVMs, tropical SVMs classify data points using a tropical hyperplane under the tropical metric with the max-plus algebra. In this paper, first we show generalization error bounds of tropical SVMs over the tropical projective torus. While the generalization error bounds attained via Vapnik-Chervonenkis (VC) dimensions in a distribution-free manner still depend on the dimension, we also show numerically and theoretically by extreme value statistics that the tropical SVMs for classifying data points from two Gaussian distributions as well as empirical data sets of different neuron types are fairly robust against the curse of dimensionality. Extreme value statistics also underlie the anomalous scaling behaviors of the tropical distance between random vectors with additional noise dimensions. Finally, we define tropical SVMs over a function space with the tropical metric.
2021-01-26
On formal concepts of random formal contexts
Published in Information Sciences 578 (2021) 615-620 • View Publication • BIB
In formal concept analysis, it is well-known that the number of formal concepts can be exponential in the worst case. To analyze the average case, we introduce a probabilistic model for random formal contexts and prove that the average number of formal concepts has a superpolynomial asymptotic lower bound.
2021-01-24
On random digraphs and cores
Published in Australasian Journal of Combinatorics, Volume 79(3) (2021), Pages 371-379 • Search Publication
An acyclic homomorphism of a digraph $C$ to a digraph $D$ is a function $ρ\colon V(C)\to V(D)$ such that for every arc $uv$ of $C$, either $ρ(u)=ρ(v)$, or $ρ(u)ρ(v)$ is an arc of $D$ and for every vertex $v\in V(D)$, the subdigraph of $C$ induced by $ρ^{-1}(v)$ is acyclic. A digraph $D$ is a core if the only acyclic homomorphisms of $D$ to itself are automorphisms. In this paper, we prove that for certain choices of $p(n)$, random digraphs $D\in D(n,p(n))$ are asymptotically almost surely cores. For digraphs, this mirrors a result from [A. Bonato and P. Prałat, The good, the bad, and the great: homomorphisms and cores of random graphs, Discrete Math., 309 (2009), no. 18, 5535-5539; MR2567955] concerning random graphs and cores.
2021-01-23
On the anti-commutator of two free random variables
Published in Indiana University Mathematics Journal 2023 • View Publication • BIB
Let $(κ_n(a))_{n\geq 1}$ denote the sequence of free cumulants of a random variable $a$ in a non-commutative probability space $(\mathcal{A},\varphi)$. Based on some considerations on bipartite graphs, we provide a formula to compute the cumulants $(κ_n(ab+ba))_{n\geq 1}$ in terms of $(κ_n(a))_{n\geq 1}$ and $(κ_n(b))_{n\geq 1}$, where $a$ and $b$ are freely independent. Our formula expresses the $n$-th free cumulant of $ab+ba$ as a sum indexed by partitions in the set $\mathcal{Y}_{2n}$ of non-crossing partitions of the form \[ σ=\{B_1,B_3,\dots, B_{2n-1},E_1,\dots,E_r\}, \quad \text{with }r\geq 0, \] such that $i\in B_{i}$ for $i=1,3,\dots,2n-1$ and $|E_j|$ even for $j\leq r$. Therefore, by studying the sets $\mathcal{Y}_{2n}$ we obtain new results regarding the distribution of $ab+ba$. For instance, the size $|\mathcal{Y}_{2n}|$ is closely related to the case when $a,b$ are free Poisson random variables of parameter 1. Our formula can also be expressed in terms of cacti graphs. This graph theoretic approach suggests a natural generalization that allows us to study quadratic forms in $k$ free random variables.
2021-01-22 v2
Girth, magnitude homology, and phase transition of diagonality
Published • View Publication • BIB
This paper studies the magnitude homology of graphs focusing mainly on the relationship between its diagonality and the girth. Magnitude and magnitude homology are formulations of the Euler characteristic and the corresponding homology, respectively, for finite metric spaces, first introduced by Leinster and Hepworth-Willerton. Several authors study them restricting to graphs with path metric, and some properties which are similar to the ordinary homology theory have come to light. However, the whole picture of their behavior is still unrevealed, and it is expected that they catch some geometric properties of graphs. In this article, we show that the girth of graphs partially determines magnitude homology, that is, the larger girth a graph has, the more homologies near the diagonal part vanish. Furthermore, applying this result to a typical random graph, we investigate how the diagonality of graphs varies statistically as the edge density increases. In particular, we show that there exists a phase transition phenomenon for the diagonality.
2021-01-21
Symbolic solutions of some linear recurrences
Published in Jour. Statist. Plann. Inference (2012), 142(2), 423--429 • View Publication • BIB
A symbolic method for solving linear recurrences of combinatorial and statistical interest is introduced. This method essentially relies on a representation of polynomial sequences as moments of a symbol that looks as the framework of a random variable with no reference to any probability space. We give several examples of applications and state an explicit form for the class of linear recurrences involving Sheffer sequences satisfying a special initial condition. The results here presented can be easily implemented in a symbolic software.
2021-01-20
Maximum induced forests in random graphs
Published • View Publication • BIB
We prove that with high probability maximum sizes of induced forests in dense binomial random graphs are concentrated in two consecutive values.
2021-01-20 v2
Sparse expanders have negative curvature
Published • View Publication • BIB
We prove that bounded-degree expanders with non-negative Ollivier-Ricci curvature do not exist, thereby solving a long-standing open problem suggested by Naor and Milman and publicized by Ollivier (2010). In fact, this remains true even if we allow for a vanishing proportion of large degrees, large eigenvalues, and negatively-curved edges. To establish this, we work directly at the level of Benjamini-Schramm limits, and exploit the entropic characterization of the Liouville property on stationary random graphs to show that non-negative curvature and spectral expansion are incompatible "at infinity". We then transfer this result to finite graphs via local weak convergence. The same approach also applies to the Bacry-Emery curvature condition CD$(0,\infty)$, thereby settling a recent conjecture of Cushing, Liu and Peyerimhoff (2019).
2021-01-20 v3
Moderate Deviations in Cycle Count
Published • View Publication • BIB
We prove moderate deviations bounds for the lower tail of the number of odd cycles in a $\calG(n, m)$ random graph. We show that the probability of decreasing triangle density by $t^3$, is $\exp(-Θ(n^2 t^2))$ whenever $n^{-3/4} \ll t^3 \ll 1$, while for $k \ge 5$ we give the same estimate for the probability of decreasing the $k$-cycle density by $t^k$, but for the larger range $n^{-1} \ll t^k \ll 1$. When $m \ge \frac 12 \binom n2$, we also find the leading coefficient in the exponent. This complements results of Goldschmidt et al., who showed that for $n^{-3/2} \ll t^k \ll n^{-1}$, the probability is $\exp(-Θ(n^3 t^{2k}))$. That is, deviations of order smaller than $n^{-1}$ behave like small deviations, and deviations of order larger than $n^{-3/4}$ (for triangles) or $n^{-1}$ (for $k$-cycles with $k \ge 5$) behave like large deviations. For triangles, we conjecture that a sharp change between the two regimes occurs for deviations of size $n^{-3/4}$, which we associate with a single large negative eigenvalue of the adjacency matrix becoming responsible for almost all of the cycle deficit. Our results can be interpreted as finite size effects in phase transitions in constrained random graphs.
2021-01-18 v2
A note on the price of bandit feedback for mistake-bounded online learning
Published • View Publication • BIB
The standard model and the bandit model are two generalizations of the mistake-bound model to online multiclass classification. In both models the learner guesses a classification in each round, but in the standard model the learner recieves the correct classification after each guess, while in the bandit model the learner is only told whether or not their guess is correct in each round. For any set $F$ of multiclass classifiers, define $opt_{std}(F)$ and $opt_{bandit}(F)$ to be the optimal worst-case number of prediction mistakes in the standard and bandit models respectively. Long (Theoretical Computer Science, 2020) claimed that for all $M > 2$ and infinitely many $k$, there exists a set $F$ of functions from a set $X$ to a set $Y$ of size $k$ such that $opt_{std}(F) = M$ and $opt_{bandit}(F) \ge (1 - o(1))(|Y|\ln{|Y|})opt_{std}(F)$. The proof of this result depended on the following lemma, which is false e.g. for all prime $p \ge 5$, $s = \mathbf{1}$ (the all $1$ vector), $t = \mathbf{2}$ (the all $2$ vector), and all $z$. Lemma: Fix $n \ge 2$ and prime $p$, and let $u$ be chosen uniformly at random from $\left\{0, \dots, p-1\right\}^n$. For any $s, t \in \left\{1, \dots, p-1\right\}^n$ with $s \neq t$ and for any $z \in \left\{0, \dots, p-1\right\}$, we have $\Pr(t \cdot u = z \mod p \text{ } | \text{ } s \cdot u = z \mod p) = \frac{1}{p}$. We show that this lemma is false precisely when $s$ and $t$ are multiples of each other mod $p$. Then using a new lemma, we fix Long's proof.
2021-01-18 v3
Large Deviation Principles for Block and Step Graphon Random Graph Models
Borgs, Chayes, Gaudio, Petti and Sen [arXiv:2007.14508] proved a large deviation principle for block model random graphs with rational block ratios. We strengthen their result by allowing any block ratios (and also establish a simpler formula for the rate function). We apply the new result to derive a large deviation principle for graph sampling from any given step graphon.
2021-01-18 v3
Phase transitions and noise sensitivity on the Poisson space via stopping sets and decision trees
Published • View Publication • BIB
Proofs of sharp phase transition and noise sensitivity in percolation have been significantly simplified by the use of randomized algorithms, via the OSSS inequality (proved by O'Donnell, Saks, Schramm and Servedio (2005)) and the Schramm-Steif inequality for the Fourier-Walsh coefficients of functions defined on the Boolean hypercube. In this article, we prove intrinsic versions of the OSSS and Schramm-Steif inequalities for functionals of a general Poisson process, and apply these new estimates to deduce sufficient conditions - expressed in terms of randomized stopping sets - yielding sharp phase transitions, quantitative noise sensitivity, exceptional times and bounds on critical windows for monotonic Boolean Poisson functions. Our analysis is based on a new general definition of `stopping set', not requiring any topological property for the underlying measurable space, as well as on the new concept of a `continuous-time decision tree', for which we establish several fundamental properties. We apply our findings to the $k$-percolation of the Poisson Boolean model and to the Poisson-based confetti percolation with bounded random grains. In these two models, we reduce the proof of sharp phase transitions for percolation, and of noise sensitivity for crossing events, to the construction of suitable randomized stopping sets and the computation of one-arm probabilities. This enables us to settle some open problem suggested by Ahlberg, Tassion and Texeira (2018) on noise sensitivity of crossing events for the planar Poisson Boolean model and also planar Confetti percolation model. Further, we also prove that critical probability is $1/2$ in certain planar confetti percolation models. A special case of this result was conjectured by Benjamini and Schramm (1998) and proved by Müller (2017). Other special cases were proven by Hirsch (2015) and Ghosh and Roy (2018).
2021-01-17 v2
An elementary approach to component sizes in critical random graphs
Published • View Publication • BIB
In this article we introduce a simple tool to derive polynomial upper bounds for the probability of observing unusually large maximal components in some models of random graphs when considered at criticality. Specifically, we apply our method to a model of random intersection graph, a random graph obtained through $p$-bond percolation on a general $d$-regular graph, and a model of inhomogeneous random graph.
2021-01-17 v2
Hamiltonicity of graphs perturbed by a random regular graph
We study Hamiltonicity and pancyclicity in the graph obtained as the union of a deterministic $n$-vertex graph $H$ with $δ(H)\geqαn$ and a random $d$-regular graph $G$, for $d\in\{1,2\}$. When $G$ is a random $2$-regular graph, we prove that a.a.s. $H\cup G$ is pancyclic for all $α\in(0,1]$, and also extend our result to a range of sublinear degrees. When $G$ is a random $1$-regular graph, we prove that a.a.s. $H\cup G$ is pancyclic for all $α\in(\sqrt{2}-1,1]$, and this result is best possible. Furthermore, we show that this bound on $δ(H)$ is only needed when $H$ is `far' from containing a perfect matching, as otherwise we can show results analogous to those of random $2$-regular graphs. Our proofs provide polynomial-time algorithms to find cycles of any length.
2021-01-15 v2
Random and quasi-random designs in group testing
Published • View Publication • BIB
For large classes of group testing problems, we derive lower bounds for the probability that all significant items are uniquely identified using specially constructed random designs. These bounds allow us to optimize parameters of the randomization schemes. We also suggest and numerically justify a procedure of constructing designs with better separability properties than pure random designs. We illustrate theoretical considerations with a large simulation-based study. This study indicates, in particular, that in the case of the common binary group testing, the suggested families of designs have better separability than the popular designs constructed from disjunct matrices. We also derive several asymptotic expansions and discuss the situations when the resulting approximations achieve high accuracy.
The number of optimal matchings for Euclidean Assignment on the line
Published in Journal of Statistical Physics 183:3 (2021) • View Publication • BIB
We consider the Random Euclidean Assignment Problem in dimension $d=1$, with linear cost function. In this version of the problem, in general, there is a large degeneracy of the ground state, i.e. there are many different optimal matchings (say, $\sim \exp(S_N)$ at size $N$). We characterize all possible optimal matchings of a given instance of the problem, and we give a simple product formula for their number. Then, we study the probability distribution of $S_N$ (the zero-temperature entropy of the model), in the uniform random ensemble. We find that, for large $N$, $S_N \sim \frac{1}{2} N \log N + N s + \mathcal{O}\left( \log N \right)$, where $s$ is a random variable whose distribution $p(s)$ does not depend on $N$. We give expressions for the asymptotics of the moments of $p(s)$, both from a formulation as a Brownian process, and via singularity analysis of the generating functions associated to $S_N$. The latter approach provides a combinatorial framework that allows to compute an asymptotic expansion to arbitrary order in $1/N$ for the mean and the variance of
2021-01-13
Loose cores and cycles in random hypergraphs
Published • View Publication • BIB
Inspired by the study of loose cycles in hypergraphs, we define the \emph{loose core} in hypergraphs as a structure which mirrors the close relationship between cycles and $2$-cores in graphs. We prove that in the $r$-uniform binomial random hypergraph $H^r(n,p)$, the order of the loose core undergoes a phase transition at a certain critical threshold and determine this order, as well as the number of edges, asymptotically in the subcritical and supercritical regimes. Our main tool is an algorithm called CoreConstruct, which enables us to analyse a peeling process for the loose core. By analysing this algorithm we determine the asymptotic degree distribution of vertices in the loose core and in particular how many vertices and edges the loose core contains. As a corollary we obtain an improved upper bound on the length of the longest loose cycle in $H^r(n,p)$.
2021-01-13 v2
Pandemic Spread in Communities via Random Graphs
Published • View Publication • BIB
Working in the multi-type Galton-Watson branching-process framework we analyse the spread of a pandemic via a general multi-type random contact graph. Our model consists of several communities, and takes, as input, parameters that outline the contacts between individuals in distinct communities. Given these parameters, we determine whether there will be an outbreak and if yes, we calculate the size of the giant connected component of the graph, thereby, determining the fraction of the population of each type that would be infected before it ends. We show that the pandemic spread has a natural evolution direction given by the Perron-Frobenius eigenvector of a matrix whose entries encode the average number of individuals of one type expected to be infected by an individual of another type. The corresponding eigenvalue is the basic reproduction number of the pandemic. We perform numerical simulations that compare homogeneous and heterogeneous spread graphs and quantify the difference between them. We elaborate on the difference between herd immunity and the end of the pandemic and the effect of countermeasures on the fraction of infected population.
2021-01-13
Unusually large components in near-critical Erdős-Rényi graphs via ballot theorems
Published • View Publication • BIB
We consider the near-critical Erdős-Rényi random graph $G(n,p)$ and provide a new probabilistic proof of the fact that, when $p$ is of the form $p=p(n)=1/n+λ/n^{4/3}$ and $A$ is large, \[\mathbb{P}(|\mathcal{C}_{\max}|>An^{2/3})\asymp A^{-3/2}e^{-\frac{A^3}{8}+\frac{λA^2}{2}-\frac{λ^2A}{2}}\] where $\mathcal{C}_{\max}$ is the largest connected component of the graph. Our result allows $A$ and $λ$ to depend on $n$. While this result is already known, our proof relies only on conceptual and adaptable tools such as ballot theorems, whereas the existing proof relies on a combinatorial formula specific to Erdős-Rényi graphs, together with analytic estimates.
2021-01-12
Independent sets in hypergraphs omitting an intersection
Published • View Publication • BIB
A $k$-uniform hypergraph with $n$ vertices is an $(n,k,\ell)$-omitting system if it does not contain two edges whose intersection has size exactly $\ell$. If in addition it does not contain two edges whose intersection has size greater than $\ell$, then it is an $(n,k,\ell)$-system. Rödl and Šiňajová proved a lower bound for the independence number of $(n,k,\ell)$-systems that is sharp in order of magnitude for fixed $2 \le \ell \le k-1$. We consider the same question for the larger class of $(n,k,\ell)$-omitting systems. For $k\le 2\ell+1$, we believe that the behavior is similar to the case of $(n,k,\ell)$-systems and prove a nontrivial lower bound for the first open case $\ell=k-2$. For $k>2\ell+1$ we give new lower and upper bounds which show that the minimum independence number of $(n,k,\ell)$-omitting systems has a very different behavior than for $(n,k,\ell)$-systems. Our lower bound for $\ell=k-2$ uses some adaptations of the random greedy independent set algorithm, and our upper bounds (constructions) for $k> 2\ell+1$ are obtained from some pseudorandom graphs. We also prove some related results where we forbid more than two edges with a prescribed common intersection size and this leads to some applications in Ramsey theory. For example, we obtain good bounds for the Ramsey number $r_{k}(F^{k},t)$, where $F^{k}$ is the $k$-uniform Fan. Here the behavior is quite different than the case $k=2$ which reduces to the classical graph Ramsey number $r(3,t)$.