tree
6861 papers tagged with this keyword
Best Match Graphs with Binary Trees
Published
• View Publication
• BIB
Best match graphs (BMG) are a key intermediate in graph-based orthology detection and contain a large amount of information on the gene tree. We provide a near-cubic algorithm to determine whether a BMG is binary-explainable, i.e., whether it can be explained by a fully resolved gene tree and, if so, to construct such a tree. Moreover, we show that all such binary trees are refinements of the unique binary-resolvable tree (BRT), which in general is a substantial refinement of the also unique least resolved tree of a BMG. Finally, we show that the problem of editing an arbitrary vertex-colored graph to a binary-explainable BMG is NP-complete and provide an integer linear program formulation for this task.
Monitoring the edges of a graph using distances
Published in Discrete Applied Mathematics 319:424-438, 2022
• View Publication
• BIB
We introduce a new graph-theoretic concept in the area of network monitoring. A set $M$ of vertices of a graph $G$ is a \emph{distance-edge-monitoring set} if for every edge $e$ of $G$, there is a vertex $x$ of $M$ and a vertex $y$ of $G$ such that $e$ belongs to all shortest paths between $x$ and $y$. We denote by $dem(G)$ the smallest size of such a set in $G$. The vertices of $M$ represent distance probes in a network modeled by $G$; when the edge $e$ fails, the distance from $x$ to $y$ increases, and thus we are able to detect the failure. It turns out that not only we can detect it, but we can even correctly locate the failing edge.
In this paper, we initiate the study of this new concept. We show that for a nontrivial connected graph $G$ of order $n$, $1\leq dem(G)\leq n-1$ with $dem(G)=1$ if and only if $G$ is a tree, and $dem(G)=n-1$ if and only if it is a complete graph. We compute the exact value of $dem$ for grids, hypercubes, and complete bipartite graphs.
Then, we relate $dem$ to other standard graph parameters. We show that $demG)$ is lower-bounded by the arboricity of the graph, and upper-bounded by its vertex cover number. It is also upper-bounded by twice its feedback edge set number. Moreover, we characterize connected graphs $G$ with $dem(G)=2$.
Then, we show that determining $dem(G)$ for an input graph $G$ is an NP-complete problem, even for apex graphs. There exists a polynomial-time logarithmic-factor approximation algorithm, however it is NP-hard to compute an asymptotically better approximation, even for bipartite graphs of small diameter and for bipartite subcubic graphs. For such instances, the problem is also unlikey to be fixed parameter tractable when parameterized by the solution size.
The clustered selected-internal Steiner tree problem
Published
• View Publication
• BIB
Given a complete graph $G=(V,E)$, with nonnegative edge costs, two subsets $R \subset V$ and $R^{\prime} \subset R$, a partition $\mathcal{R}=\{R_1,R_2,\ldots,R_k\}$ of $R$, $R_i \cap R_j=φ$, $i \neq j$ and $\mathcal{R}^{\prime}=\{R^{\prime}_1,R^{\prime}_2,\ldots,R^{\prime}_k\}$ of $R^{\prime}$, $R^{\prime}_i \subset R_i$, a clustered Steiner tree is a tree $T$ of $G$ that spans all vertices in $R$ such that $T$ can be cut into $k$ subtrees $T_i$ by removing $k-1$ edges and each subtree $T_i$ spanning all vertices in $R_i$, $1 \leq i \leq k$. The cost of a clustered Steiner tree is defined to be the sum of the costs of all its edges. A clustered selected-internal Steiner tree of $G$ is a clustered Steiner tree for $R$ if all vertices in $R^{\prime}_i$ are internal vertices of $T_i$, $1 \leq i \leq k$. The clustered selected-internal Steiner tree problem is concerned with the determination of a clustered selected-internal Steiner tree $T$ for $R$ and $R^{\prime}$ in $G$ with minimum cost. In this paper, we present the first known approximation algorithm with performance ratio $(ρ+4)$ for the clustered selected-internal Steiner tree problem, where $ρ$ is the best-known performance ratio for the Steiner tree problem.
Spanning trees at the connectivity threshold
Published
• View Publication
• BIB
We present an explicit connected spanning structure that appears in a random graph just above the connectivity threshold with high probability.
On the maximum mean subtree order of trees
Published in European Journal of Combinatorics 2021
• View Publication
• BIB
A subtree of a tree is any induced subgraph that is again a tree (i.e., connected). The mean subtree order of a tree is the average number of vertices of its subtrees. This invariant was first analyzed in the 1980s by Jamison. An intriguing open question raised by Jamison asks whether the maximum of the mean subtree order, given the order of the tree, is always attained by some caterpillar. While we do not completely resolve this conjecture, we find some evidence in its favor by proving different features of trees that attain the maximum. For example, we show that the diameter of a tree of order $n$ with maximum mean subtree order must be very close to $n$. Moreover, we show that the maximum mean subtree order is equal to $n - 2\log_2 n + O(1)$. For the local mean subtree order, which is the average order of all subtrees containing a fixed vertex, we can be even more precise: we show that its maximum is always attained by a broom and that it is equal to $n - \log_2 n + O(1)$.
Random walks on stochastic uniform growth trees: Analytical formula for mean first-passage time
Published
• View Publication
• BIB
As known, the commonly-utilized ways to determine mean first-passage time $\overline{\mathcal{F}}$ for random walk on networks are mainly based on Laplacian spectra. However, methods of this type can become prohibitively complicated and even fail to work when the Laplacian matrix of network under consideration is difficult to describe in the first place. In this paper, we propose an effective approach to determining quantity $\overline{\mathcal{F}}$ on some widely-studied tree networks. To this end, we first build up a general formula between Wiener index $\mathcal{W}$ and $\overline{\mathcal{F}}$ on a tree. This enables us to convert issues to answer into calculation of $\mathcal{W}$ on networks in question. As opposed to most of previous work focusing on deterministic growth trees, our goal is to consider stochastic case. Towards this end, we establish a principled framework where randomness is introduced into the process of growing trees. As an immediate consequence, the previously published results upon deterministic cases are thoroughly covered by formulas established in this paper. Additionally, it is also straightforward to obtain Kirchhoff index on our tree networks using the proposed approach. Most importantly, our approach is more manageable than many other methods including spectral technique in situations considered herein.
Trees and cycles
Let $T$ be a tree on $n$ vertices. We can regard the edges of $T$ as transpositions of the vertex set; their product (in any order) is a cyclic permutation. All possible cyclic permutations arise (each exactly once) if and only if the tree is a star. In this paper we find the number of realised cycles, and obtain some results on the number of realisations of each cycle, for other trees. We also solve the inverse problem of the number of trees which give rise to a given cycle. On the way, we meet some familiar number sequences including the Euler and Fuss--Catalan numbers.
A tree expansion formula of a homology intersection numbers on the configuration space $\mathcal{M}_{0,n}$
Published
• View Publication
• BIB
In \cite{M}, Sebastian Mizera discovered a tree expansion formula of a homology intersection number on the configuration space $\mathcal{M}_{0,n}$. The formula originates in a study of Kawai-Lewellen-Tye relation in string theory. In this paper, we give an elementary proof of the formula. The basic ingredients are the combinatorics of the real moduli space $\overline{\mathcal{M}}_{0,n}(\R)$ and a combinatorial identity related to the face number of the associahedron.
The p-Airy distribution
In this manuscript we consider the set of Dyck paths equipped with the uniform measure, and we study the statistical properties of a deformation of the observable "area below the Dyck path" as the size $N$ of the path goes to infinity. The deformation under analysis is apparently new: while usually the area is constructed as the sum of the heights of the steps of the Dyck path, here we regard it as the sum of the lengths of the connected horizontal slices under the path, and we deform it by applying to the lengths of the slices a positive regular function $ω(\ell)$ such that $ω(\ell) \sim \ell^p$ for large argument. This shift of paradigm is motivated by applications to the Euclidean Random Assignment Problem in Random Combinatorial Optimization, and to Tree Hook Formulas in Algebraic Combinatorics.
For $p \in \mathbb{R}^+ \smallsetminus \left\{ \frac{1}{2}\right\}$, we characterize the statistical properties of the deformed area as a function of the deformation function $ω(\ell)$ by computing its integer moments, finding a generalization of a well-known recursion for the moments of the area-Airy distribution, due to Takács. Most of the properties of the distribution of the deformed area are \emph{universal}, meaning that they depend on the deformation parameter $p$, but not on the microscopic details of the function $ω(\ell)$. We call \emph{$p$-Airy distribution} this family of universal distributions.
Using graph theory to compute Laplace operators arising in a model for blood flow in capillary network
Maintaining cerebral blood flow is critical for adequate neuronal function. Previous computational models of brain capillary networks have predicted that heterogeneous cerebral capillary flow patterns result in lower brain tissue partial oxygen pressures. It has been suggested that this may lead to number of diseases such as Alzheimer's disease, acute ischemic stroke, traumatic brain injury and ischemic heart disease. We have previously developed a computational model that was used to describe in detail the effect of flow heterogeneities on tissue oxygen levels. The main result in that paper was that, for a general class of capillary networks, perturbations of segment diameters or conductances always lead to decreased oxygen levels. This result was verified using both numerical simulations and mathematical analysis. However, the analysis depended on a novel conjecture concerning the Laplace operator of functions related to the segment flow rates and how they depend on the conductances. The goal of this paper is to give a mathematically rigorous proof of the conjecture for a general class of networks. The proof depends on determining the number of trees and forests in certain graphs arising from the capillary network.
Impartial Achievement Games on Convex Geometries
Published
• View Publication
• BIB
We study a game where two players take turns selecting points of a convex geometry until the convex closure of the jointly selected points contains all the points of a given winning set. The winner of the game is the last player able to move. We develop a structure theory for these games and use it to determine the nim number for several classes of convex geometries, including one-dimensional affine geometries, vertex geometries of trees, and games with a winning set consisting of extreme points.
Enumerative and planar combinatorics of trivariate monomial resolutions
The canonical sylvan resolution is a resolution of an arbitrary monomial ideal over a polynomial ring that is minimal and has an explicit combinatorial formula for the differential. The differential is a weighted sum over lattice paths of weights of chain-link fences, which are sequences of faces that are linked to each other via higher-dimensional analogues of spanning trees. Along a lattice path in the three-variable case, these weights can be condensed to a single weight contributing to the combinatorial formula for the differential that bypasses any computation of chain-link fences. The main results in this paper express the sylvan matrix entries for monomial ideals in three variables as a sum over lattice paths of simpler weights that depend only on the number of specific Koszul simplicial complexes that lie along the corresponding lattice path. Certain entries have numerators equal to the number of lattice paths in $\mathbb{N}^2$ that follow specific restrictions.
On the Markov numbers: fixed numerator, denominator, and sum conjectures
Published
• View Publication
• BIB
The Markov numbers are the positive integer solutions of the Diophantine equation $x^2 + y^2 + z^2 = 3xyz$. Already in 1880, Markov showed that all these solutions could be generated along a binary tree. So it became quite usual (and useful) to index the Markov numbers by the rationals between 0 and 1 which stand at the same place in the Stern-Brocot binary tree. The Frobenius conjecture claims that each Markov number appears at most once in the tree.
In particular, if the conjecture is true, the order of Markov numbers would establish a new strict order on the rationals. Aigner suggested three conjectures to better understand this order. The first one has already been solved for a few months. We prove that the other two conjectures are also true.
Along the way, we generalize Markov numbers to any couple (p,q) of nonnegative integers (not only when they are relatively primes) and conjecture that the unicity is still true as soon as $p \leq q$. Finally, we show that the three conjectures are in fact true for this superset.
Laplacian Fractional Revival on Graphs
Published
• View Publication
• BIB
We develop the theory of fractional revival in the quantum walk on a graph using its Laplacian matrix as the Hamiltonian. We first give a spectral characterization of Laplacian fractional revival, which leads to a polynomial time algorithm to check this phenomenon and find the earliest time when it occurs. We then apply the characterization theorem to special families of graphs. In particular, we show that no tree admits Laplacian fractional revival except for the paths on two and three vertices, and the only graphs on a prime number of vertices that admit Laplacian fractional revival are double cones. Finally, we construct, through Cartesian products and joins, several infinite families of graphs that admit Laplacian fractional revival; some of these graphs exhibit polygamous fractional revival.
Log-rank and lifting for AND-functions
Published
• View Publication
• BIB
Let $f: \{0,1\}^n \to \{0, 1\}$ be a boolean function, and let $f_\land (x, y) = f(x \land y)$ denote the AND-function of $f$, where $x \land y$ denotes bit-wise AND. We study the deterministic communication complexity of $f_\land$ and show that, up to a $\log n$ factor, it is bounded by a polynomial in the logarithm of the real rank of the communication matrix of $f_\land$. This comes within a $\log n$ factor of establishing the log-rank conjecturefor AND-functions with no assumptions on $f$. Our result stands in contrast with previous results on special cases of the log-rank conjecture, which needed significant restrictions on $f$ such as monotonicity or low $\mathbb{F}_2$-degree. Our techniques can also be used to prove (within a $\log n$ factor) a lifting theorem for AND-functions, stating that the deterministic communication complexity of $f_\land$ is polynomially-related to the AND-decision tree complexity of $f$.
The results rely on a new structural result regarding boolean functions $f:\{0, 1\}^n \to \{0, 1\}$ with a sparse polynomial representation, which may be of independent interest. We show that if the polynomial computing $f$ has few monomials then the set system of the monomials has a small hitting set, of size poly-logarithmic in its sparsity. We also establish extensions of this result to multi-linear polynomials $f:\{0,1\}^n \to \mathbb{R}$ with a larger range.
Revisiting Shao and Sokal's $B_2$ index of phylogenetic balance
Published in Journal of Mathematical Biology 83:52 (2021)
• View Publication
• BIB
Measures of phylogenetic balance, such as the Colless and Sackin indices, play an important role in phylogenetics. Unfortunately, these indices are specifically designed for phylogenetic trees, and do not extend naturally to phylogenetic networks (which are increasingly used to describe reticulate evolution). This led us to consider a lesser-known balance index, whose definition is based on a probabilistic interpretation that is equally applicable to trees and to networks. This index, known as the $B_2$ index, was first proposed by Shao and Sokal in 1990. Surprisingly, it does not seem to have been studied mathematically since. Likewise, it is used only sporadically in the biological literature, where it tends to be viewed as arcane. In this paper, we study mathematical properties of $B_2$ such as its expectation and variance under the most common models of random trees and its extremal values over various classes of phylogenetic networks. We also assess its relevance in biological applications, and find it to be comparable to that of the Colless and Sackin indices. Altogether, our results call for a reevaluation of the status of this somewhat forgotten measure of phylogenetic balance.
The Horton-Strahler Number of Conditioned Galton-Watson Trees
Published in Electron. J. Probab. 26 1 - 29, 2021
• View Publication
• BIB
The Horton-Strahler number of a tree is a measure of its branching complexity; it is also known in the literature as the register function. We show that for critical Galton-Watson trees with finite variance conditioned to be of size $n$, the Horton-Strahler number grows as $\frac{1}{2}\log_2 n$ in probability. We further define some generalizations of this number. Among these are the rigid Horton-Strahler number and the $k$-ary register function, for which we prove asymptotic results analogous to the standard case.
Complexity Measures on the Symmetric Group and Beyond
We extend the definitions of complexity measures of functions to domains such as the symmetric group. The complexity measures we consider include degree, approximate degree, decision tree complexity, sensitivity, block sensitivity, and a few others. We show that these complexity measures are polynomially related for the symmetric group and for many other domains.
To show that all measures but sensitivity are polynomially related, we generalize classical arguments of Nisan and others. To add sensitivity to the mix, we reduce to Huang's sensitivity theorem using "pseudo-characters", which witness the degree of a function.
Using similar ideas, we extend the characterization of Boolean degree 1 functions on the symmetric group due to Ellis, Friedgut and Pilpel to the perfect matching scheme. As another application of our ideas, we simplify the characterization of maximum-size $t$-intersecting families in the symmetric group and the perfect matching scheme.
A conjecture about spectral distances between cycles, paths and certain trees
Published
• View Publication
• BIB
We confirm the following conjecture which has been proposed in [{\em Linear Algebra and its Applications}, {\bf 436} (2012), No. 5, 1425-1435.]: $$ 0.945\approx\displaystyle\lim_{n\longrightarrow \infty}σ(P_n,Z_n)=\displaystyle\lim_{n\longrightarrow \infty}σ(W_n,Z_n)=\frac{1}{2}\displaystyle\lim_{n\longrightarrow \infty}σ(P_n,W_n);\ \displaystyle\lim_{n\longrightarrow \infty}σ(C_{2n},Z_{2n})=2,$$ where $σ(G_1,G_2)=\sum_{i=1}^n |λ_i(G_1)-λ_i(G_2)|$ is the spectral distance between $n$ vertex non-isomorphic graphs $G_1$ and $G_2$ with adjacency spectra $λ_1(G_i) \geq λ_2(G_i) \geq \cdots \geq λ_n(G_i)$ for $i=1,2$, and $P_n$ and $C_n$ denote the path and cycle on $n$ vertices, respectively; $Z_n$ denotes the coalescence of $P_{n-2}$ and $P_3$ on one of the vertices of degree 1 of $P_{n-2}$ and the vertex of degree $2$ of $P_3$; and $W_n$ denotes the coalescence of $Z_{n-2}$ and $P_3$ on the vertex of degree 1 of $Z_{n-2}$ which is adjacent to a vertex of degree $2$ and the vertex of degree $2$ of $P_3$.
On edge-weighted mean eccentricity of graphs
Let $G$ be a connected edge-weighted graph of order $n$ and size $m$. Let $w:E(G)\rightarrow \mathbb{R}^{\geq 0}$ be the weighting function. We assume that $w$ is normalised, that is, $\sum_{e\in E(G)} w(e)=m$. The weighted distance $d_w(u,v)$ between any two vertices $u$ and $v$ is the least weight between them and the eccentricity $e_w(v)$ of a vertex $v$ is the weighted distance from $v$ to a vertex farthest from it in $G$. The mean(average) eccentricity of $G$, $avec(G,w)$, is the (weighted) mean of all eccentricities in $G$. We obtain upper and lower bounds on $avec(G,w)$ in terms of $n$, $m$ or edge-connectivity $λ$ for two cases: $G$ is a tree and $G$ is connected but not a tree. In addition, we obtain the Nordhaus-Gaddum-type results for edge-weighted average eccentricity.