uniform distribution
178 papers tagged with this keyword
On the distribution of the entries of a fixed-rank random matrix over a finite field
Published
• View Publication
• BIB
Let $r > 0$ be an integer, let $\mathbb{F}_q$ be a finite field of $q$ elements, and let $\mathcal{A}$ be a nonempty proper subset of $\mathbb{F}_q$. Moreover, let $\mathbf{M}$ be a random $m \times n$ rank-$r$ matrix over $\mathbb{F}_q$ taken with uniform distribution. We prove, in a precise sense, that, as $m, n \to +\infty$ and $r,q,\mathcal{A}$ are fixed, the number of entries of $\mathbf{M}$ that belong to $\mathcal{A}$ approaches a normal distribution.
The secretary problem with items arriving according to a random permutation avoiding a pattern of length three
In the classical secretary problem, $n$ ranked items arrive one by one, and each item's rank relative to its predecessors is noted. The observer must select or reject each item as it arrives, with the object of selecting the item of highest rank. For $M_n\in\{0,1,\cdots, n-1\}$, let $\mathcal{S}(n,M_n)$ denote the strategy whereby the observer rejects the first $M_n$ items, and then selects the first later-arriving item whose rank is higher than that of any of the first $M_n$ items (if such an item exists). If the ranked items arrive in a uniformly random order, it is well-known that the limiting optimal probability of success is $\frac1e$, which occurs if $M_n\sim\frac ne$. It has been shown that when the ranked items arrive according to certain non-uniform distributions on the set of permutations, $\frac1e$ serves as a lower bound for the optimal probability. There is a fundamental reason for this phenomenon. We consider certain distributions for which that reason does not apply. We begin by noting a cooked-up class of distributions for which $\mathcal{S}(n,M)$ yields the lowest possible probability of success -- namely $\frac1n$, for all $M$. We then consider the uniform distribution over all permutations avoiding a particular pattern of length three. In the case of the pattern 231 or 132, for any choice of $M_n$, the strategy $\mathcal{S}(n,M_n)$ yields the very same probability of success; namely $\frac{n+1}{2(2n-1)}$, which gives a limiting probability of $\frac14$. For the pattern 213, the optimal strategy is obtained for $M\in\{0,1\}$, also yielding a limiting probability of $\frac14$. For the pattern 123, the optimal strategy is obtained for $M=1$, yielding a limiting probability of $\frac34$. For the other two patterns, 312 and 321, an optimal strategy will yield a limiting probability of at least $\frac7{16}$.
Optimal mixing of the down-up walk on independent sets of a given size
Published
• View Publication
• BIB
Let $G$ be a graph on $n$ vertices of maximum degree $Δ$. We show that, for any $δ> 0$, the down-up walk on independent sets of size $k \leq (1-δ)α_c(Δ)n$ mixes in time $O_{Δ,δ}(k\log{n})$, thereby resolving a conjecture of Davies and Perkins in an optimal form. Here, $α_{c}(Δ)n$ is the NP-hardness threshold for the problem of counting independent sets of a given size in a graph on $n$ vertices of maximum degree $Δ$. Our mixing time has optimal dependence on $k,n$ for the entire range of $k$; previously, even polynomial mixing was not known. In fact, for $k = Ω_Δ(n)$ in this range, we establish a log-Sobolev inequality with optimal constant $Ω_{Δ,δ}(1/n)$.
At the heart of our proof are three new ingredients, which may be of independent interest. The first is a method for lifting $\ell_\infty$-independence from a suitable distribution on the discrete cube -- in this case, the hard-core model -- to the slice by proving stability of an Edgeworth expansion using a multivariate zero-free region for the base distribution. The second is a generalization of the Lee-Yau induction to prove log-Sobolev inequalities for distributions on the slice with considerably less symmetry than the uniform distribution. The third is a sharp decomposition-type result which provides a lossless comparison between the Dirichlet form of the original Markov chain and that of the so-called projected chain in the presence of a contractive coupling.
Random Algebraic Graphs and Their Convergence to Erdos-Renyi
Published
• View Publication
• BIB
A random algebraic graph is defined by a group $G$ with a uniform distribution over it and a connection $σ:G\longrightarrow[0,1]$ with expectation $p,$ satisfying $σ(g)=σ(g^{-1}).$ The random graph $\mathsf{RAG}(n,G,p,σ)$ with vertex set $[n]$ is formed as follows. First, $n$ independent vectors $x_1,\ldots,x_n$ are sampled uniformly from $G.$ Then, vertices $i,j$ are connected with probability $σ(x_ix_j^{-1}).$ This model captures random geometric graphs over the sphere and the hypercube, certain regimes of the stochastic block model, and random subgraphs of Cayley graphs. The main question of interest to the current paper is: when is a random algebraic graph statistically and/or computationally distinguishable from $\mathsf{G}(n,p)$? Our results fall into two categories. 1) Geometric. We focus on the case $G =\{\pm1\}^d$ and use Fourier-analytic tools. For hard threshold connections, we match [LMSY22b] for $p = ω(1/n)$ and for $1/(r\sqrt{d})$-Lipschitz connections we extend the results of [LR21b] when $d = Ω(n\log n)$ to the non-monotone setting. We study other connections such as indicators of interval unions and low-degree polynomials. 2) Algebraic. We provide evidence for an exponential statistical-computational gap. Consider any finite group $G$ and let $A\subseteq G$ be a set of elements formed by including each set of the form $\{g, g^{-1}\}$ independently with probability $1/2.$ Let $Γ_n(G,A)$ be the distribution of random graphs formed by taking a uniformly random induced subgraph of size $n$ of the Cayley graph $Γ(G,A).$ Then, $Γ_n(G,A)$ and $\mathsf{G}(n,1/2)$ are statistically indistinguishable with high probability over $A$ if and only if $\log|G|\gtrsim n.$ However, low-degree polynomial tests fail to distinguish $Γ_n(G,A)$ and $\mathsf{G}(n,1/2)$ with high probability over $A$ when $\log |G|=\log^{Ω(1)}n.$
Sampling planar tanglegrams and pairs of disjoint triangulations
Published
• View Publication
• BIB
A tanglegram consists of two rooted binary trees and a perfect matching between their leaves, and a planar tanglegram is one that admits a layout with no crossings. We show that the problem of generating planar tanglegrams uniformly at random reduces to the corresponding problem for irreducible planar tanglegram layouts, which are known to be in bijection with pairs of disjoint triangulations of a convex polygon. We extend the flip operation on a single triangulation to a flip operation on pairs of disjoint triangulations. Interestingly, the resulting flip graph is both connected and regular, and hence a random walk on this graph converges to the uniform distribution. We also show that the restriction of the flip graph to the pairs with a fixed triangulation in either coordinate is connected, and give diameter bounds that are near optimal. Our results furthermore yield new insight into the flip graph of triangulations of a convex $n$-gon with a geometric interpretation on the associahedron.
Strong spatial mixing for colorings on trees and its algorithmic applications
Published
• View Publication
• BIB
Strong spatial mixing (SSM) is an important quantitative notion of correlation decay for Gibbs distributions arising in statistical physics, probability theory, and theoretical computer science. A longstanding conjecture is that the uniform distribution on proper $q$-colorings on a $Δ$-regular tree exhibits SSM whenever $q \ge Δ+1$. Moreover, it is widely believed that as long as SSM holds on bounded-degree trees with $q$ colors, one would obtain an efficient sampler for $q$-colorings on all bounded-degree graphs via simple Markov chain algorithms. It is surprising that such a basic question is still open, even on trees, but then again it also highlights how much we still have to learn about random colorings. In this paper, we show the following:
(1) For any $Δ\ge 3$, SSM holds for random $q$-colorings on trees of maximum degree $Δ$ whenever $q \ge Δ+ 3$. Thus we almost fully resolve the aforementioned conjecture. Our result substantially improves upon the previously best bound which requires $q \ge 1.59Δ+γ^*$ for an absolute constant $γ^* > 0$.
(2) For any $Δ\ge 3$ and girth $g = Ω_Δ(1)$, we establish optimal mixing of the Glauber dynamics for $q$-colorings on graphs of maximum degree $Δ$ and girth $g$ whenever $q \ge Δ+3$. Our approach is based on a new general reduction from spectral independence on large-girth graphs to SSM on trees that is of independent interest.
Using the same techniques, we also prove near-optimal bounds on weak spatial mixing (WSM), a closely-related notion to SSM, for the antiferromagnetic Potts model on trees.
Boltzmann Distribution on "Short" Integer Partitions with Power Parts: Limit Laws and Sampling
Published in Advances in Applied Mathematics, Volume 159, August 2024, 102739
• View Publication
• BIB
The paper is concerned with the asymptotic analysis of a family of Boltzmann (multiplicative) distributions over the set $\check{\varLambda}^{q}$ of strict integer partitions (i.e., with unequal parts) into perfect $q$-th powers. A combinatorial link is provided via a suitable conditioning by fixing the partition weight (the sum of parts) and length (the number of parts), leading to uniform distribution on the corresponding subspaces of partitions. The Boltzmann measure is calibrated through the hyper-parameters $\langle N\rangle$ and $\langle M\rangle$ controlling the expected weight and length, respectively. We study ``short'' partitions, where the parameter $\langle M\rangle$ is either fixed or grows slower than for typical plain (unconstrained) partitions. For this model, we obtain a variety of limit theorems including the asymptotics of the cumulative cardinality in the case of fixed $\langle M\rangle$ and a limit shape result in the case of slow growth of $\langle M\rangle$. In both cases, we also characterize the joint distribution of the weight and length, as well as the growth of the smallest and largest parts. Using these results we construct suitable sampling algorithms and analyse their performance.
Optimizers of three-point energies and nearly orthogonal sets
Published
• View Publication
• BIB
This paper is devoted to spherical measures and point configurations optimizing three-point energies. Our main goal is to extend the classic optimization problems based on pairs of distances between points to the context of three-point potentials. In particular, we study three-point analogues of the sphere packing problem and the optimization problem for $p$-frame energies based on three points. It turns out that both problems are inherently connected to the problem of nearly orthogonal sets by Erdős. As the outcome, we provide a new solution of the Erdős problem from the three-point packing perspective. We also show that the orthogonal basis uniquely minimizes the $p$-frame three-point energy when $0<p<1$ in all dimensions. The arguments make use of multivariate polynomials employed in semidefinite programming and based on the classical Gegenbauer polynomials. For $p=1$, we completely solve the analogous problem on the circle. As for higher dimensions, we show that the Hausdorff dimension of minimizers is not greater than $d-2$ for measures on $\mathbb{S}^{d-1}$. As the main ingredient of our proof, we show that the only isotropic measure without obtuse angles is the uniform distribution over an orthonormal basis.
A proof of the union-closed sets conjecture
We provide a proof of the union-closed sets conjecture, by means of a suitable refinement of the breakthrough entropy-approach introduced by Gilmer. The novelty here is to consider a convex combination of $A$ and $A\cup B$, where $A,B$ are independent samples from the uniform distribution over a union-closed family.
Triangle processes on graphs with given degree sequence
Published
• View Publication
• BIB
The switch chain is a well-studied Markov chain which generates random graphs with a given degree sequence and has uniform stationary distribution. Motivated by the high number of triangles seen in some real-world networks, we study a variant of the switch chain which is more likely to produce graphs with higher numbers of triangles. Specifically, we apply a Metropolis scheme designed to have the following stationary distribution: graph $G$ has probability proportional to $λ^{\min\{t(G),ν\}}$, where $t(G)$ is the number of triangles in $G$ and $ν$ is a cut-off value introduced to moderate the impact of graphs with a very high number of triangles. We assume that the "activity" $λ$ satisfies $λ\geq 1$, and call the resulting chain the modified Metropolis switch chain. We prove that the modified Metropolis switch chain is rapidly mixing whenever the (standard) switch chain is rapidly mixing, provided that the activity and maximum degree are not too large.
The triangle switch (or "$\triangle$-switch") chain is a restriction of the switch chain which only performs switches that change the set of triangles in the graph. We prove that the $\triangle$-switch chain is irreducible for any degree sequence with minimum degree at least 3, and prove a rapid mixing result for the modified Metropolis $\triangle$-switch chain.
Finally, we investigate the distribution of triangles in random graphs with given degrees, under both the uniform distribution and the distribution in which graph $G$ has probability proportional to $λ^{t(G)}$. Our analysis implies that the imposition of the cut-off $ν$ does not significantly impact the behaviour of these modified Metropolis chains over polynomially many steps
Fast approximation of search trees on trees with centroid trees
Search trees on trees (STTs) generalize the fundamental binary search tree (BST) data structure: in STTs the underlying search space is an arbitrary tree, whereas in BSTs it is a path. An optimal BST of size $n$ can be computed for a given distribution of queries in $O(n^2)$ time [Knuth 1971] and centroid BSTs provide a nearly-optimal alternative, computable in $O(n)$ time [Mehlhorn 1977].
By contrast, optimal STTs are not known to be computable in polynomial time, and the fastest constant-approximation algorithm runs in $O(n^3)$ time [Berendsohn, Kozma 2022]. Centroid trees can be defined for STTs analogously to BSTs, and they have been used in a wide range of algorithmic applications. In the unweighted case (i.e., for a uniform distribution of queries), a centroid tree can be computed in $O(n)$ time [Brodal et al. 2001; Della Giustina et al. 2019]. These algorithms, however, do not readily extend to the weighted case. Moreover, no approximation guarantees were previously known for centroid trees in either the unweighted or weighted cases.
In this paper we revisit centroid trees in a general, weighted setting, and we settle both the algorithmic complexity of constructing them, and the quality of their approximation. For constructing a weighted centroid tree, we give an output-sensitive $O(n\log h)\subseteq O(n\log n)$ time algorithm, where $h$ is the height of the resulting centroid tree. If the weights are of polynomial complexity, the running time is $O(n\log\log n)$. We show these bounds to be optimal, in a general decision tree model of computation. For approximation, we prove that the cost of a centroid tree is at most twice the optimum, and this guarantee is best possible, both in the weighted and unweighted cases. We also give tight, fine-grained bounds on the approximation-ratio for bounded-degree trees and on the approximation-ratio of more general $α$-centroid trees.
Karp's patching algorithm on random perturbations of dense digraphs
Published
• View Publication
• BIB
We consider the following question. We are given a dense digraph $D_0$ with minimum in- and out-degree at least $αn$, where $α>0$ is a constant. We then add random edges $R$ to $D_0$ to create a digraph $D$. Here an edge $e$ is placed independently into $R$ with probability $n^{-ε}$ where $ε>0$ is a small positive constant. The edges $E(D)$ of $D$ are given independent edge costs $C=C(e),e\in E(D)$, where $C$ has a density $f(x)=a+bx+o(x)$ as $x\to 0$. Here $a>0,b$ are constants. The prime examples will be the uniform $[0,1]$ distribution ($a=1,b=0$) and the exponential mean 1 distribution $EXP(1)$ ($a=1,b=-1$). Let $C(i,j),i,j\in[n]$ be the associated $n\times n$ cost matrix where $C(i,j)=\infty$ if $(i,j)\notin E(D)$. We show that w.h.p.\ the patching algorithm of Karp finds a tour for the asymmetric traveling salesperson problem whose cost is asymptotically equal to the cost of the associated assignment problem. Karp's algorithm runs in polynomial time.
Monotone Subsequences in Locally Uniform Random Permutations
A locally uniform random permutation is generated by sampling $n$ points independently from some absolutely continuous distribution $ρ$ on the plane and interpreting them as a permutation by the rule that $i$ maps to $j$ if the $i$th point from the left is the $j$th point from below. As $n$ tends to infinity, decreasing subsequences in the permutation will appear as curves in the plane, and by interpreting these as level curves, a union of decreasing subsequences give rise to a surface. We show that, under the correct scaling, for any $r\ge0$, the largest union of $\lfloor r\sqrt{n}\rfloor$ decreasing subsequences approaches a limit surface as $n$ tends to infinity, and the limit surface is a solution to a specific variational problem. As a corollary, we prove the existence of a limit shape for the Young diagram associated to the random permutation under the Robinson-Schensted correspondence. In the special case where $ρ$ is the uniform distribution on the diamond $|x|+|y|<1$ we conjecture that the limit shape is triangular, and assuming the conjecture is true we find an explicit formula for the limit surfaces of a uniformly random permutation and recover the famous limit shape of Vershik, Kerov and Logan, Shepp.
Random Walks, Equidistribution and Graphical Designs
Published
• View Publication
• BIB
Let $G=(V,E)$ be a $d$-regular graph on $n$ vertices and let $μ_0$ be a probability measure on $V$. The act of moving to a randomly chosen neighbor leads to a sequence of probability measures supported on $V$ given by $μ_{k+1} = A D^{-1} μ_k$, where $A$ is the adjacency matrix and $D$ is the diagonal matrix of vertex degrees of $G$. Ordering the eigenvalues of $ A D^{-1}$ as $1 = λ_1 \geq |λ_2| \geq \dots \geq |λ_n| \geq 0$, it is well-known that the graphs for which $|λ_2|$ is small are those in which the random walk process converges quickly to the uniform distribution: for all initial probability measures $μ_0$ and all $k \geq 0$, $$ \sum_{v \in V} \left| μ_k(v) - \frac{1}{n} \right|^2 \leq λ_2^{2k}.$$ One could wonder whether this rate can be improved for specific initial probability measures $μ_0$. We show that if $G$ is regular, then for any $1 \leq \ell \leq n$, there exists a probability measure $μ_0$ supported on at most $\ell$ vertices so that $$ \sum_{v \in V} \left| μ_k(v) - \frac{1}{n} \right|^2 \leq λ_{\ell+1}^{2k}.$$ The result has applications in the graph sampling problem: we show that these measures have good sampling properties for reconstructing global averages.
Cheeger Inequalities for Vertex Expansion and Reweighted Eigenvalues
Published
• View Publication
• BIB
The classical Cheeger's inequality relates the edge conductance $φ$ of a graph and the second smallest eigenvalue $λ_2$ of the Laplacian matrix. Recently, Olesker-Taylor and Zanetti discovered a Cheeger-type inequality $ψ^2 / \log |V| \lesssim λ_2^* \lesssim ψ$ connecting the vertex expansion $ψ$ of a graph $G=(V,E)$ and the maximum reweighted second smallest eigenvalue $λ_2^*$ of the Laplacian matrix.
In this work, we first improve their result to $ψ^2 / \log d \lesssim λ_2^* \lesssim ψ$ where $d$ is the maximum degree in $G$, which is optimal assuming the small-set expansion conjecture. Also, the improved result holds for weighted vertex expansion, answering an open question by Olesker-Taylor and Zanetti. Building on this connection, we then develop a new spectral theory for vertex expansion. We discover that several interesting generalizations of Cheeger inequalities relating edge conductances and eigenvalues have a close analog in relating vertex expansions and reweighted eigenvalues. These include an analog of Trevisan's result on bipartiteness, an analog of higher order Cheeger's inequality, and an analog of improved Cheeger's inequality.
Finally, inspired by this connection, we present negative evidence to the $0/1$-polytope edge expansion conjecture by Mihail and Vazirani. We construct $0/1$-polytopes whose graphs have very poor vertex expansion. This implies that the fastest mixing time to the uniform distribution on the vertices of these $0/1$-polytopes is almost linear in the graph size. This does not provide a counterexample to the conjecture, but this is in contrast with known positive results which proved poly-logarithmic mixing time to the uniform distribution on the vertices of subclasses of $0/1$-polytopes.
Virtual permutations and polymorhisms
There is a natural map from a symmetric group $S_n$ to a smaller symmetric group $S_{n-1}$, we write a decomposition of a permutation into a product of disjoint cycles and remove the element $n$ from this expression. For this reason there exists the inverse limit $\mathfrak{S}$ of sets $S_n$. We equip $S_n$ with the uniform distribution (or more generally with an Ewens distribution) and get a structure of a measure space on $\mathfrak{S}$ (it is called 'virtual permutations' or 'Chinese restaurant process'), a double $S_\infty\times S_\infty $ of an infinite symmetric group acts on $\mathfrak{S}$ by left and right 'multiplications'. We discuss the closure of $S_\infty\times S_\infty $ in the semigroup of polymorphisms (spreading maps with spreaded Radon--Nikodym derivatives) of $\mathfrak{S}$. We get formulas for some polymorphisms, in particular for the center of the closure. Expressions are sums of multiple convolutions of Dirichlet distributions, summation sets are certain collections of dessins d'enfant.
Uniform distribution and geometric incidence theory
Published
• View Publication
• BIB
A celebrated unit distance conjecture due to Erd\H os says that that the unit distances cannot arise more than $C_εn^{1+ε}$ times (for any $ε>0$) among $n$ points in the Euclidean plane (see e.g. \cite{SST84} and the references contained therein). In three dimensions, the conjectured bound is $Cn^{\frac{4}{3}}$ (see e.g. \cite{KMSS12} and \cite{Z19}). In dimensions four and higher, this problem, in its general formulation, loses meaning because the Lens example shows that one can construct a set of $n$ points in dimension $4$ and higher where the unit distance arises $\approx n^2$ times (see e.g. \cite{B97}). However, the Lens example is one-dimension in nature, which raises the possibility that the unit distance conjecture is still quite interesting in higher dimensions under additional structural assumptions on the point set. This point of view was explored in \cite{I19}, \cite{IS16}, \cite{IMT12}, \cite{IRU14}, \cite{OO15} and has led to some interesting connections between the unit distance problem and its continuous counterparts, especially the Falconer distance conjecture (\cite{Falc85}).
In this paper, we study the unit distance problem and its variants under the assumption that the underlying family of point sets is uniformly distributed. We prove several incidence bounds in this setting and clarify some key properties of uniformly distributed sequences in the context of incidence problems in combinatorial geometry.
The sharp form of the Kolmogorov--Rogozin inequality and a conjecture of Leader--Radcliffe
Let $X$ be a random variable and define its concentration function by $$\mathcal{Q}_{h}(X)=\sup_{x\in \mathbb{R}}\mathbb{P}(X\in (x,x+h]).$$ For a sum $S_n=X_1+\cdots+X_n$ of independent real-valued random variables the Kolmogorov-Rogozin inequality states that $$\mathcal{Q}_{h}(S_n)\leq C\left(\sum_{i=1}^{n}(1-\mathcal{Q}_{h}(X_i))\right)^{-\frac{1}{2}}.$$
In this paper we give an optimal bound for $\mathcal{Q}_{h}(S_n)$ in terms of $\mathcal{Q}_{h}(X_i)$, which settles a question posed by Leader and Radcliffe in 1994. Moreover, we show that the extremal distributions are mixtures of two uniform distributions each lying on an arithmetic progression.
Approximate counting and sampling via local central limit theorems
Published
• View Publication
• BIB
We give an FPTAS for computing the number of matchings of size $k$ in a graph $G$ of maximum degree $Δ$ on $n$ vertices, for all $k \le (1-δ)m^*(G)$, where $δ>0$ is fixed and $m^*(G)$ is the matching number of $G$, and an FPTAS for the number of independent sets of size $k \le (1-δ) α_c(Δ) n$, where $α_c(Δ)$ is the NP-hardness threshold for this problem. We also provide quasi-linear time randomized algorithms to approximately sample from the uniform distribution on matchings of size $k \leq (1-δ)m^*(G)$ and independent sets of size $k \leq (1-δ)α_c(Δ)n$.
Our results are based on a new framework for exploiting local central limit theorems as an algorithmic tool. We use a combination of Fourier inversion, probabilistic estimates, and the deterministic approximation of partition functions at complex activities to extract approximations of the coefficients of the partition function. For our results for independent sets, we prove a new local central limit theorem for the hard-core model that applies to all fugacities below $λ_c(Δ)$, the uniqueness threshold on the infinite $Δ$-regular tree.
From Trees to Barcodes and Back Again II: Combinatorial and Probabilistic Aspects of a Topological Inverse Problem
Published
• View Publication
• BIB
In this paper we consider two aspects of the inverse problem of how to construct merge trees realizing a given barcode. Much of our investigation exploits a recently discovered connection between the symmetric group and barcodes in general position, based on the simple observation that death order is a permutation of birth order. The first important outcome of our study is a clear combinatorial distinction between the space of phylogenetic trees (as defined by Billera, Holmes and Vogtmann) and the space of merge trees. Generic BHV trees on $n+1$ leaf nodes fall into $(2n-1)!!$ distinct strata, but the analogous number for merge trees is equal to the number of maximal chains in the lattice of partitions, i.e., $(n+1)!n!2^{-n}$. The second aspect of our study is the derivation of precise formulas for the distribution of tree realization numbers (the number of merge trees realizing a given barcode) when we assume that barcodes are sampled using a uniform distribution on the symmetric group. We are able to characterize some of the higher moments of this distribution, thanks in part to a reformulation in terms of Dirichlet convolution. This characterization provides a type of null hypothesis, apparently different from the distributions observed in real neuron data and opens the door to doing more precise science.