arXiv++ Combinatorics

Browse math.CO papers from arXiv

boolean function

334 papers tagged with this keyword
2026-10-02
Small Influential Coalitions in $[0,1]^n$ via the Junta Theorem
Every monotone Boolean function on the continuous cube admits a coalition of $O(n/(\varepsilon\log n))$ coordinates that can force a fixed output (either zero or one) with probability at least $1-\varepsilon$. This old conjecture of mine was recently proven by Chattopadhyay and Gurumukhani~\cite{CG}. This writeup contains a simplification of their proof, generated by AI after suggesting the use of the junta theorem of~\cite{F98} and the discretization procedure of ~\cite{F04}.
2026-10-01
Non-Malleable Affine Extractors with Small Error and Complexity Lower Bounds
We construct explicit non-malleable affine extractors for every constant entropy rate, with linear output length and exponentially small error, against any fixed number of affine tamperings without fixed points. For every fixed $0<η<1$ and $t$, we also obtain entropy threshold $C_{η,t}n/\log n$, output length $\lfloor n^{1-η}\rfloor$, and error $2^{-n^{1-η}}$ against $t$ tamperings. Our extractors, as well as the directional affine extractors of Li and Zhong (CCC 2024), yield explicit Boolean functions with correlation $2^{-Ω(n)}$ against weakly read-once linear branching programs of size $2^{Ω(n)}$. For non-oblivious decision trees, we prove linear depth lower bounds for queries of each fixed degree $r\ge2$. Applying Li's sumset extractor (FOCS 2023) gives depth $Ω_δ((n/\ell)\log\ell)$ for growing locality $\ell\le n^{1-δ}$, where $0<δ<1$ is fixed. In the same range, directional affine extractors give correlation $2^{-Ω(n/\sqrt\ell)}$ against local trees of depth $c(n/\ell)\log\ell/\log\log\ell$, for a sufficiently small constant $c>0$. Our extractors derandomize the lossless lifting of Efremenko and Itsykson (STOC 2026). For every fixed $0<ξ<1$, this gives explicit polynomial-size unsatisfiable CNFs on $N$ variables whose $\mathrm{Res}(\oplus)$ refutations of resolution depth at most $N$ require size at least $2^{(1-ξ)N}$. Separately, parity substitutions give polynomial-size CNFs on $N$ variables with polynomial-size ordinary-resolution proofs for which every $\mathrm{Res}(\oplus)$ refutation of size $S$ and depth $d$ satisfies $d\log(2S)=Ω(N^2)$. This removes the $\log^2 N$ loss in the tradeoff of Itsykson, Podolskii, and Shekhovtsov (CCC 2026).
2026-09-29 v2
A weak Hellinger inequality for noisy Boolean channels
A weak form of the Hellinger conjecture of Anantharam, Bogdanov, Chakrabarti, Jayram, and Nair for the binary symmetric channel is proved: dictator functions maximize Hellinger $Φ$-entropy among all Boolean functions of the input and all one-bit statistics of the output of a noisy channel. The technical heart of the matter is an explicit inequality in three real parameters, which is proved using explicit polynomial approximations and computer-assisted positivity checks. The results are also formally verified in Lean 4.
2026-09-26 v2
Certificate Complexity of Elementary Symmetric Boolean Functions of Arbitrary Degree
Let $σ_{n,d}$ denote the elementary symmetric Boolean function of $n$ variables and degree $d$. Previous work determined $C(σ_{n,d})$ when $d$ is odd or a power of two, but the general even non-power-of-two case remained open. We determine the certificate complexity for every degree $1\le d\le n$, thereby completing the classification for elementary symmetric Boolean functions. Writing $d=2^t m$ with $m$ odd, we obtain an explicit formula in which the possible deficit from the maximal value $n$ is controlled by $2^t$, while the exact value is determined by a binary containment condition involving $m$. In particular, \[ n-2^{ν_2(d)}+1\le C(σ_{n,d})\le n, \] and we characterize when the upper bound is attained. Moreover, we determine the least positive period of the certificate-complexity deficit $Δ_d(n)=n-C(σ_{n,d})$: it is $1$ for odd $d$, equals $d$ when $d$ is a power of two, and equals $2^{\lfloor\log_2 d\rfloor+1}$ for even non-power-of-two $d$. This least-period problem is distinct from the classical periodicity of the underlying value sequence $\binom{j}{d}\bmod 2$. The known odd-degree and power-of-two formulas are recovered as special cases.
2026-09-25
Stability of the Courtade-Kumar inequality
We prove dimension-independent stability for the Courtade-Kumar inequality: a Boolean function $f:\{-1,1\}^n\to\{-1,1\}$ whose information is close to the dictator value is close in probability to a signed dictator. The correlation dependence is sharp in order near zero and, for increasing functions, also at the noiseless endpoint.
2026-09-25 v2
On Average Distance, Level-1 Fourier Weight, and Chang's Lemma
We study the maximum level-$1$ Fourier weight of Boolean functions, which is equivalent to the minimum average-distance problem on the hypercube. We first determine the dimension-free optimum at sufficiently small densities: there exists a universal $a_{0}>0$ such that the maximum level-$1$ Fourier weight for Boolean functions of mean smaller than $a_{0}$ is asymptotically attained by Hamming balls as the dimension $n\to \infty$. The key ingredient is an eventual Gaussian stop-loss domination inequality for normalized Rademacher sums. We then use an induction argument to improve the classical level-$1$ bound (Chang's lemma) in both the small- and large-density regimes. We apply these estimates to strengthen the Friedgut--Kalai--Naor theorem, study the corresponding average-distance problem in Euclidean space, and derive a sharp form of Chang's original lemma for $\mathbb{F}_{2}^{n}$: Hamming balls maximize the dimension of the span of the large Fourier coefficients.
2026-09-21
Dictators are most informative
We prove the Courtade-Kumar conjecture: among all Boolean functions $f\colon \{-1,1\}^n\to\{-1,1\}$, a dictator retains the most information about a uniformly random input observed through independent binary noise.
2026-09-20
Sensitivity and Block Sensitivity of Elementary Symmetric Boolean Functions of Arbitrary Degree
Let $σ_{n,d}$ denote the elementary symmetric Boolean function of $n$ variables and degree $d$. We completely determine the sensitivity, average sensitivity, and block sensitivity of $σ_{n,d}$ for every $1\le d\le n$. Using Lucas' theorem, we obtain a uniform binary description of the Hamming-weight value sequence, from which the sensitivity and average-sensitivity formulas follow and the computation of block sensitivity reduces to at most four explicit candidates. Combining these results with the arbitrary-degree formula for certificate complexity, we determine the exact relations among sensitivity, block sensitivity, and certificate complexity. We also prove a general result for symmetric Boolean functions: every nonconstant symmetric Boolean function $f$ satisfies \[ \bs(f)\le \max\{s(f),C(f)-1\}. \] Consequently, only the three patterns \[ s=\bs=C,\qquad s=\bs<C,\qquad s<\bs<C \] can occur for nonconstant symmetric Boolean functions. For elementary symmetric Boolean functions, we give necessary and sufficient conditions for each of these three patterns, thereby completely classifying the relations among $s(σ_{n,d})$, $\bs(σ_{n,d})$, and $C(σ_{n,d})$. In particular, we obtain a necessary and sufficient characterization of the full strict hierarchy \[ s(σ_{n,d})<\bs(σ_{n,d})<C(σ_{n,d}), \] and exhibit infinite families for which it holds.
2026-09-19
Counting and Covering in Nearest-Neighbour Representations of Boolean Functions
We study the number of prototypes needed to represent Boolean functions by nearest-neighbour classification. There are two distinct settings: the prototypes may be arbitrary points of Euclidean space, or they may themselves be required to lie in the Boolean cube. For unrestricted prototypes, we strengthen a known lower bound for almost all Boolean functions. The bound applies simultaneously to nearest-neighbour voting rules with any number of voting neighbours, and substantially narrows the gap with the known general upper bound. We obtain a VC-dimension bound for classes with a bounded number of prototypes, and show that it is sharp in order in dimensions at least four. We then study Boolean prototypes, beginning with symmetric threshold functions. A connection with covering designs expresses the minimum number of prototypes at every threshold level exactly in terms of a covering number, and leads to further exact results for related monotone functions, including disjunctive extensions and a characterisation of when a representation with a single negative prototype is possible. For a uniformly random Boolean function, the Boolean nearest-neighbour complexity, as a proportion of the cube, is asymptotically close either to one half or to one, with explicit limiting probabilities. In particular, almost every Boolean function requires at least approximately half as many prototypes as there are points in the cube, and one half is the largest proportion for which such a lower bound holds. Finally, we consider arbitrary symmetric Boolean functions. Their Boolean nearest-neighbour complexity is closely approximated by a weighted vertex-cover problem on paths. As a consequence, a uniformly random symmetric function typically requires prototypes amounting to $11/20$ of the cube. This is much larger than the upper bounds known when the prototypes are allowed to lie anywhere in Euclidean space.
2026-09-18
On the Fourier Entropy-Influence Conjecture for Boolean Plateaued Functions
We prove the following inequality for Boolean functions: $2\sum_{x\in F_2^n}f(x)wt(x)\geq wt(f)(n-deg(f))$. Using this inequality, we establish the Fourier Entropy-Influence (FEI) conjecture for Boolean plateaued functions. In particular, we show that the sharp FEI constant for the class of plateaued functions is 4. We also prove the FEI conjecture for partially bent functions and show that the corresponding sharp constant is 2. Finally, we derive several estimates for the p-biased distribution on the Boolean hypercube. Keywords: Fourier entropy, total influence, average sensitivity, plateaued function, algebraic degree, Reed-Muller code, p-biased distribution.
2026-09-16
A proof of Chvátal's conjecture via a sharp correlation inequality
We prove Chvátal's conjecture, posed in 1972: every hereditary family of subsets of a finite set has a largest intersecting subfamily that is a star. More generally, we prove a sharp correlation inequality for increasing Boolean functions $f,g:\{0,1\}^n\to\{0,1\}$. Writing $g^*(x)=1-g(1-x)$, we show that $$ \sum_{\varnothing\ne S\subseteq[n]}\hat{g}(S)^2\max_{i\in S}\mathrm{Inf}_i[f]\le\frac{2\mathrm{Cov}(f,g)\mathrm{Cov}(f,g^*)}{\mathrm{Cov}(f,g)+\mathrm{Cov}(f,g^*)}. $$ When $g$ is antipodal, that is, $g=g^*$, this yields $\mathrm{Cov}(f,g)\ge\frac{1}{4}\min_{i\in[n]}\mathrm{Inf}_i[f]$, the correlation formulation of Chvátal's conjecture due to Friedgut, Kahn, Kalai and Keller.
2026-09-15 v5
Bounds on the realizations of zero-nonzero patterns and sign conditions of polynomials restricted to varieties and applications
We obtain upper bounds, independent of the ambient dimension, for the number of realizable zero-nonzero patterns and (over ordered fields) sign conditions of a finite family of polynomials $\mathcal P$ restricted to an algebraic subset $V$ of affine or projective space. The bounds depend only on $\mathrm{card}(\mathcal P)$ and the degrees of the polynomials in $\mathcal P$, together with $\mathrm{deg}(V)$ and $\dim(V)$, and not on the dimension of the space in which $V$ is embedded. This feature is particularly useful when $V$ has small intrinsic dimension but is presented in a very high-dimensional ambient space. We describe several applications. First, we extend existing results on bounding the $\varepsilon$-entropy of real algebraic varieties. Second, we derive lower bounds (in terms of the number of connected components) for membership testing in semi-algebraic sets in the algebraic computation tree model. Finally, motivated by quantum complexity theory, we introduce additive and multiplicative notions of \emph{relative rank} in finite-dimensional vector spaces and algebras with respect to a fixed algebraic subset, generalizing the classical notion of tensor rank. We prove a general lower bound on the maximum relative rank of finite subsets with respect to algebraic sets of bounded degree and dimension that is again independent of the ambient dimension. As an illustration, we obtain a quantum analog of Shannon's classical lower bound. We prove that if $Δ_n$ is an algebraic family of allowed gates with $$ \dim Δ_n=\operatorname{poly}(n), \qquad \log\mathrm{deg} Δ_n=\operatorname{poly}(n), $$ then some Boolean functions require $2^n/\operatorname{poly}(n)$ gates from $Δ_n$, even when individual gates in $Δ_n$ may be highly nonlocal.
2026-09-11
An upper bound on the number of relevant variables in a bounded degree Boolean function on the Hamming graph
In this work, we prove that any Boolean function of degree $d$ on $\mathbb{Z}_{q}^n$, $q\geq 3$, has at most $m_qq^d$ relevant variables, where $m_q=\frac{2q(q^2+4q+1)}{(q-1)^4}$. For $q\in \{3,4,5,6,7\}$, we improve this bound to $2.854\cdot 3^d$, $1.749\cdot 4^d$, $1.263\cdot 5^d$, $0.994\cdot 6^d$, and $0.814\cdot 7^d$, respectively.
2026-09-11 v2
Perfect Combinatorial Structures in Coding Theory and Cryptography
This book develops algebraic and combinatorial methods for studying discrete structures. It brings together graph theory, Boolean functions, Fourier analysis on finite groups, coding theory, perfect colorings and perfect codes, association schemes, Latin squares, and related topics. A central theme is the interaction between different representations of the same object: combinatorial, algebraic, spectral, and coding-theoretic. The main mathematical object studied in this book is a perfect coloring of a graph, or, equivalently, an equitable partition of a graph. The book is intended for advanced undergraduate and graduate students in mathematics and computer science, as well as for researchers in discrete mathematics, combinatorics, coding theory, and related fields.
2026-09-06
Exponential Sampling Lower Bounds for Polynomial Sources
A degree-$d$ polynomial source is the output of a polynomial map of degree at most $d$ over $\mathbb{F}_2$ on arbitrarily many uniform random bits. Khodabandeh and Shinkar (FOCS '26) proved that $\mathrm{Ber}(1/3)^{\otimes N}$ has statistical distance $1-o(1)$ from every constant-degree polynomial source and conjectured exponentially small overlap. Independently of Khodabandeh and Shinkar, Byramji, Kane, Morris, and Ostuni (RANDOM '26) asked for an explicit target distribution at distance $1-\exp(-N^{Ω_d(1)})$. We resolve both questions. For every fixed $d\geq1$, every degree-$d$ polynomial source has overlap at most $\exp(-c_dN)$ with $\mathrm{Ber}(1/3)^{\otimes N}$, where $c_d>0$ is independent of the seed length. For quadratics, $c_2=2^{-26}$ suffices. We amplify Khodabandeh and Shinkar's uniform separation of acceptance probabilities from non-dyadic parameters (numbers not of the form $a/2^b$ for integers $a$ and $b\geq0$). The result extends to other non-dyadic Bernoulli parameters and to coordinates that are Boolean functions of boundedly many bounded-degree polynomials. We also give a uniform deterministic hierarchy between adjacent degrees. Appending the outputs of disjoint AND gates on $d+1$ inputs to uniform seed bits yields flat degree-$(d+1)$ target distributions of entropy $k$ with overlap $\exp(-Ω_d(\min\{k,N-k\}))$ against every degree-$d$ source, for $\min\{k,N-k\}\geq2(d+1)$. This entropy dependence is optimal up to constants in the exponent among flat target distributions for fixed $d$. The construction has locality $d+1$ and uses $O(N)$ field operations to sample. At $k=\lfloor N/2\rfloor$, it handles $d\leq(1-\varepsilon)\log_2N/3$ with overlap $\exp(-N^{\varepsilon-o(1)})$ for fixed $0<\varepsilon<1$. The proof combines monotonicity of Gowers uniformity norms, pairwise independence of points in a random affine cube, and relative entropy.
2026-08-30
A Sharp Small-Coefficient Variant of Khintchine's Inequality and the Sharp $π/2$ Theorem
We prove a refined quadratic normal approximation for the first absolute moment of normalized weighted Rademacher sums with bounded maximal coefficients. For any weight vector $w\in\mathbb{R}^{n}$ satisfying $\|w\|_{2}=1$ and $\|w\|_{\infty}\leqβ$ with sufficiently small $β>0$, we establish the uniform error bound $|\mathbb{E}|\sum_{i=1}^{n}w_{i}X_{i}|-\sqrt{2/π}|=O(β^{2})$ over all admissible weight configurations. Our proof combines zero-bias Stein's method and refined small-ball probability estimates to exploit symmetry cancellation and control the non-smooth residual of the absolute-value test function. An explicit extremal construction further verifies the optimality of this quadratic convergence rate. As an application, we establish an asymptotically sharp refinement of the Friedgut--Kalai--Naor (FKN) theorem for Boolean functions, also known as the sharp $π/2$ theorem, characterizing the level-1 Fourier energy for functions deviating far from dictatorships.
2026-08-25
Forbidden stars in multidimensional $0$-$1$ matrices and visibility of lattice points
A $d$-dimensional $0$-$1$ matrix $M$ of size $n_1\times n_2\times \dots \times n_d$ can be considered as a Boolean function $M: B(n_1\times n_2\times \dots \times n_d) \to \{ 0,1\}$, where $B$ is the $d$-dimensional box of lattice points $(x_1, \dots, , x_d)\in Z^d$ with $0\leq x_i \leq n_i-1$, $1\leq i\leq d$. The $0$-$1$ matrix $M$ can also be described as a subset $P:=P(M)$ of $B$ such that $x\in P$ if and only if $M(x)=1$. A $k$-star with center $p$ in $M$ corresponds to a $(k+1)$-element subset $\{ p, p_1, \dots , p_k\} \subset B$ such that $p$ and $p_i$ differ only in one coordinate (for all $1\leq i\leq k$) and these $k$ coordinates are distinct. Here we consider the problem of determining the maximum number of $1$-entries of a $0$-$1$ matrix $M$ of dimension $d$ and size $n\times n \times \dots \times n$ that avoids all $k$-stars. Our main results are the asymptotical solution of the problem for every $d$ and $k$ (as $n\to \infty$), very close bounds for $k=d$, and the exact solution of the $d=k=3$ case. This problem has connections to several other areas of discrete mathematics, including $k$-partite hypergraphs, independent set problems, dominating set problems and covering codes. One of our tools (concerning maximal packings of induced copies of a given hypergraph) might have independent interest.
2026-08-24
Resolving a conjecture on quadratic APN functions and a new quadratic $(n,n)$-function associated to crooked functions
We say an $(n,n)$-function $F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ is a crooked function if for any nonzero $a \in \mathbb{F}_2^n$, the image of $D_aF(x)=F(x)+F(x+a)$ is an affine hyperplane. The only known examples of crooked functions are all quadratic almost perfect nonlinear (APN), or equivalently, for every known crooked function, $D_aF$ is affine for all $a \in \mathbb{F}_2^n$. The ortho-derivative $π_F \colon\mathbb{F}_2^n \to \mathbb{F}_2^n$ of a crooked function $F$ is the function such that $π_F(0)=0$, and for any nonzero $a$, the set $\{0,π_F(a)\}^\perp$ is the underlying vector space of $\mathrm{Im}(D_aF)$. We prove that for $n \geq 4$ and a crooked function $F$, if $k$ is a non-negative integer such that $F$ has $2^k$ quadratic component functions, $π_F$ has at least $2^n-2^{n-k}$ nonzero components of algebraic degree $n-2$. In particular, we resolve Gorodilova's conjecture that every nonzero component of $π_F$ has algebraic degree $n-2$ when $F$ is quadratic APN. As a corollary, we prove that for any even $n \geq 4$, any crooked $(n,n)$-function with at least one quadratic component has at least $5$ semi-bent components. As a second main result, for $n \geq 4$, we associate to a crooked function $F$ a quadratic function $\varepsilon_F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ that satisfies a strong geometric-combinatorial condition regarding the sums of $F$ over $2$-dimensional linear subspaces. Furthermore, we obtain a congruence result on a problem on $m$-sequences introduced by Johansen, Helleseth, and Kholosha, and we determine the exact algebraic degrees of some Boolean functions associated to the bent and near-bent components of particular classes of plateaued vectorial functions.
2026-08-20
Quantitative bounds for regular $3$-wise intersecting families
Frankston, Kahn and Narayanan proved that every regular increasing $3$-wise intersecting family of subsets of $[n]$ has cardinality $o(2^n)$ using Friedgut's junta theorem. We give a short quantitative proof using elementary tools from the analysis of Boolean functions and entropy. More precisely, if $\mathcal{A}\subseteq\mathcal{P}_n$ is a nonempty $3$-wise intersecting family that is both regular and increasing, then $$ \log\frac{2^n}{|\mathcal{A}|}\ge \frac{n}{2}\left(\frac{|\mathcal{A}|}{2^n-|\mathcal{A}|}\right)^2, $$ and consequently $|\mathcal{A}|\le 2^n\sqrt{W(n)/n}$, where $W$ is the principal Lambert function defined by $W(x)e^{W(x)}=x$ for $x\ge0$. We also give a purely Fourier-analytic proof of the weaker estimate $$ |\mathcal{A}|\le \frac{2^n}{1+n^{1/3}}. $$
Finite-n Estimate of Dedekind Numbers by Layer-Ratio Monte Carlo
Published in Mathematics 14(18), 3300 (2026) • View Publication • BIB
Dedekind's problem counts monotone Boolean functions, equivalently downsets of a Boolean lattice. We recast this enumeration as a finite layer-ratio reconstruction problem for the Whitney numbers of the ranked ideal lattice. An exact adjacent-layer double count expresses each layer ratio through local averages of the number of addable elements and the number of removable elements. Reversible fixed-layer Markov chains estimate these averages and hence estimate the Dedekind number $M(n)$. Backtests at $M(8)$ and $M(9)$ calibrate seed-level variability under the fixed protocol and measure the observed Monte Carlo budget scaling. The resulting estimate probes the Whitney-number sequence of the ideal lattice. Although these rows have previously been described empirically as unimodal, the high-precision $n=9$ estimate has a shallow two-shoulder feature around the central rank, contrary to that empirical description; $n=11$ and $n=13$ center-window estimates show a larger-contrast analogous pattern. The protocol estimate for $M(10)$ is \[ \widehat M(10)=(8.9360\pm0.0010)\times 10^{78}, \] where the displayed uncertainty is the budget-based forecast scale from the cross-$n$ scaling law under the production budget.