arXiv++ Combinatorics

Browse math.CO papers from arXiv

cs.IT ↗ arXiv

213 papers in this category
2026-08-31
Shannon's problem on the monotonicity of entropy and a Conjecture of Tao
Let $X_1,X_2,\ldots$ be i.i.d. finitely supported random variables in a torsion-free abelian group, and write $S_k=X_1+\cdots+X_k$, and $H(S_k)$ is the Shannon entropy $S_k$, for all $k \ge 1$. We prove that, for every fixed $n\geq1$, \[ H(S_{n+1})-H(S_n) \geq \frac12\log\frac{n+1}{n} -o_{H(X_1)\to\infty}(1), \] uniformly over the ambient group and the input law. This proves a conjecture of Tao [29] in 2010.
2026-08-31
Strengthening Recursive Constructions for Zero-Error Shannon Capacity
The exact Shannon capacity is unknown for every odd cycle beyond the five-cycle $C_5$, making odd cycles a central open problem in zero-error information theory. Improving the known lower bounds requires constructing large independent sets in strong powers of these graphs. Recent AI-assisted work has produced a rapid sequence of improvements: building on the construction of Itty et al., Gao developed a recursive product construction for combining structured independent sets, and Buys, Polak, and Zuiddam (BPZ) subsequently strengthened this through a richer recursion framework. We continue this line of AI-assisted exploration and introduce a heterogeneous refinement of these constructions. The central observation is that the usefulness of an intermediate construction depends not only on the size of its current main independent set, but also on the auxiliary structure it carries into subsequent recursion. Consequently, different parts of that auxiliary structure need not use the same independent set, and different occurrences in a recursion need not use the same intermediate representation. We formalize this for Gao's binary product and derive explicit propagation rules showing how heterogeneous choices strengthen the resulting gadget while leaving its current code size unchanged, then extend the principle to the more general BPZ framework, tailoring constructions to the distinct roles they play within the recursion. Applying these refinements to the seven-cycle $C_7$, we obtain an independent set in $C_7^{\boxtimes 500}$ yielding $Θ(C_7)\ge 3.25883262\ldots$, improving the best known lower bound. Beyond the numerical gain, the results illustrate a general principle for recursive zero-error constructions: intermediate structures with the same dimension and current code size can have different downstream value depending on where and how they are used in the recursion.
Edge codes constructed from unicyclic graphs
Jaramillo-Velez recently introduced edge codes, a new class of toric evaluation codes constructed from the edges of a (hyper)graph $\mathcal{H}$. In the case that $\mathcal{H}$ is a tree, Jaramillo-Velez computed both the minimum distance and the weight distribution of the associated code. In this paper, we study edge codes associated to unicyclic graphs. Our most striking result is that computing the parameters of these codes is subtle in the case that the induced cycle has an even length because these values will depend on certain conditions regarding the length of the cycle and the size of the base field.
2026-08-31
Asymptotic Bounds on Generalized Covering Radii of Binary Primitive BCH Codes
Fix integers $e\ge2$ and $r\ge1$. In this paper we study the $r$-th generalized covering radius $ρ_r\left(BCH(e,m)\right)$ of the binary primitive $e$-error-correcting BCH code $BCH(e,m)$. By using an algebraic-geometric reformulation of the covering problem together with an explicit Lang-Weil estimate, we prove that \[ρ_r\bigl(\BCH(e,m)\bigr)\le(r+1)e-1\] for all sufficiently large $m$. For $e\ge7$, this improves a recent result of Belinsky--Zabokritskiy. Our proof gives a substantially simpler geometric approach to this upper bound. In particular it implies that \[ρ_2\bigl(BCH(e,m)\bigr)=3e-1\] for all sufficiently large $m$. Previously it was only known that \[ρ_2\bigl(\BCH(e,m)\bigr) \in \left\{3e-1,3e\right\}\] for all sufficiently large $m$.
Efficient Polynomial-Time Decoding of Simplicial Anticodes with Near-Optimal Performance
In this work, we propose an efficient decoding algorithm for codes arising from simplicial complexes, a family of binary linear codes for which no decoding method of this type was previously known. Although the algorithm does not always attain the maximum theoretical error-correcting capability, it provides an explicit bound that can be computed directly from the structure of the complex. Moreover, this bound is asymptotically optimal: the ratio between the guaranteed correcting capability and the theoretical maximum converges to $1$ as the code length increases, under natural assumptions on the dimension of the maximal faces. The correction capability is also presented in specific examples. Finally, we introduce specific families of simplicial complexes where the algorithm successfully reaches this theoretical bound.
Mutually orthogonal anti-Latin squares
Anti-Latin squares were introduced in connection with non-linear secure network coding, and the extremal problem for large mutually orthogonal families is motivated by that setting. We study the maximum size $N_A(d)$ of a family of mutually orthogonal anti-Latin squares of order $d$. We prove that $N_L(d)+1\le N_A(d)\le N_L(d)+2$ for every $d\ge 3$, where $N_L(d)$ denotes the classical maximum size of a family of mutually orthogonal Latin squares of order $d$, and we show that in fact $N_A(3)=N_L(3)+1$ whereas $N_A(d)=N_L(d)+2$ for every $d\ge 4$. The upper bound is obtained by passing through balanced matrices, while the lower bound is given by a deterministic permutation argument. For all $d\ge 8$, and also for the exceptional order $d=6$, the upper bound is shown to be attainable by a general probabilistic construction. On the structural side, we show that a saturated family of size $d+1$ induces an affine plane of order $d$, and that the saturated case is characterized by the existence of an anti-coordinate grid decomposition; after transporting this condition to the fixed cell set $[d]^2$, it becomes a direction-completeness condition on the corresponding row-blocks and column-blocks. The remaining small orders are treated separately: $d=3$ is handled by direct analysis and classification of orthogonal triples, $d=4$ by an explicit saturated construction and an analysis of its finite-geometric structure, and $d=5$ and $d=7$ by explicit saturated examples arising from the random-grid framework. Thus $N_A(d)$ is determined in terms of $N_L(d)$ for every $d\ge3$, and its numerical value is obtained explicitly for every $3\le d\le9$.
2026-08-28
Fine Difference Structure and Prime-Power Depth of Bent Partitions
A $p$-ary bent partition of $\mathbb{F}_p^n$ is a partition into $K$ nonempty cells such that every balanced assignment of its cells to $\mathbb{F}_p$ produces a bent function. It was asked whether every possible depth $K$ is a power of $p$; for general $p$, previous affirmative results required regularity or cell-symmetry hypotheses. We prove the stronger unconditional statement that, for every nonzero $h$, exactly $p^n/K$ points remain in the same fine cell under translation by $h$. Thus the fine cells form a partitioned difference family and the fine label map is zero-difference balanced. Consequently $K\mid p^n$, so $K=p^t$; nonempty cells further give $1\le t<n$. In even dimension, the classical cell-size theorem yields $K\mid p^{n/2}$. Together with the known odd-dimensional ternary three-fibre parameter restriction, this gives the global bound $t\le\lfloor n/2\rfloor$. The proof is an exact finite average over balanced coarsenings. The main counting identity and selected consequences are formalized and kernel-checked in Lean 4.
Decoding Algorithms for MDS Array Codes
We study decoding procedures for a family of MDS array codes previously constructed from the Kronecker product of a superregular matrix and a non-singular matrix over a finite field. By exploiting the particular structure of their parity-check matrices, we develop decoding algorithms for different channel models. For the erasure channel, we provide an algorithm capable of recovering any pattern of up to $n-k$ symbol erasures. For the $q$-ary symmetric channel, we investigate the decoding of one and two symbol errors and give explicit procedures for determining their locations and values. We also consider the particular case in which the superregular matrix is a Vandermonde matrix, showing how its additional algebraic structure can be exploited in the decoding process. Explicit examples over different finite fields are provided to illustrate the proposed procedures.
Weight Distributions of Single Parity-Check Product Codes via Character Sums
We investigate structural and enumerative properties of binary single parity-check product codes. For each $n\geq 2$, $\operatorname{SPC}(n)$ denotes the binary single parity-check code of length $n$, consisting of all binary vectors of length $n$ having even Hamming weight. We determine the generalized Hamming weight hierarchy of the product code $\mathcal{C}_{m,n}=\operatorname{SPC}(m)\otimes\operatorname{SPC}(n)$, whose codewords can be represented as $m\times n$ binary matrices in which every row and every column has even Hamming weight. For the square product $\mathcal{C}_n =\operatorname{SPC}(n)\otimes\operatorname{SPC}(n)$, we also determine the maximum codeword weight and prove that its homogeneous weight enumerator is symmetric if and only if $n$ is even. After characterizing the dual code, we apply the MacWilliams identity in its Walsh--Hadamard formulation to derive an exact closed-form expression for the weight enumerator. By grouping the auxiliary binary vectors according to their Hamming weights, we obtain an explicit formula for each coefficient in terms of binomial coefficients and alternating convolutions. Finally, using Krawtchouk polynomials, we present an exact procedure for computing the full weight distribution without exhaustively enumerating all codewords. Numerical examples illustrate the formulas and verify the resulting computations.
2026-08-25
A quaternionic construction behind $841$-point kissing arrangement in ${\mathbb R}^{12}$
Recently, a new record kissing arrangement of $841$ points in $\mathbb R^{12}$ was obtained numerically by optimization (Takhanov-Assylbekov-Yun, 2026). The configuration was released as a coordinate file, without a mathematical description of its structure. The purpose of this paper is to provide such a description. The key observation is that the geometry becomes transparent once we regard $\mathbb R^{12}\cong \mathbb H^3$ as the Cartesian product of three copies of the quaternion algebra. We first introduce a new $840$-point kissing arrangement with a certain quaternionic structure. It consists of three mutually orthogonal regular $24$-cells, supported on the three quaternionic coordinate factors $\mathbb H\times\{0\}\times\{0\}$, $\{0\}\times\mathbb H\times\{0\}$, $\{0\}\times\{0\}\times\mathbb H$, together with two $384$-point families obtained by lifting affine sets of the form $$\{(u,v,w)\in (\mathbb F_2^2)^3\mid u+v+w=η\},$$ to quaternionic triples (whose components belong to the binary octahedral group $2O$) and then applying suitable component-wise rotations and weightings. A characteristic feature of this construction is a pronounced asymmetry among the three quaternionic factors. For the $816$ vectors obtained after removing the third $24$-cell, most of the squared norm is concentrated in the first two quaternionic coordinates, while the third coordinate carries systematically less mass. Thus, the third four-dimensional factor contains more available space than the first two. We then show that this $840$-point configuration provides a natural structural model for the numerical $841$-point record. Finally, we introduce a notion of the general quaternionic construction in dimensions divisible by $4$, and check that record kissing arrangements in ${\mathbb R}^{4k}$, $k\leq 5$, admit a quaternionic construction.
2026-08-24
A Comment on Local Hypercube Inequalities
This comment gives dimension-independent analytic proofs of the local inequalities arising in the odd- and even-dimensional constructions of higher-dimensional partition charge functions. The prior works established these inequalities by exhaustive computation in low dimensions and tested them numerically in selected higher dimensions.
2026-08-24
Resolving a conjecture on quadratic APN functions and a new quadratic $(n,n)$-function associated to crooked functions
We say an $(n,n)$-function $F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ is a crooked function if for any nonzero $a \in \mathbb{F}_2^n$, the image of $D_aF(x)=F(x)+F(x+a)$ is an affine hyperplane. The only known examples of crooked functions are all quadratic almost perfect nonlinear (APN), or equivalently, for every known crooked function, $D_aF$ is affine for all $a \in \mathbb{F}_2^n$. The ortho-derivative $π_F \colon\mathbb{F}_2^n \to \mathbb{F}_2^n$ of a crooked function $F$ is the function such that $π_F(0)=0$, and for any nonzero $a$, the set $\{0,π_F(a)\}^\perp$ is the underlying vector space of $\mathrm{Im}(D_aF)$. We prove that for $n \geq 4$ and a crooked function $F$, if $k$ is a non-negative integer such that $F$ has $2^k$ quadratic component functions, $π_F$ has at least $2^n-2^{n-k}$ nonzero components of algebraic degree $n-2$. In particular, we resolve Gorodilova's conjecture that every nonzero component of $π_F$ has algebraic degree $n-2$ when $F$ is quadratic APN. As a corollary, we prove that for any even $n \geq 4$, any crooked $(n,n)$-function with at least one quadratic component has at least $5$ semi-bent components. As a second main result, for $n \geq 4$, we associate to a crooked function $F$ a quadratic function $\varepsilon_F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ that satisfies a strong geometric-combinatorial condition regarding the sums of $F$ over $2$-dimensional linear subspaces. Furthermore, we obtain a congruence result on a problem on $m$-sequences introduced by Johansen, Helleseth, and Kholosha, and we determine the exact algebraic degrees of some Boolean functions associated to the bent and near-bent components of particular classes of plateaued vectorial functions.
Entropy power inequalities in compact groups
Suppose $X,Y$ are independent random variables with values in a compact abelian group $(G,+)$. We examine the following two entropy power-type inequalities: $h(X+Y)\geq \frac{1}{2}h(X)+\frac{1}{2}h(Y)$ and $h(X+Y)\geq \max\{h(X),h(Y)\}$, where the entropy $h(Z)$ of a $G$-valued random variable $Z$ is defined in terms of its density with respect to Haar measure on $G$. For groups that are either connected or finite with no nontrivial subgroups, we precisely characterize the cases of equality and establish explicit, quantitative stability estimates in terms of relative entropy for these two inequalities. The main tools are a generalization of an entropic inequality obtained by Green, Manners and Tao (2023) for discrete entropy, and a harmonic-analytic estimate for the chi-squared contraction coefficient in connected compact groups. As an application, we derive exponential convergence rates to the uniform distribution in relative entropy for random walks on connected compact abelian groups.
2026-08-23
Average-Radius List-Decodability of Random Linear Codes
We prove that for every prime power $q$ and every $p \in (0, 1-1/q)$, a random $\mathbb{F}_q$-linear code of rate $1 - h_q(p) - ε$ is $(p, C_{p,q}/ε)$-average-radius list-decodable with probability at least $1 - q^{-Ω(n)}$, i.e., for every center $y \in \mathbb{F}_q^n$, the $C_{p,q}/ε$ codewords closest to $y$ have average fractional Hamming distance at least $p$ from $y$. This extends a similar result for (standard) list-decoding due to Guruswami, Håstad, and Kopparty (2010) to the stronger average-radius guarantee, with the same $O(1/ε)$ list size. For average-radius list-decoding, such a result was previously known only for binary linear codes (Guruswami, Li, Mosheiff, Resch, Silas, and Wootters, 2021) and for general (non-linear) random codes over arbitrary alphabets (Elias, 1991).
2026-08-20
Weak arcs and applications to the DNA-based storage access problem
Weak arcs are point sets in PG$(n-1,q)$ meeting every general hyperplane (those are the hyperplanes not going through one of the points given by the standard basis vectors) in at most $n-1$ points. In this paper, we study weak arcs together with balanced variants which are contained on the sides of the fundamental simplex. We give an upper bound on the size of weak arcs, characterise the largest balanced quasi-arcs in the plane and construct large balanced quasi-arcs in PG$(3, q)$. We then use these configurations to build point sets for the random-access problem in DNA-based storage. The constructions are explicit, work over small fields, and attain recovery expectations matching the best known asymptotic bounds.
2026-08-20
New upper bounds on covering codes K_q(n,R) for alphabets of size six and seven
We present improved upper bounds for nine entries of the standard tables of bounds on K_q(n,R), the minimum cardinality of a q-ary code of length n with covering radius R, for q in {6,7}: K_6(7,3)<=232, K_6(8,3)<=1045, K_6(8,4)<=167, K_6(9,4)<=703, K_6(9,5)<=123, K_6(10,4)<=2951, K_6(10,5)<=610, K_7(8,4)<=329, and K_7(9,4)<=1743. The previous best bounds, recorded in Keri's tables (last updated 2011), all arose from general constructions (direct sums and related product rules) rather than from explicit search; to our knowledge these are the first improvements to any upper bound on K_q(n,R) with q>=5 since 2011. The new bounds were found by focused local search seeded with the construction-based incumbents. All nine codes are given explicitly in the ancillary files, together with a standalone verifier; each code was checked by four independent exhaustive verification methods.
The Generalized Random Access Problem for Linear Codes
Random access is a central requirement in DNA-based storage systems: one would like to recover selected information symbols without sequencing the whole encoded object. A recent combinatorial model associates to a generator matrix $G\in F_q^{k\times n}$ the random variable $τ_i(G)$, measuring the number of sampled columns needed to recover the information vector $e_i$. We study the cardinality-based extremal and finite-geometric aspects of simultaneous multi-symbol recovery. For a nonempty set $I\subseteq[k]$, let $τ_I(G)$ denote the number of random column samples needed until all vectors $e_i$, $i\in I$, lie in the span of the observed columns. This variable interpolates between the singleton random access problem and the full-recovery problem underlying coverage depth. For each $m$, we introduce uniform worst-case and average parameters over all requested sets $I$ with $|I|=m$. Using the known subset-counting formula for $E[τ_I(G)]$, we establish general upper and lower bounds for these parameters. In particular, the lower bounds are expressed through order statistics of the singleton recovery variables and specialize to the known singleton bounds when $m=1$. For systematic MDS encoders, we record an equivalent form of the known multi-symbol expectation formula and derive monotonicity and asymptotic consequences. For simplex encoders in arbitrary dimension, we obtain closed formulae in terms of Gaussian binomial coefficients; the full-recovery endpoint agrees with the known coverage-depth formula for simplex codes. Finally, in dimension three we study balanced quasi-arcs and compare their values with the simplex and MDS benchmarks.
2026-08-18
Near-MDS codes of lengths q+6 and q+7 from conics in PG(2,q), q odd
Near maximum distance separable (NMDS) codes of dimension 3 and length n over the finite field with q elements are equivalent to (n,3)-arcs in PG(2,q). For every odd prime power q we construct, by adding five suitable points to a conic of PG(2,q), a family of [q+6,3,q+3] NMDS codes and determine their weight distributions completely; three distinct weight enumerators occur, governed by two explicit quadratic-character conditions on the parameters. For every odd prime power q, no code of our family is monomially equivalent to a code of the recent [q+6,3,q+3] NMDS family of Fan, Wang and Xu, even where the weight enumerators of the two families coincide: the separating invariant is a triple of geometric data attached to the underlying arc. Extending the configuration by a sixth point on a distinguished external line, we further obtain [q+7,3,q+4] NMDS codes, together with their weight distributions, for every odd prime power q >= 11 (admissible parameters exist for no q <= 9); the existence proof combines exact and Weil-type character sum estimates with a finite computer verification. Our proofs are purely geometric and rest on a simple counting identity for the trisecant lines of a point set obtained by extending a conic. All the codes constructed are optimal locally recoverable codes with locality 2.
2026-08-14
New lower bounds for constant-weight codes via seeded bit-swap tabu search
A binary constant-weight code is a set of binary words of length $n$ such that each word has exactly weight $w$ and is at least Hamming distance $d$ from every other word in the set. $A(n,d,w)$ denotes the maximum size of a binary constant-weight code with parameters $(n,d,w)$. Using seeded initialization with bit-swap tabu search, we found 124 new constructions that improve existing lower bounds for $A(n,d,w)$. As a corollary of stronger bounds on $A(n,8,8)$ for $n \in \{ 32,33,34,37 \}$, we also improve lower bounds on kissing numbers $τ_{32}$, $τ_{33}$, $τ_{34}$, and $τ_{37}$.
2026-08-12
New optimal linear codes over $\ZZ_4$
Published in Bulletin of the Australian Mathematical Society, 2023, 107(1), pp. 158-169 • Search Publication
In this work, we present novel approaches for constructing linear codes over $\ZZ_4$ from the known ones. We succeeded in obtaining new linear codes, many of which are optimal. In particular, we found all optimal codes for $k_1=2,~k_2=0$ and many optimal codes for $k_1=3,~k_2=0.$