cs.IT ↗ arXiv
154 papers in this category
Minimal additive codes and additive strong blocking sets
Additive codes over $\mathbb{F}_{q^h}$ generalize linear codes by relaxing linearity over the alphabet while retaining linearity over the subfield $\mathbb{F}_q$. In this paper, we introduce minimal additive codes and we initiate their study from a geometric perspective. We define the concept of additive strong blocking sets, a class of $h$-projective systems whose union forms a strong blocking set. We establish a one-to-one correspondence between equivalence classes of nondegenerate minimal additive codes and equivalence classes of additive strong blocking sets. We also compare this framework with the theory of outer strong blocking sets, showing that the latter arises as a special case. Finally, we provide constructions and existence results for minimal additive codes, and derive upper, lower, and asymptotic bounds on their minimum length.
Discrepancy for Random Linear Codes
We prove that random linear codes have nearly optimal discrepancy properties in a broad range of regimes. Our main results are two general theorems: one controlling all translates of a fixed test, and another controlling large families of Fourier-pseudorandom tests. Two motivating applications follow.
First, random linear codes match unstructured random codes for list-decoding from errors above capacity. If $C\subseteq\mathbb F_q^n$ is a random linear code of rate $1-\frac1n\log_q |B_ρ|+ε$, where $B_ρ$ is a radius-$ρ$ Hamming ball, then with high probability $$ |C\cap B|=(1\pm o(1))\frac{|C||B|}{q^n} $$ simultaneously for all radius-$ρ$ Hamming balls $B\subseteq\mathbb F_q^n$. This extends the classical result that such codes have covering radius at most $ρn$ whp (Blinovsky, 1987).
Second, over prime fields, random linear codes match unstructured random codes for zero-error list-recovery above capacity. For prime $q>2$ and $2\le \ell\le q-1$, a random linear code of rate $1-\log_q\ell+ε$ satisfies, with high probability, $$ |C\cap S|=(1\pm o(1))\frac{|C|\ell^n}{q^n} $$ simultaneously for all rectangles $S=S_1\times\cdots\times S_n$ with $|S_i|=\ell$. As a consequence, there are abundant $n$-party linear ramp secret sharing schemes over $\mathbb F_q$ with privacy threshold about $n/(2\log q)$ and reconstruction threshold about $5n/(2\log q)$, resilient to balanced local leakage; prior existence results required thresholds above $n/2$ even in this case.
The translate result, hence the list-decoding application, holds over arbitrary finite fields, even growing with $n$. The list-recovery and leakage applications hold over prime fields under moderate growth, e.g. $q\le n^{1/5-o(1)}$. The proofs use a refined second-moment analysis tracking intersection sizes as random generators are added to $C$.
An Erdős Matching Conjecture for Vector Spaces
We study a vector-space analogue of the Erdős Matching Conjecture. Let $m_q(n,k,s)$ denote the maximum cardinality of a family of $k$-dimensional subspaces of an $n$-dimensional vector space over $\mathbb F_q$ with no $s+1$ members whose sum is direct. Two natural constructions provide lower bounds. The first consists of all $k$-subspaces contained in a fixed $((s+1)k-1)$-dimensional subspace; the second consists of all $k$-subspaces that intersect a fixed $s$-dimensional subspace nontrivially. These constructions motivate the following vector-space analogue of the Erdős Matching Conjecture: for all $n\ge (s+1)k$, $$m_q(n,k,s)=\max\left\{\genfrac{[}{]}{0pt}{}{(s+1)k-1}{k}_q,~\genfrac{[}{]}{0pt}{}{n}{k}_q-q^{ks}\genfrac{[}{]}{0pt}{}{n-s}{k}_q\right\}.$$ We prove this conjecture when $k=2$, when $n=(s+1)k$, and when $n$ is sufficiently large. In particular, the case $k=2$ may be viewed as a vector-space analogue of the Erdős--Gallai theorem. In the large-$n$ range, we also prove a Hilton--Milner-type stability theorem, determining the largest nontrivial families with this property. Finally, we connect this problem with $t$-cover-free families in vector spaces and determine their extremal number up to a lower-order term, extending a recent result of Shan and Zhou for the special case $t=2$. The proofs combine Lovász's minimax theorem for matroid matchings, a high-dimensional Hoffman bound for uniform hypergraphs, and packing-design arguments in vector spaces.
An eigenvalue proof of Hegedüs's bound for codes with a single Hamming distance
We give a short, self-contained linear-algebra proof of a bound of Hegedüs [Australasian Journal of Combinatorics, 2026; arXiv:2409.07877]: if all pairwise Hamming distances in a family of subsets of $\{1,\ldots,n\}$ equal a fixed value $λ\ne(n+1)/2$, then the family has at most $n$ members. Our proof uses the same Gram matrix as in Hegedüs's argument, but reads its eigenvalues in place of its determinant, and keys off of a single fact about vectors of equal norm and equal pairwise inner product. That fact applies verbatim over an alphabet of size $q$, where it yields the bound $n(q-1)$ for $λ\ne\bigl((q-1)n+1\bigr)/q$ -- the corrected form of a conjecture of Hegedüs, recently established by Hu, Huang, and Yu [arXiv:2504.07036].
Optimal Small Set Expanders and Their Codes
A left-regular bipartite graph $G$ of degree $d$ is called a $(t,α)$-small-set-expander if every subset $X$ of left vertices of size at most $t$ has at least $α|X|$ neighbors. Such a graph is an optimal small-set expander if small subsets have as many neighbors as possible. We characterize optimal expanders combinatorially via girth and prove the existence of $s$-optimal expanders for every $s$. We also prove that $s$-optimality yields new "transfer" lower bounds on the number of neighbors of sets of size $h\geq s$. Finally, as an application, we discuss the use of optimal small-set expanders in building good codes for key exchange protocols in post-quantum cryptography.
Perfect Sphere Packing In The Boolean Space
Perfect sphere packing in the Boolean space is a fundamental and complex problem with significant implications for coding theory, cryptography, and discrete mathematics. The classical solution to the perfect sphere packing problem was provided by Hamming via his well-known perfect codes. However, a major limitation of the traditional Hamming metric is its strict applicability, as it allows perfect partitioning only for spaces with specific, highly constrained dimensions. To address this structural limitation, this article introduces a novel distance metric specifically designed for Boolean hypercubes. The proposed metric modifies the topological properties of the space, making it mathematically viable to partition a Boolean space of any arbitrary dimension into disjoint, perfect spheres. We rigorously define the algebraic properties of this new distance function and demonstrate its consistency across various dimensions. Furthermore, we explore the structural characteristics of the resulting packings. This approach bypasses the classical dimensional constraints of Hamming codes, potentially opening new avenues for designing error-correcting codes and cryptographic primitives in non-traditional dimensions.
A geometric approach to generalized covering radii of linear codes
Covering problems in coding theory are closely related to finite geometry through the interpretation of the columns of parity-check matrices as point sets in finite vector spaces. Motivated by the recent notion of generalized covering radii of linear codes introduced by Elimelech, Firer and Schwartz, we develop a geometric framework for these parameters. We introduce $(ρ,t)$-saturating sets and show that they are precisely the finite-geometric counterparts of linear codes whose $t$-th generalized covering radius is at most $ρ$. We study the structure of these sets and show that the extremal case $ρ=t$ coincides with the notion of $t$-strong blocking sets. Thus, $(ρ,t)$-saturating sets interpolate between classical saturating sets and strong blocking sets. We provide several equivalent formulations, including affine and dual Grassmannian criteria, derive lower bounds on their size, and give constructions from strong blocking sets, graphs and projective configurations.
Block Tensor Rank of Sum-Rank Metric Codes
Sum-rank codes provide a generalized framework for Hamming and rank-metric codes, with codewords represented as tuples of matrices and weight given by the sum of the block ranks. In this paper, we introduce and study a block-tensor-rank invariant for sum-rank metric codes. To each code, we associate its \emph{block tensor rank}: the smallest number of block-simple tensors, namely rank-one matrices supported inside single blocks, whose linear span contains the code. In general, determining the block tensor rank of a sum-rank code is challenging. Our main structural result shows that the block tensor rank decomposes additively across the blocks of the code, thereby reducing its computation to a tensor-rank problem on each block projection. Consequently, we derive two complementary lower bounds on the block tensor rank, referred to as the \emph{projection-wise bound} and the \emph{coordinate-code bound}. Moreover, by combining the coordinate-code bound with the classical Singleton and Griesmer bounds for codes in the Hamming metric, we obtain explicit lower bounds, called the \emph{Singleton coordinate-code bound} and the \emph{Griesmer coordinate-code bound}, respectively. We further construct families of sum-rank codes whose block tensor ranks attain the Singleton or Griesmer coordinate-code bounds. These constructions are based on Hamming-metric codes achieving the corresponding classical bounds. Finally, we show that, in certain cases, the block tensor ranks of two known families of sum-rank codes in the literature do not attain the Singleton coordinate-code bound.
Counting contiguous superregular $4 \times 4$ matrices
This short paper has two goals. First, explaining a simple procedure (which is essentially folklore) that, sometimes, makes it possible to obtain a formula for the number of solutions to a system of multivariate polynomial inequalities over a finite field. Second, applying that procedure to prove a formula for the number of contiguous superregular $4 \times 4$ matrices over a finite field. The formula was previously conjectured by Appuswamy, Bazzani, Connelly, Ekaireb, Congero, and Zeger [Probability of super-regular matrices and MDS codes over finite fields, arXiv:2603.20983]. In addition, the same procedure is used to provide formulas for the number of contiguous superregular $3 \times 4$, $3 \times 5$, and $3 \times 6$ matrices over a finite field.
Degree-Four Vector-Coordinate SoS Cannot Detect the MUB Upper Bound
We prove a degree-four Sum-of-Squares lower bound for the standard vector-coordinate formulations of mutually unbiased bases. For every dimension $d$ and every proposed number $m$ of bases, we construct a degree-four pseudoexpectation satisfying the orthonormality constraints and the cross-unbiasedness constraints in the quartic equality formulation. The construction is expectation over $m$ independent Haar-random orthonormal bases. We also prove that the same pseudoexpectation satisfies the degree-four localizing constraints for the natural $2\times 2$ Hermitian semidefinite formulation of the cross-coherence inequalities. Consequently, degree-four vector-coordinate SoS cannot refute the existence of $m$ mutually unbiased bases, even when $m>d+1$. In particular, under the two vector-coordinate encodings explicitly described in Randomstrasse101 Open Problem 23, degree-four SoS cannot prove that seven mutually unbiased bases do not exist in $\mathbb C^6$. We contrast this with a centered projector-coordinate Gram formulation, where degree-four SoS already recovers the elementary upper bound $m\le d+1$, giving a simple separation between vector-coordinate and projector-coordinate degree-four relaxations.
On perfect flag-rank metric codes
Flag-rank-metric codes arise as a natural generalization of rank-metric codes in the context of network communication. While recent research has mainly focused on algebraic and structural properties of these codes, the combinatorial geometry underlying the flag-rank metric remains largely unexplored. In this paper, we initiate a detailed investigation of this geometry. We explicitly determine the size of spheres of small flag-rank radius in the space $\mathrm{U}(n,\mathbb{F}_q)$ of upper triangular matrices over the finite field $\mathbb{F}_q$, and consequently obtain formulas for the size of balls of radius at most $3$. Using these enumerative results, we derive a sphere-packing bound for flag-rank-metric codes and introduce the notion of perfect codes with respect to the flag-rank metric. We observe that no non-trivial perfect flag-rank-metric codes exist in $\mathrm{U}(n,\mathbb{F}_q)$ for $n\in\{2,3\}$. We then investigate the possible parameters of perfect codes in higher dimensions. For minimum distance $3$, we obtain a characterization in terms of the codimension of the code, and show that suitable maximum flag-rank distance codes with minimum distance $3$ yield non-trivial perfect codes. For minimum distances $5$ and $7$, we derive explicit quadratic and cubic conditions, respectively, that any perfect code must satisfy. Finally, using asymptotic estimates for balls of fixed radius, we prove that for fixed length $n$ and $δ\in\{3,5,7,9,11\}$, perfect linear flag-rank-metric codes with minimum distance $δ$ do not exist over $\mathbb{F}_q$ for all sufficiently large $q$.
A $q$-analogue of the rational normal curve and linearized Reed-Solomon codes
The relationship between linear codes in the Hamming metric and projective algebraic varieties has led to deep interactions between coding theory and algebraic geometry, with classical examples such as Reed-Solomon codes and the rational normal curve. On the other hand, the sum-rank metric has recently gained attention due to applications in network coding, distributed storage, and post-quantum cryptography, with linearized Reed-Solomon codes emerging as optimal constructions. Despite recent advances, their structural and geometric properties are still not fully understood, and existing distinguishers remain limited. In this paper, we develop a geometric framework for linearized Reed-Solomon codes by considering a $q$-analogue of the rational normal curve. This yields a geometric characterization for certain parameter choices and reveals that the corresponding sets of points satisfy unexpectedly many $(q+1)$-degree hypersurface conditions. Our approach extends Schur-product-based techniques from the Hamming and rank-metric settings to the sum-rank metric case. Finally, we study the Hilbert function of the associated coordinate ring, providing a detailed description of its behavior and identifying its regularity, which also sheds new light on Gabidulin codes.
Quadratic APN Functions in Dimension 8 via Gröbner Basis Search in a Self-Equivalence Subspace
We describe a computational search for quadratic APN (Almost Perfect Nonlinear) functions in dimension 8 within a structured self-equivalence subspace. The search space is a 40-dimensional binary linear subspace consisting of all functions commuting with a linear automorphism of order 5 (class 22 in the taxonomy of Beierle, Brinkmann, and Leander, 2021), previously reported to contain no APN functions. Our approach combines random sampling via an explicit RREF parameterization (approximately 600 fresh APN-positive evaluations per core-hour) with Gröbner basis computation in Magma to enumerate all APN functions in a 24-dimensional hyperplane through each center (approximately 10 minutes per hyperplane). From 428 hyperplane computations, covering 0.65% of all 65,536 hyperplanes, we obtained 566 quadratic APN functions forming six CCZ-equivalence classes under the ortho-derivative invariant. Four classes, comprising 500 functions, match no entry in the 2025 database of 3,775,599 quadratic APN functions or in the pre-2020 compilation of 12,921 instances. Two classes (66 functions) are CCZ-equivalent to the Gold functions x^3 and x^9, confirming the correctness of the search pipeline. A membership analysis shows that the three new classes (B, C, D) lie entirely outside the original subspace and occur only in Gold-centered slices, demonstrating the essential role of the Gröbner basis stage. In 532 experiments using database functions as slice centers and 20 experiments with random centers, no APN neighbors were found, indicating that the gateway phenomenon is specific to the self-equivalence structure of the search space. Since the ortho-derivative invariant is a complete CCZ-invariant for quadratic APN functions, the absence of matching signatures provides a rigorous proof of CCZ-inequivalence.
Graphical Analysis of Lifted Product Code Constructions
Lifted product codes are an important family of quantum low-density parity-check (QLDPC) codes, as they were the first QLDPC code family shown to be asymptotically good. Understanding the structure of their parity-check matrices $H_{\mathsf{X}}$ and $H_{\mathsf{Z}}$, as well as the associated Tanner graphs, is essential for analyzing their decoding behavior and error-floor performance. In this work, we show that the Tanner graphs of $H_{\mathsf{X}}$ and $H_{\mathsf{Z}}$ are indeed isomorphic, and investigate their graph-theoretical structure. We establish conditions ensuring the connectivity of these graphs and provide bounds on their minimal absorbing sets, providing new insight into the combinatorial structures influencing decoding performance.
Handbook of Error-Correcting Codes
Barcode scans, clear phone calls, reliable data storage, satellite communication, and large-scale quantum computation are all made possible by error correction. We present a handbook version of The Error Correction Zoo, a curated reference of methods for protecting classical or quantum information from errors during storage and transmission. The handbook includes descriptions of these error-correcting codes and a classification according to the symbols they use. It also catalogues relations among codes and related objects such as sphere packings, lattices, designs, groups, and classical and quantum phases of matter. The collection is intended both as a rigorous reference and as a practical aid for tracing the web of code relationships and uncovering new connections.
Finite-n Estimate of Dedekind Numbers by Layer-Ratio Monte Carlo
Dedekind's problem counts monotone Boolean functions, equivalently downsets of a Boolean lattice. We recast this enumeration as a finite layer-ratio reconstruction problem for the Whitney numbers of the ranked ideal lattice. An exact adjacent-layer double count expresses each layer ratio through local averages of the number of addable elements and the number of removable elements. Reversible fixed-layer Markov chains estimate these averages and hence estimate the Dedekind number M(n). Backtests at M(8) and M(9) calibrate seed-level variability under the fixed protocol and measure the observed Monte Carlo budget scaling. The resulting estimate probes the Whitney-number sequence of the ideal lattice. Although these rows have previously been described empirically as unimodal, the high-precision n=9 estimate has a shallow two-shoulder feature around the central rank, contrary to that empirical description; n=11 and n=13 center-window estimates show a larger-contrast analogous pattern. The protocol estimate for M(10) is \[
\widehat M(10)=(8.9360\pm0.0010)\times 10^{78}, \] where the displayed uncertainty is the budget-based forecast scale from the cross-n scaling law under the production budget.
New Codes from Cyclic and Negacyclic Codes of Even Length over $\mathbb{Z}_4$
This paper uses theoretical results previously established in the literature to design search algorithms to find new linear codes over $\mathbb{Z}_4$ from cyclic and negacyclic codes of even length. As a result of these searches, we have found 2500 new cyclic codes and 730 negacyclic codes. These new codes exhibit improved parameters compared to previously known codes. Additionally, we have obtained binary quantum codes with good parameters from such $\mathbb{Z}_4$ codes.
Majorization and Gaussian-Mass Maximality for Construction-A Lattices from Binary Self-Dual Codes
Regev and Stephens-Davidowitz conjectured that the integer lattice maximizes Gaussian mass among integral lattices of a given rank. We prove this, including the equality case, for all unimodular Construction-A lattices arising from binary self-dual codes. The proof reduces the theta-series inequality to a sharp majorization statement for codes: if $C$ is a binary self-dual $[2k,k]$ code, then the half-weight distribution of $C$ is dominated in convex order by $\operatorname{Bin}(k,1/2)$, which is the corresponding distribution for the repetition-code model of $\mathbb{Z}^{2k}$. Indeed, after putting $C$ in systematic form $[I\mid A]$, self-duality gives $AA^T=I$ over $\mathbb{F}_2$, so for a uniformly random message $a$ the two weights $\operatorname{wt}(a)$ and $\operatorname{wt}(aA)$ have the same binomial law. The half-weight of the resulting codeword is their average, and Jensen's inequality then gives convex-order domination. Applied to the convex test functions that build the theta series, this yields a sum-of-squares formula for the Gaussian-mass gap; applied to hinge functions, it gives coefficientwise nonnegativity of the reduced gap polynomial.
Classification of independent sets in signed Johnson graphs and applications to kissing arrangements
Johnson graph are a family of graphs that play an important role in the theory of constant-weight codes, extremal combinatorics, and combinatorial geometry. We study signed analogues of classical Johnson graphs, denoted by $J_\pm(n,k)$, whose vertices are vectors of the form $\pm e_{i_1}\pm\cdots\pm e_{i_k}$, where two vertices are adjacent whenever their dot product equals $k-1$. We are particularly interested in maximum independent sets in the case $k=4$. An example of such an independent set in $J_\pm(n,4)$, which we call \emph{classical}, is obtained by lifting an arbitrary optimal $(n,4,4)$-code. Such independent sets naturally define kissing arrangements in ${\mathbb R}^n$.
We develop an algorithm that is practical for computing all maximum independent sets in $J_\pm(n,4)$ up to signed permutations for $n\le 12$, $n\ne 11$. In addition to obtaining complete lists, we provide structural characterizations of all types of maximum independent sets in these dimensions, excluding $n=5$ and $n=11$. Our most striking results concern the case $n=12$. We identify $1579$ non-isomorphic maximum independent sets in $J_\pm(12,4)$, all corresponding to non-isometric kissing arrangements of size $840$ in ${\mathbb R}^{12}$. Structurally, $1575$ of these independent sets arise from three different constructions, the rest are liftings of one of four $(12,4,4)$-codes. To our knowledge, this is the first dimension in which such a large diversity of potentially optimal kissing arrangements has been observed.
Beyond this finite range, we prove that for $n\equiv 2$ or $4 \pmod 6$, every maximum independent set arises from a Steiner quadruple system. We also obtain a characterization of the so-called \emph{nontrivially self-compatible} codes, namely optimal $(n,4,4)$-codes from which non-classical maximum independent sets can be constructed.
Upper Bounds on Multiple $b$-Burst Deletion-Correcting Codes
Motivated by their applications in DNA-based storage systems, codes capable of correcting consecutive deletions have attracted significant attention. An important class of such codes consists of those that can correct multiple consecutive deletion errors, commonly referred to as multiple $b$-burst deletion-correcting codes. In this paper, we investigate the fundamental limits of multiple $b$-burst deletion-correcting codes. Specifically, we first characterize several structural properties of the associated deletion balls. Then, leveraging these properties, we derive several upper bounds and a combinatorial lower bound on the maximum size of such codes. As a consequence, our bounds improve upon the previously known results for general parameter regimes and are shown to be asymptotically optimal for certain cases.