arXiv++ Combinatorics

Browse math.CO papers from arXiv

boolean function

322 papers tagged with this keyword
2026-09-07
Perfect Combinatorial Structures in Coding Theory and Cryptography
This book develops algebraic and combinatorial methods for studying discrete structures. It brings together graph theory, Boolean functions, Fourier analysis on finite groups, coding theory, perfect colorings and perfect codes, association schemes, Latin squares, and related topics. A central theme is the interaction between different representations of the same object: combinatorial, algebraic, spectral, and coding-theoretic. The main mathematical object studied in this book is a perfect coloring of a graph, or, equivalently, an equitable partition of a graph. The book is intended for advanced undergraduate and graduate students in mathematics and computer science, as well as for researchers in discrete mathematics, combinatorics, coding theory, and related fields.
2026-09-06
Exponential Sampling Lower Bounds for Polynomial Sources
A degree-$d$ polynomial source is the output of a polynomial map of degree at most $d$ over $\mathbb{F}_2$ on arbitrarily many uniform random bits. Khodabandeh and Shinkar (FOCS '26) proved that $\mathrm{Ber}(1/3)^{\otimes N}$ has statistical distance $1-o(1)$ from every constant-degree polynomial source and conjectured exponentially small overlap. Independently of Khodabandeh and Shinkar, Byramji, Kane, Morris, and Ostuni (RANDOM '26) asked for an explicit target distribution at distance $1-\exp(-N^{Ω_d(1)})$. We resolve both questions. For every fixed $d\geq1$, every degree-$d$ polynomial source has overlap at most $\exp(-c_dN)$ with $\mathrm{Ber}(1/3)^{\otimes N}$, where $c_d>0$ is independent of the seed length. For quadratics, $c_2=2^{-26}$ suffices. We amplify Khodabandeh and Shinkar's uniform separation of acceptance probabilities from non-dyadic parameters (numbers not of the form $a/2^b$ for integers $a$ and $b\geq0$). The result extends to other non-dyadic Bernoulli parameters and to coordinates that are Boolean functions of boundedly many bounded-degree polynomials. We also give a uniform deterministic hierarchy between adjacent degrees. Appending the outputs of disjoint AND gates on $d+1$ inputs to uniform seed bits yields flat degree-$(d+1)$ target distributions of entropy $k$ with overlap $\exp(-Ω_d(\min\{k,N-k\}))$ against every degree-$d$ source, for $\min\{k,N-k\}\geq2(d+1)$. This entropy dependence is optimal up to constants in the exponent among flat target distributions for fixed $d$. The construction has locality $d+1$ and uses $O(N)$ field operations to sample. At $k=\lfloor N/2\rfloor$, it handles $d\leq(1-\varepsilon)\log_2N/3$ with overlap $\exp(-N^{\varepsilon-o(1)})$ for fixed $0<\varepsilon<1$. The proof combines monotonicity of Gowers uniformity norms, pairwise independence of points in a random affine cube, and relative entropy.
2026-08-30
A Sharp Small-Coefficient Variant of Khintchine's Inequality and the Sharp $π/2$ Theorem
We prove a refined quadratic normal approximation for the first absolute moment of normalized weighted Rademacher sums with bounded maximal coefficients. For any weight vector $w\in\mathbb{R}^{n}$ satisfying $\|w\|_{2}=1$ and $\|w\|_{\infty}\leqβ$ with sufficiently small $β>0$, we establish the uniform error bound $|\mathbb{E}|\sum_{i=1}^{n}w_{i}X_{i}|-\sqrt{2/π}|=O(β^{2})$ over all admissible weight configurations. Our proof combines zero-bias Stein's method and refined small-ball probability estimates to exploit symmetry cancellation and control the non-smooth residual of the absolute-value test function. An explicit extremal construction further verifies the optimality of this quadratic convergence rate. As an application, we establish an asymptotically sharp refinement of the Friedgut--Kalai--Naor (FKN) theorem for Boolean functions, also known as the sharp $π/2$ theorem, characterizing the level-1 Fourier energy for functions deviating far from dictatorships.
2026-08-25
Forbidden stars in multidimensional $0$-$1$ matrices and visibility of lattice points
A $d$-dimensional $0$-$1$ matrix $M$ of size $n_1\times n_2\times \dots \times n_d$ can be considered as a Boolean function $M: B(n_1\times n_2\times \dots \times n_d) \to \{ 0,1\}$, where $B$ is the $d$-dimensional box of lattice points $(x_1, \dots, , x_d)\in Z^d$ with $0\leq x_i \leq n_i-1$, $1\leq i\leq d$. The $0$-$1$ matrix $M$ can also be described as a subset $P:=P(M)$ of $B$ such that $x\in P$ if and only if $M(x)=1$. A $k$-star with center $p$ in $M$ corresponds to a $(k+1)$-element subset $\{ p, p_1, \dots , p_k\} \subset B$ such that $p$ and $p_i$ differ only in one coordinate (for all $1\leq i\leq k$) and these $k$ coordinates are distinct. Here we consider the problem of determining the maximum number of $1$-entries of a $0$-$1$ matrix $M$ of dimension $d$ and size $n\times n \times \dots \times n$ that avoids all $k$-stars. Our main results are the asymptotical solution of the problem for every $d$ and $k$ (as $n\to \infty$), very close bounds for $k=d$, and the exact solution of the $d=k=3$ case. This problem has connections to several other areas of discrete mathematics, including $k$-partite hypergraphs, independent set problems, dominating set problems and covering codes. One of our tools (concerning maximal packings of induced copies of a given hypergraph) might have independent interest.
2026-08-24
Resolving a conjecture on quadratic APN functions and a new quadratic $(n,n)$-function associated to crooked functions
We say an $(n,n)$-function $F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ is a crooked function if for any nonzero $a \in \mathbb{F}_2^n$, the image of $D_aF(x)=F(x)+F(x+a)$ is an affine hyperplane. The only known examples of crooked functions are all quadratic almost perfect nonlinear (APN), or equivalently, for every known crooked function, $D_aF$ is affine for all $a \in \mathbb{F}_2^n$. The ortho-derivative $π_F \colon\mathbb{F}_2^n \to \mathbb{F}_2^n$ of a crooked function $F$ is the function such that $π_F(0)=0$, and for any nonzero $a$, the set $\{0,π_F(a)\}^\perp$ is the underlying vector space of $\mathrm{Im}(D_aF)$. We prove that for $n \geq 4$ and a crooked function $F$, if $k$ is a non-negative integer such that $F$ has $2^k$ quadratic component functions, $π_F$ has at least $2^n-2^{n-k}$ nonzero components of algebraic degree $n-2$. In particular, we resolve Gorodilova's conjecture that every nonzero component of $π_F$ has algebraic degree $n-2$ when $F$ is quadratic APN. As a corollary, we prove that for any even $n \geq 4$, any crooked $(n,n)$-function with at least one quadratic component has at least $5$ semi-bent components. As a second main result, for $n \geq 4$, we associate to a crooked function $F$ a quadratic function $\varepsilon_F \colon \mathbb{F}_2^n \to \mathbb{F}_2^n$ that satisfies a strong geometric-combinatorial condition regarding the sums of $F$ over $2$-dimensional linear subspaces. Furthermore, we obtain a congruence result on a problem on $m$-sequences introduced by Johansen, Helleseth, and Kholosha, and we determine the exact algebraic degrees of some Boolean functions associated to the bent and near-bent components of particular classes of plateaued vectorial functions.
2026-08-20
Quantitative bounds for regular $3$-wise intersecting families
Frankston, Kahn and Narayanan proved that every regular increasing $3$-wise intersecting family of subsets of $[n]$ has cardinality $o(2^n)$ using Friedgut's junta theorem. We give a short quantitative proof using elementary tools from the analysis of Boolean functions and entropy. More precisely, if $\mathcal{A}\subseteq\mathcal{P}_n$ is a nonempty $3$-wise intersecting family that is both regular and increasing, then $$ \log\frac{2^n}{|\mathcal{A}|}\ge \frac{n}{2}\left(\frac{|\mathcal{A}|}{2^n-|\mathcal{A}|}\right)^2, $$ and consequently $|\mathcal{A}|\le 2^n\sqrt{W(n)/n}$, where $W$ is the principal Lambert function defined by $W(x)e^{W(x)}=x$ for $x\ge0$. We also give a purely Fourier-analytic proof of the weaker estimate $$ |\mathcal{A}|\le \frac{2^n}{1+n^{1/3}}. $$
2026-08-01
Block Sensitivity can exceed Spectral Sensitivity Squared
The spectral sensitivity $λ(f)$ of a Boolean function is the largest eigenvalue of the adjacency matrix of its sensitivity graph. It lower-bounds every standard measure of query complexity, and Aaronson, Ben-David, Kothari, Rao and Tal, who introduced it, asked whether block sensitivity is at most quadratic in it: is $bs(f)=O(λ(f)^{2})$? We show that it is not. We construct a total Boolean function on $2017584$ variables with $bs(f)\ge 14011$ and $λ(f)\le 89.0162$, so that $bs(f)\geλ(f)^{2.127}$, and hence by composition a family with $λ(f_n)\to\infty$ and $bs(f_n)=Ω(λ(f_n)^{2.127})$. The function is the indicator of a union of $k$ subcubes indexed by the vertices of a doubly regular tournament, and the freedom left in the construction is fixed by the Lovász local lemma. The main result has been formally verified in Lean. We also give numerical evidence that a member of the same family on $1255$ variables reaches an exponent near $2.20$, and exhibit a member on $30$ variables whose exponent already exceeds $2$ and whose spectral sensitivity can be computed exactly.
2026-07-10
The complete cubic Walsh spectrum of a permutation-inverse Boolean family
Let $q=2^e$ with $e\ge2$ even, put $d=(q^2+q+1)/3$, and let $σ(X)=X+X^d+X^{dq}$ be the permutation of $\mathbb F_{q^2}$ introduced by Ding, Qu, Wang, Yuan, and Yuan. For $α\in\mathbb F_q^*$, define the Boolean function \[ f_α(x)=\operatorname{Tr}_{q^2}\bigl(α(σ^{-1}(x))^3\bigr), \qquad x\in\mathbb F_{q^2}. \] In this paper, we determine the complete Walsh distribution of $f_α$ in the remaining cubic case $α\in(\mathbb F_q^*)^3$. More precisely, these functions are not bent but are $2$-plateaued: their Walsh values are precisely $0$ and $\pm 2q$, with exact multiplicities. The main new tool is a completion method for the outside Walsh coefficients: the punctured Fourier transform arising from the outside reduction is filled on the missing line, a modification invisible to outside frequencies, and the completed function is then identified with a Boolean component of a Kasami APN monomial. The APN property supplies a fourth-moment identity which, together with the known subfield spectrum and a Hasse divisibility congruence, forces the pointwise cubic spectrum.
Statistical Estimation of higher Dedekind Numbers
We provide highly accurate estimations of the 10th through 15th Dedekind Numbers, to a precision of 4 digits for $D(10)$, to 2 digits for $D(15)$. These estimates were obtained using three methods, including pair matching on large quantities of 9-dimensional monotone Boolean functions for $D(10)$, Reference Subsets for $D(10)$, $D(11)$, and $D(12)$. And our best method "Weight Layer Branching" which provided accurate estimates for all $D(10)$ through $D(15)$, strongly improving on the previous best known estimates by Korshunov and Tian-Shun Chen et al. arXiv:2606.09795
2026-06-30
The sharp diagonal spectral correlation inequality on the discrete cube
We prove the sharp diagonal spectral correlation conjecture of Friedgut, Kahn, Kalai and Keller, proposed in their Fourier-analytic approach to Chvátal's conjecture. For every pair of increasing Boolean functions $f,g:\{0,1\}^n\to\{0,1\}$, $$\mathrm{Cov}(f,g)\ge4\sum_{\varnothing\ne S\subseteq[n]}|S|\hat{f}(S)^2\hat{g}(S)^2.$$ Thus covariance controls the degree-weighted collision of the two nonconstant Fourier spectra, giving a sharp Fourier strengthening of the Harris--Kleitman inequality. The theorem also implies the unweighted diagonal conjecture of Friedgut--Kahn--Kalai--Keller for an increasing family and a maximal intersecting family. The factor $4$ is optimal, and we determine all equality cases. Apart from pairs whose relevant coordinate sets are disjoint, equality occurs only for a common dictatorship and, up to relabelling coordinates and interchanging $f$ and $g$, for the two-coordinate AND-OR pair $(f,g)=(x_i x_j,\,x_i\vee x_j).$ The main novelty is a correlated four-restriction induction and a sharp endpoint convolution inequality. The usual two-restriction induction behind Harris--Kleitman sees only the parallel restricted pairs and loses the mixed Fourier information needed to control the degree-weighted diagonal spectral energy. We instead couple the four codimension-one restricted pairs with correlation $1/2$; this precise correlation extracts the missing degree-weighted energy as a nonnegative square.
2026-06-28
Toward a KKL Theorem for any HDX
The KKL Theorem, a seminal result in boolean function analysis, characterizes the structure of low-influence (non-expanding) functions on the hypercube. While recent years have seen breakthrough results across a variety of areas relying on analogs of the KKL Theorem beyond the cube (e.g., on product spaces, Grassmann graphs), further progress has been inhibited by our poor understanding of the phenomenon across more general domains. Motivated in this context, Bafna, Hopkins, Kaufman, and Lovett (STOC 2022) and Gur, Lifshitz, and Liu (STOC 2022) proved a generalized KKL-type Theorem for spectral high dimensional expanders (HDX). Their results, however, remain highly restricted due to strong quantitative expansion requirements on the underlying complex. In this work, we introduce a simple local-to-global method for analyzing low influence functions on simplicial complexes. Using this method we prove a local-to-global KKL-type Theorem: any simplicial complex whose links satisfy a KKL-Theorem also satisfies such a result globally. Building on Gotlib and Kaufman (RANDOM 2023), we also prove a weaker dimension-dependent KKL-type Theorem for simplicial complexes with any non-trivial (two-sided) expansion. As concrete applications of our framework, we give the first characterization of non-expanding functions on `combinatorial' HDX such as dense clique complexes and a corresponding Kruskal-Katona Theorem, as well as a small-set expansion theorem for the Ramanujan Complexes of Lubotzky, Samuels, and Vishne (EJC '05).
2026-06-08
A spectral correlation inequality for increasing Boolean functions
Talagrand's correlation inequality provides a quantitative strengthening of the Harris--Kleitman inequality for increasing Boolean functions. Motivated by a Fourier-analytic conjecture of Friedgut, Kahn, Kalai, and Keller, we prove that $$ \mathrm{Cov}(f,g)\ge 2\sum_{S\neq\emptyset}|S|\hat f(S)^2\hat g(S)^2 $$ holds for all increasing Boolean functions $f,g:\{0,1\}^n\to\{0,1\}$. The proof combines the reverse Bonami--Beckner inequality with Young's convolution inequality. We also establish a sharp pointwise inequality: for every $n\ge1$, every $0\leρ\le1$, and every $f,g:\{0,1\}^n\to[0,1]$, the optimal constant $c_{ρ,n}$ for which $$ \left\langle f,T_ρg \right\rangle\ge c_{ρ,n}\|f*g\|_2^2 $$ holds for all such $f,g$ is $1$ for $0\leρ\le1/2$, $(2(1-ρ))^n$ for $1/2<ρ<1$, and $0$ for $ρ=1$. Integrating this pointwise inequality yields, for $n\ge1$, the slightly improved bound $$ \mathrm{Cov}(f,g)\ge 4\cdot\frac{n+1}{2n}\sum_{S\neq\emptyset}|S|\hat f(S)^2\hat g(S)^2. $$
2026-06-08
Finite-n Estimate of Dedekind Numbers by Layer-Ratio Monte Carlo
Dedekind's problem counts monotone Boolean functions, equivalently downsets of a Boolean lattice. We recast this enumeration as a finite layer-ratio reconstruction problem for the Whitney numbers of the ranked ideal lattice. An exact adjacent-layer double count expresses each layer ratio through local averages of the number of addable elements and the number of removable elements. Reversible fixed-layer Markov chains estimate these averages and hence estimate the Dedekind number M(n). Backtests at M(8) and M(9) calibrate seed-level variability under the fixed protocol and measure the observed Monte Carlo budget scaling. The resulting estimate probes the Whitney-number sequence of the ideal lattice. Although these rows have previously been described empirically as unimodal, the high-precision n=9 estimate has a shallow two-shoulder feature around the central rank, contrary to that empirical description; n=11 and n=13 center-window estimates show a larger-contrast analogous pattern. The protocol estimate for M(10) is \[ \widehat M(10)=(8.9360\pm0.0010)\times 10^{78}, \] where the displayed uncertainty is the budget-based forecast scale from the cross-n scaling law under the production budget.
2026-06-04
A unified abstract regularity lemma
The goal of this short note is to prove a unified abstract regularity lemma which recovers Szemerédi's graph regularity lemma, Green's arithmetic regularity lemma, and a regularity lemma for Boolean functions as direct corollaries.
Low Soundness Linearity Testing on the Half-Slice
Let $f: T\to \{ 0,1 \}$ be a Boolean function on the Boolean half-slice, $T$, \ie elements of $\{0,1\}^n$ with Hamming weight $n/2$. We show that if $f(x)+f(y)=f(x+y)$ holds with probability $\frac{1+δ}{2}$ over a uniform pair $(x,y)$ such that $x,y,x+y\in T$, then $f$ agrees with some linear function on at least $\frac{1+δ}{2}-o(1)$ fraction of the points in $T$. More generally, we show that if $f$ passes the natural $k$-query BLR test with probability $\frac{1+δ}{2}$ for any $k\geq3$, then it must agree with some affine function at $\frac{1+δ^{\frac{1}{k-2}}}{2}-o(1)$ fraction of the points in $T$. The only other known linearity test for the slice in the low soundness regime (i.e., when $δ$ can be arbitrarily small) was given by Kalai, Lifshitz, Minzer, and Ziegler [FOCS'24]. Our result improves upon this result in two significant ways: firstly, it works for $k=3$ queries, instead of requiring $k\geq4$; secondly, our result is sharper, e.g., when $k=4$, we are able to conclude an agreement of $\frac{1+\sqrtδ}{2}-o(1)$ instead of $\frac{1+c\sqrtδ}{2}$ for $c\approx.0035$. In particular, our result matches (up to the $o(1)$ term) the conclusion one obtains over the full hypercube via the classical BLR analysis. Our main technical contribution is a new dense model theorem using bounds on Krawtchouk polynomials. Using these Krawtchouk polynomial bounds, we also obtain a simple $k$-query test ($k\geq 5$) that avoids any use of the dense model machinery. This simplified test naturally extends to the slice over the $q$-ary hypercube, giving the first such result over larger alphabets.
2026-05-21
Holographic functions and neural networks
A fuzzy Boolean function is a map $f:\cube^n\to [0,1]$, where $n\in\mathbb N$. We introduce and compare three ways of saying that such a function has bounded complexity. The first is a sampling property: the value $f(x)$ can be recovered, up to small error and with high probability, from the values of a bounded number of randomly chosen coordinates of $x$. We call this the holographic property. The second is a structural property: $f$ is uniformly close to a bounded-degree polynomial in boundedly many bounded linear coordinate forms. The third is computational: $f$ is uniformly close to the output of a neural network with a bounded number of non-input neurons, bounded Lipschitz activation functions and bounded incoming weights. We prove that these three properties are equivalent up to quantitative changes of the parameters. The implication from holography to polynomial structure uses a variant of a weak version of hypergraph regularity.
2026-05-02
The Banach-Butterfly Invariant: Influence-Adaptive Walsh Geometry for Ternary Polynomial Threshold Functions
We introduce the Banach-Butterfly Invariant (BBT), an influence-adaptive Banach geometry on the Walsh-Hadamard butterfly factorization. For a Boolean function $f:\{-1,+1\}^n\to\{-1,+1\}$ with coordinate influences $\mathrm{Inf}_\ell(f)$, BBT assigns exponent $p_\ell = 1+\mathrm{Inf}_\ell(f)$ to butterfly layer $\ell$, yielding the contraction invariant $μ(f)=\prod_\ell 2^{-\mathrm{Inf}_\ell/(1+\mathrm{Inf}_\ell)}$. We prove a Jensen lower bound $\log_2μ(f) \ge -I(f)/(1+I(f)/n)$ and that $μ$ is strictly Schur-convex in the influence vector (modulo permutation), giving scaling classes $μ\sim 2^{-n/2}$ (parity), $2^{-Θ(\sqrt{n})}$ (majority), $2^{-1/2}$ (dictators). $\log_2μ$ is rational but not polynomial in the Fourier coefficients while $μ$ is algebraic, and $μ$ separates functions with identical total influence (122 pairs at $n=3$). Using the certified $n \le 4$ ternary Walsh-threshold universe from a companion synthesis manuscript as a finite testbed, we compute exact MILP minimum-support certificates for all 65,536 Boolean functions at $n=4$ (mean 6.42, max 9, all-odd by a parity argument) and on 10,000 of the 616,126 NPN-canonical representatives we enumerate at $n=5$ (matching OEIS A000370). Conditional Spearman $ρ(μ,|\mathrm{supp}|)$ at fixed total influence is $+0.571$ in the largest stratum at $n=4$ but reverses to $-0.38$ at $n=5$ under both function-uniform and NPN-canonical sampling: $μ$ is a valid Schur-convex concentration invariant, not a universal monotone predictor of minimum support across $n$. A companion application paper validates a real-valued WHT activation-energy proxy inspired by this theory on five pretrained LLMs at W2A16, cutting wikitext-2 perplexity by 15-58% versus vanilla auto-round; the transfer from Boolean theory to the real-valued proxy is qualitative, not formal.
2026-03-30
Nonvanishing $k$-flats of Boolean and vectorial functions
$k$th-order sum-free functions are a natural generalization of APN functions using the concept of (non)vanishing flats. In this paper, we introduce a new combinatorial technique to study the nonvanishing flats of Boolean functions. This approach allows us to determine the number of nonvanishing flats for an infinite family of Boolean functions. We moreover use it to show that any $k$th-order sum-free $(n,n)$-function of algebraic degree $k$ gives rise to an $(n-k)$th-order sum-free $(n,n)$-function of algebraic degree $n-k$. This implies the existence of millions of $(n-2)$th-order sum-free functions.
2026-03-19
Improvement on the Erdős-Kleitman conjecture via the KKL theorem
In 1974, Erdős and Kleitman conjectured that if a family $\mathcal{F}\subseteq 2^{[n]}$ contains no matching of size \(s\) and is maximal with respect to this property, then $ |\mathcal{F}|\ge \left(1-2^{-(s-1)}\right)\cdot 2^{n}. $ For decades, the best general lower bound remained the trivial $2^{n-1}$. About a decade ago, Frankl and Tokushige emphasized that obtaining a bound of the form $\left(\frac{1}{2}+\varepsilon\right)\cdot 2^n$ for some $\varepsilon>0$ is a challenging problem. A breakthrough of Bucič, Letzter, Sudakov and Tran in 2018 showed that $ |\mathcal{F}|\ge \left(1-\frac{1}{s}\right)\cdot 2^n $ via two very elegant and quite different approaches. Our main result shows that $$ |\mathcal{F}|\ge \left( 1 - \frac{1}{s + (s-2)\frac{\log n}{2\sqrt{5}n}} \right)\cdot 2^n $$ by exploiting a connection to the cornerstone result of Kahn, Kalai and Linial on influences of Boolean functions. Independently, we can also obtain a weaker improvement combining the linear algebra method with a combinatorial twist.
2026-02-16
Multidimensional convolution matrices and perfect colorings of subspace hypergraphs applied for bent functions and related designs
The main aim of the present paper is to introduce new methods for the study of combinatorial designs related to bent functions. They are based on interpretations of convolution on finite abelian groups as multiplication by a multidimensional matrix and designs as perfect colorings of subspace hypergraphs of $\mathbb{F}_2^n$. We establish a correspondence between eigenfunctions of convolution matrices and perfect colorings of subspace hypergraphs, show that perfect colorings of subspace hypergraphs admit a characterization in terms of convolution and that two-valued eigenfunctions of subspace hypergraphs correspond to perfect colorings. As applications, we represent partial difference sets, bent and plateaued Boolean functions, spreads, and strong bent partitions of $\mathbb{F}_2^n$ as eigenfunctions of convolution matrices and as perfect colorings of subspace hypergraphs. We also find some eigenvalues of convolution matrices over $\mathbb{F}_2^n$ and $\mathbb{F}_3^n$.