arXiv++ Combinatorics

Browse math.CO papers from arXiv

boolean function

322 papers tagged with this keyword
2026-02-14
The bipartite analogue of a classical spanning tree enumeration formula, Boolean functions, and their applications to counting odd spanning trees
Recently, Zheng and Wu defined the concept of odd spanning tree of a graph, meaning a spanning tree in which every vertex has odd degree. Similar to Cayley's formula, Feng, Chen and Wu counted the number of odd spanning trees in complete graphs via Prüfer code and the exponential generating function. In this note, we give a simple proof via a classical spanning tree enumeration formula and the Boolean function.We also generalize it to complete bipartite graphs.
2026-01-20
Bialgebraic structures on boolean functions
We study several bialgebraic structures on boolean functions, that is to say maps defined on the set of subsets of a finite set $X$, taking the value $0$ on $\emptyset$. Examples of boolean functions are given by the indicator function of the hyperedges of a given hypergraph, or the rank function of a matroid. We give the species of boolean functions a two-parameters family of products and a coproduct, and this defines a two-parameters family of twisted bialgebras. We then try to define a second coproduct on boolean functions, based on contractions, in order to obtain a double bialgebra. We show that this is not possible on the whole species of boolean functions, but that there exists a maximal subspecies where this is possible. This subspecies being rather mysterious, we introduce rigid boolean functions and show that this subspecies has indeed a second coproduct, as wished, and that it contains rank functions of matroids and indicator functions associated to hypergraphs. As a consequence, we obtain a unique polynomial invariant on rigid boolean functions, which is a generalization of the chromatic polynomial of graphs.
2025-12-05
Fourier Sparsity of Delta Functions and Matching Vector PIRs
In this paper we study a basic and natural question about Fourier analysis of Boolean functions, which has applications to the study of Matching Vector based Private Information Retrieval (PIR) schemes. For integers m and r, define a delta function on {0,1}^r to be a function f: Z_m^r -> C with f(0) = 1 and f(x) = 0 for all nonzero Boolean x. The basic question we study is how small the Fourier sparsity of a delta function can be; namely how sparse such an f can be in the Fourier basis? In addition to being intrinsically interesting and natural, such questions arise naturally when studying "S-decoding polynomials" for the known matching vector families. Finding S-decoding polynomials of reduced sparsity, which corresponds to finding delta functions with low Fourier sparsity, would improve the current best PIR schemes. We show nontrivial upper and lower bounds on the Fourier sparsity of delta functions. Our proofs are elementary and clean. These results imply limitations on improving Matching Vector PIR schemes simply by finding better S-decoding polynomials. In particular, there are no S-decoding polynomials that can make Matching Vector PIRs based on the known matching vector families achieve polylogarithmic communication with a constant number of servers. Many interesting questions remain open.
2025-11-11 v2
A Lower Bound for the Fourier Entropy of Boolean Functions on the Biased Hypercube
We study Boolean functions on the $p$-biased hypercube $(\{0,1\}^n,μ_p^n)$ through the lens of Fourier/spectral entropy, i.e., the Shannon entropy of the squared Fourier coefficients. Motivated by recent progress on upper bounds toward the Fourier-Entropy-Influence (FEI) conjecture, we prove a complementary lower bound in terms of squared influences: for every $f:(\{0,1\}^n,μ_p^n)\to \{-1,1\}$ we have $$ {\rm Ent}_p(f)\ge 4p(1-p)(2p-1)^2\cdot\sum_{k=1}^n{\rm Inf}^{(p)}_k(f)^2.$$
2025-10-30
On the number of non-degenerate canalizing Boolean functions
Canalization is a key organizing principle in complex systems, particularly in gene regulatory networks. It describes how certain input variables exert dominant control over a function's output, thereby imposing hierarchical structure and conferring robustness to perturbations. Degeneracy, in contrast, captures redundancy among input variables and reflects the complete dominance of some variables by others. Both properties influence the stability and dynamics of discrete dynamical systems, yet their combinatorial underpinnings remain incompletely understood. Here, we derive recursive formulas for counting Boolean functions with prescribed numbers of essential variables and given canalizing properties. In particular, we determine the number of non-degenerate canalizing Boolean functions -- that is, functions for which all variables are essential and at least one variable is canalizing. Our approach extends earlier enumeration results on canalizing and nested canalizing functions. It provides a rigorous foundation for quantifying how frequently canalization occurs among random Boolean functions and for assessing its pronounced over-representation in biological network models, where it contributes to both robustness and to the emergence of distinct regulatory roles.
2025-10-25
Talagrand-Type Correlation Inequalities for Supermodular and Submodular Functions on the Hypercube
Talagrand initiated a quantitative program by lower-bounding the correlation of any two increasing Boolean functions in terms of their influences, thereby capturing how strongly the functions depend on the exact coordinates. We strengthen this line of results by proving Talagrand-type correlation lower bounds that hold whenever the increasing functions additionally satisfy super/submodularity. In particular, under super/submodularity, we establish the ``dream inequality'' $$\mathbb{E}[fg]-\mathbb{E}[f]\mathbb{E}[g]\ge \frac{1}{4}\cdot\sum\limits_{i=1}^n\mathrm{Inf}_i[f]\mathrm{Inf}_i[g].$$ Thereby confirming a conjectural direction suggested by Kalai--Keller--Mossel. Our results also clarify the connection to the antipodal strengthening considered by Friedgut, Kahn, Kalai, and Keller, who showed that a famous Chvátal's conjecture is equivalent to a certain reinforcement of Talagrand-type correlation inequality when one function is antipodal. Thus, our inequality verifies the Friedgut--Kahn--Kalai--Keller conjectural bound in this structured regime (super/submodular). Our approach uses two complementary methods: (1) a semigroup proof based on a new heat-semigroup representation via second-order discrete derivatives, and (2) an induction proof that avoids semigroup argument entirely.
2025-10-22
Counterexample to majority optimality in NICD with erasures
We asked GPT-5 Pro to look for counterexamples among a public list of open problems (the Simons ``Real Analysis in Computer Science'' collection). After several numerical experiments, it suggested a counterexample for the Non-Interactive Correlation Distillation (NICD) with erasures question: namely, a Boolean function on 5 bits that achieves a strictly larger value of $\mathbb{E}|f(z)|$ than the 5-bit majority function when the erasure parameter is $p=0.40.$ In this very short note we record the finding, state the problem precisely, give the explicit function, and verify the computation step by step by hand so that it can be checked without a computer. In addition, we show that for each fixed odd $n$ the majority is optimal (among unbiased Boolean functions) in a neighborhood of $p=0$. We view this as a little spark of an AI contribution in Theoretical Computer Science: while modern Large Language Models (LLMs) often assist with literature and numerics, here a concrete finite counterexample emerged.
2025-10-15 v2
VC-Dimension vs Degree: An Uncertainty Principle for Boolean Functions
In this paper, we uncover a new uncertainty principle that governs the complexity of Boolean functions. This principle manifests as a fundamental trade-off between two central measures of complexity: a combinatorial complexity of its supported set, captured by its Vapnik-Chervonenkis dimension ($\mathrm{VC}(f)$), and its algebraic structure, captured by its polynomial degree over various fields. We establish two primary inequalities that formalize this trade-off: $\mathrm{VC}(f)+\mathrm{deg}(f)\ge n,$ and $\mathrm{VC}(f)+\mathrm{deg}_{\mathbb{F}_2}(f)\ge n$. In particular, these results recover the classical uncertainty principle on the discrete hypercube, as well as the Sziklai--Weiner's bound in the case of $\mathbb{F}_2$.
The paired construction for Boolean functions on the slice
Let $V$ be a finite set of size $n$. We consider real functions on the "slice" $\binom{V}{k}$, which are also known as functions in the Johnson scheme. For $I \subseteq J \subseteq V$, the characteristic function of the set of all $K\in\binom{V}{k}$ with $I \subseteq K \subseteq J$ is called "basic". In this article, we investigate a construction arising as the sum of two "opposite" basic functions. In essentially all cases, these "paired" functions are Boolean. Our main result is the determination of the exact degree -- regarding a representation by an $n$-variable polynomial -- of all paired functions. The proof is elementary and does not involve any spectral methods. First, we settle the middle layer case $n=2k$ by identifying and combining various relations among the degrees involved. Then the general case is reduced to the middle layer situation by means of derived, reduced, and dual functions. Remarkably, in certain situations, the degree is strictly smaller than what is guaranteed by the elementary upper bound for the sum of functions. This makes paired functions good candidates for fixed-degree Boolean functions of small support size. As it turns out, for $n = 2k$ and even degree $t \notin \{0,k\}$, paired functions provide the smallest known non-zero Boolean functions, surpassing the $t$-pencils, which is the smallest known construction in all other cases.
2025-08-20 v3
A lower bound on the number of bent squares
Bent functions are Boolean functions that are maximally nonlinear. They can be represented as bent squares, i.e., square matrices for which each row and each column is the Walsh spectrum of a Boolean function. Using this representation, it is shown in this note that the number of bent functions in $n$ variables is at least $2^{n \cdot 2^{\frac{n}{2}} \left(1 + O\left(\frac{1}{n}\right)\right)}$ for even integers $n$.
Almost Maiorana-McFarland bent functions
In this article, we study bent functions on $\mathbb{F}_2^{2m}$ of the form $f(x,y) = x \cdot φ(y) + h(y)$, where $x \in \mathbb{F}_2^{m-1} $ and $ y \in \mathbb{F}_2^{m+1}$, which form the generalized Maiorana-McFarland class (denoted by ${GMM}_{m+1}$) and are referred to as almost Maiorana-McFarland bent functions. We provide a complete characterization of the bent property for such functions and determine their duals. Specifically, we show that $f$ is bent if and only if the mapping $φ$ partitions $ \mathbb{F}_2^{m+1}$ into 2-dimensional affine subspaces, on each of which the function $ h $ has odd weight. We investigate which properties of mappings $φ\colon \mathbb{F}_2^{m+1} \to \mathbb{F}_2^{m-1}$ lead to bent functions of the form $ f(x,y) = x \cdot φ(y) + h(y) $ both inside and outside ${M}^\# $ and provide construction methods for suitable Boolean functions $ h $ on $\mathbb{F}_2^{m+1}$. We present a simple algorithm for constructing partitions of the vector space $\mathbb{F}_2^{m+1}$ together with appropriate Boolean functions $ h $ that generate bent functions outside ${M}^\# $. When $ 2m = 8 $, we explicitly identify many such partitions that produce at least $ 2^{78} $ distinct bent functions on $\mathbb{F}_2^8$ that do not belong to ${M}^\# $, thereby generating more bent functions outside ${M}^\#$ than the total number of 8-variable bent functions in ${M}^\#$. Additionally, we demonstrate that concatenating four almost Maiorana-McFarland bent functions outside ${M}^\# $, can result in a bent function ${M}^\# $. This finding answers an open problem posed recently in Kudin et al. (IEEE Trans. Inf. Theory 71(5): 3999-4011, 2025). Conversely, using a similar approach to concatenate four functions each in ${M}^\#$, we generate bent functions that are provably outside ${M}^\#$.
2025-08-07 v3
NP-Hardness and ETH-Based Inapproximability of Communication Complexity via Relaxed Interlacing
We prove that computing the deterministic communication complexity D(f) of a Boolean function is NP-hard in the standard protocol-tree model, answering, independently and concurrently with Hirahara-Llango-Loff (arXiv:2507.10426), a question first posed by Yao (1979). Our reduction builds and expands on a suite of structural "interlacing" lemmas introduced by Mackenzie and Saffidine (arXiv:2411.19003); these lemmas can be reused as black boxes in future lower-bound constructions. The instances produced by our reduction admit optimal protocols for self-similar constructions with strong structural properties, giving a flexible framework for the design of reductions showing NP-hardness of deciding the communication complexity of a Boolean matrix. This complements the work by Hirahara, Ilango, and Loff, which establishes NP-hardness in the same model via a different route; our analysis additionally yields reusable structural guarantees and underpins further consequences concerning inapproximability. Because the gadgets in our construction are self-similar, they can be recursively embedded. We sketch how this yields, under the Exponential-Time Hypothesis, an additive inapproximability gap that grows without bound. Furthermore we outline a route toward NP-hardness of approximating D(f) within a fixed constant additive error. Full details of the ETH-based inapproximability results will appear in a future version. Beyond settling the complexity of deterministic communication complexity itself, the modular framework we develop opens the door to a wider class of reductions and, we believe, will prove useful in tackling other long-standing questions in communication complexity.
Millions of inequivalent quadratic APN functions in eight variables
The only known example of an almost perfect nonlinear (APN) permutation in even dimension was obtained by applying CCZ-equivalence to a specific quadratic APN function. Motivated by this result, there have been numerous recent attempts to construct new quadratic APN functions. Currently, 32,892 quadratic APN functions in dimension 8 are known and two recent conjectures address their possible total number. The first, proposed by Y. Yu and L. Perrin (Cryptogr. Commun. 14(6): 1359-1369, 2022), suggests that there are more than 50,000 such functions. The second, by A. Polujan and A. Pott (Proc. 7th Int. Workshop on Boolean Functions and Their Applications, 2022), argues that their number exceeds that of inequivalent quadratic (8,4)-bent functions, which is 92,515. We computationally construct 3,775,599 inequivalent quadratic APN functions in dimension 8 and estimate the total number to be about 6 million.
2025-08-02
Quantum Algorithms for Gowers Norm Estimation, Polynomial Testing, and Arithmetic Progression Counting over Finite Abelian Groups
We propose a family of quantum algorithms for estimating Gowers uniformity norms $ U^k $ over finite abelian groups and demonstrate their applications to testing polynomial structure and counting arithmetic progressions. Building on recent work for estimating the $ U^2 $-norm over $ \mathbb{F}_2^n $, we generalize the construction to arbitrary finite fields and abelian groups for higher values of $ k $. Our algorithms prepare quantum states encoding finite differences and apply Fourier sampling to estimate uniformity norms, enabling efficient detection of structural correlations. As a key application, we show that for certain degrees $ d = 4, 5, 6 $ and under appropriate conditions on the underlying field, there exist quasipolynomial-time quantum algorithms that distinguish whether a bounded function $ f(x) $ is a degree-$ d $ phase polynomial or far from any such structure. These algorithms leverage recent inverse theorems for Gowers norms, together with amplitude estimation, to reveal higher-order algebraic correlations. We also develop a quantum method for estimating the number of 3-term arithmetic progressions in Boolean functions $ f : \mathbb{F}_p^n \to \{0,1\} $, based on estimating the $ U^2 $-norm. Though not as query-efficient as Grover-based counting, our approach provides a structure-sensitive alternative aligned with additive combinatorics. Finally, we demonstrate that our techniques remain valid under certain quantum noise models, due to the shift-invariance of Gowers norms. This enables noise-resilient implementations within the NISQ regime and suggests that Gowers-norm-based quantum algorithms may serve as robust primitives for quantum property testing, learning, and pseudorandomness.
Some more constructions of $n-$cycle permutation polynomials
$ n-$cycle permutation polynomials with small n have the advantage that their compositional inverses are efficient in terms of implementation. These permutation polynomials have significant applications in cryptography and coding theory. In this article, we propose criteria for the construction of $ n-$cycle permutation using linearized polynomial $ L(x) $ for larger $ n $. Furthermore, we investigate and generalize certain novel forms of $ n-$cycle permutation polynomials. Finally, we demonstrate our approach by constructing explicit $ n-$cycle permutation of the form $ L(x)+γh(Tr_{q^{m}/q}(x)) $, and $ G(x)+γf(x) $ with a Boolean function $ f(x) $. The polynomial $ x^{d}+γf(x) $ with $ f(x) $ being a Boolean function is shown to be quadruple and quintuple permutation polynomials. Moreover, linear binomial triple-cycle permutation polynomials are constructed.
2025-05-19
A near-optimal Quadratic Goldreich-Levin algorithm
In this paper, we give a quadratic Goldreich-Levin algorithm that is close to optimal in the following ways. Given a bounded function $f$ on the Boolean hypercube $\mathbb{F}_2^n$ and any $\varepsilon>0$, the algorithm returns a quadratic polynomial $q: \mathbb{F}_2^n \to \mathbb{F}_2$ so that the correlation of $f$ with the function $(-1)^q$ is within an additive $\varepsilon$ of the maximum possible correlation with a quadratic phase function. The algorithm runs in $O_\varepsilon(n^3)$ time and makes $O_\varepsilon(n^2\log n)$ queries to $f$, which matches the information-theoretic lower bound of $Ω(n^2)$ queries up to a logarithmic factor. As a result, we obtain a number of corollaries: - A near-optimal self-corrector of quadratic Reed-Muller codes, which makes $O_\varepsilon(n^2\log n)$ queries to a Boolean function $f$ and returns a quadratic polynomial $q$ whose relative Hamming distance to $f$ is within $\varepsilon$ of the minimum distance. - An algorithmic polynomial inverse theorem for the order-3 Gowers uniformity norm. - An algorithm that makes a polynomial number of queries to a bounded function $f$ and decomposes $f$ as a sum of poly$(1/\varepsilon)$ quadratic phase functions and error terms of order $\varepsilon$. Our algorithm is obtained using ideas from recent work on quantum learning theory. Its construction deviates from previous approaches based on algorithmic proofs of the inverse theorem for the order-3 uniformity norm (and in particular does not rely on the recent resolution of the polynomial Freĭman-Ruzsa conjecture).
2025-04-06 v2
Clonoids of Boolean functions with a linear source clone and a semilattice or 0- or 1-separating target clone
Extending Sparks's theorem, we determine the cardinality of the lattice of $(C_1,C_2)$-clonoids of Boolean functions for certain pairs $(C_1,C_2)$ of clones of Boolean functions. Namely, when $C_1$ is a subclone (a proper subclone, resp.) of the clone of all linear (affine) functions and $C_2$ is a subclone of the clone generated by a semilattice operation and constants (a subclone of the clone of all $0$- or $1$-separating functions, resp.), then the lattice of $(C_1,C_2)$-clonoids is uncountable. Combining this fact with several earlier results, we obtain a complete classification of the cardinalities of the lattices of $(C_1,C_2)$-clonoids for all pairs $(C_1,C_2)$ of clones on $\{0,1\}$.
2025-03-05
On the minimum Hamming distance between vectorial Boolean and affine functions
In this paper, we study the Hamming distance between vectorial Boolean functions and affine functions. This parameter is known to be related to the non-linearity and differential uniformity of vectorial functions, while the calculation of it is in general difficult. In 2017, Liu, Mesnager and Chen conjectured an upper bound for this metric. We prove this bound for two classes of vectorial bent functions, obtained from finite quasigroups in characteristic two, and we improve the known bounds for two classes of monomial functions of differential uniformity two or four. For many of the known APN functions of dimension at most nine, we compute the exact distance to affine functions.
2025-02-15
Recursions for quadratic rotation symmetric functions weights
A Boolean function in $n$ variables is rotation symmetric (RS) if it is invariant under powers of $ρ(x_1, \ldots, x_n) = (x_2, \ldots, x_n, x_1)$. An RS function is called monomial rotation symmetric (MRS) if it is generated by applying powers of $ρ$ to a single monomial. The author showed in $2017$ that for any RS function $f_n$ in $n$ variables, the sequence of Hamming weights $wt(f_n)$ for all values of $n$ satisfies a linear recurrence with associated recursion polynomial given by the minimal polynomial of a {\em rules matrix}. Examples showed that the usual formula for the weights $wt(f_n)$ in terms of powers of the roots of the minimal polynomial always has simple coefficients. The conjecture that this is always true is the Easy Coefficients Conjecture (ECC). The present paper proves the ECC if the rules matrix satisfies a certain condition. Major applications include an enormous decrease in the amount of computation that is needed to determine the values of $wt(f_n)$ for a quadratic RS function $f_n$ if either $n$ or the order of the recursion for the weights is large, and a simpler way to determine the Dickson form of $f_n.$ The ECC also enables rapid computation of generating functions which give the values of $wt(f_n)$ as coefficients in a power series.
2025-02-09
On Shapley Values and Threshold Intervals
Let $f\colon \{0,1\}^n\to \{0,1\}$ be a monotone Boolean functions, let $ψ_k(f)$ denote the Shapley value of the $k$th variable and $b_k(f)$ denote the Banzhaf value (influence) of the $k$th variable. We prove that if we have $ψ_k(f) \le t$ for all $k$, then the threshold interval of $f$ has length $\displaystyle O \left(\frac {1}{\log (1/t)}\right)$. We also prove that if $f$ is balanced and $b_k(f) \le t$ for every $k$, then $\displaystyle \max_{k} ψ_k(f) \le O\left(\frac {\log \log (1/t)}{\log(1/t)}\right) $.