- The paper introduces continuous-valued non-redundancy (CVNRD), a block-diagonal obstruction that characterizes sparsifier size for real-valued codes up to polynomial factors in ε⁻¹ and polylogarithmic factors in the dimension.
- The framework unifies Boolean, spectral, hypergraph, submodular, and valued-CSP sparsification, including nearly tight Õ(n²/ε⁴) bounds for bounded submodular sums and near-linear bounds for higher-power graph and hypergraph energies.
- The paper replaces dependence on codeword counts with Sauer–Shelah and covering arguments, enabling results for infinite codes while leaving open whether the randomized lower bound extends to every sparsification method in the unbounded-range setting.
Overview
This paper develops a structural theory of sparsification for real-valued codes, unifying combinatorial sparsification (cut and CSP sparsifiers) with continuous sparsification (spectral graph and hypergraph sparsification) under a single abstraction. A code C⊆R≥0m is an arbitrary collection of nonnegative vectors; a (1±ϵ)-sparsifier is a reweighted subset of coordinates preserving the weight of every codeword. The central contribution is a single parameter — continuous-valued non-redundancy (CVNRD), measuring the largest approximately block-diagonal obstruction inside a code — that governs sparsifiability up to polynomial factors in ϵ−1 and polylogarithmic factors in m (2607.16126).
The framework strictly generalizes prior work: Boolean codes C⊆{0,1}m recover the non-redundancy characterization of Brakensiek–Guruswami (STOC 2025); the spectral code c(u,v)(x)=(xu−xv)2 recovers Spielman–Teng spectral sparsification; and the t-spectral code c(u,v)(x)=∣xu−xv∣t yields new results for higher powers of graph energies.
The main classification theorem
The paper defines three notions of sparsifiability: unweighted sparsifiability US(C,ϵ) (worst-case over coordinate restrictions), weighted sparsifiability (1±ϵ)0, and random sparsifiability (1±ϵ)1, where a random sparsifier is any distribution over weight functions satisfying only coordinate-wise unbiasedness (1±ϵ)2. This model is deliberately broad: it permits arbitrary nonuniform, correlated sampling-and-reweighting schemes (the standard paradigm underlying most sparsification algorithms), excluding only deterministic constructions such as Batson–Spielman–Srivastava.
The main theorem states that, setting (1±ϵ)3,
(1±ϵ)4
and conversely,
(1±ϵ)5
A (1±ϵ)6 witness consists of disjoint coordinate blocks (1±ϵ)7 with subcodes (1±ϵ)8 such that each block is "completely shattered": for every subset (1±ϵ)9, some codeword realizes a low value CVNRD0 on CVNRD1 and a high value CVNRD2 on CVNRD3, while off-block entries are small both cumulatively and entrywise. This is the real-valued analogue of a diagonal submatrix: for Boolean codes, every such witness forces exact block diagonality, so CVNRD4 reduces to classical non-redundancy, recovering the qualitative content of Brakensiek–Guruswami's theorem as a corollary.
Two features distinguish this theorem from prior work. First, it pays no factor depending on CVNRD5 — essential because real-valued codes may be infinite (as in spectral sparsification), and even finite obstructions can require exponentially many codewords. The paper demonstrates this with the code CVNRD6: for CVNRD7, every sparsifier needs CVNRD8 coordinates, yet any finite witness requires CVNRD9 codewords, since replacing nonzero entries by ϵ−10 collapses the entire code to a single word. Second, the lower bound holds against all unbiased randomized schemes, not merely independent sampling.
Limitations of the general characterization
The lower bound in full generality is proved only for the randomized model rather than for all sparsifiers; whether ϵ−11 always holds is left open, though the authors conjecture the gap is an artifact of the proof. This gap disappears in bounded-aspect-ratio settings, where lower bounds apply to arbitrary sparsifiers.
Bounded-aspect-ratio codes and applications
For codes ϵ−12 with constant ϵ−13, the paper introduces ϵ−14, a cleaner obstruction requiring exact zeros off-block, and proves nearly tight bounds:
ϵ−15
Two consequences follow. For submodular sparsification, sums of ϵ−16-bounded submodular functions ϵ−17 admit ϵ−18-sparsifiers of size ϵ−19, essentially settling the complexity for bounded functions: the best previous upper bound was m0 (Kenneth–Krauthgamer), while the known m1 lower bound via directed cuts shows m2 is likely optimal. Extension to arbitrary nonnegative submodular functions remains open.
For valued CSPs, the paper defines m3 — a predicate-level analogue of non-redundancy with complete blocks over pairs of values in m4 — and proves that worst-case sparsifier size satisfies m5 for constant m6. Notably, complete blocks arise even for simple predicates: the predicate m7 has m8, whereas its Boolean support projection has non-redundancy only m9. This shows that magnitude structure, not merely zero/nonzero patterns, fundamentally governs VCSP sparsifiability.
Spectral consequences
Applying the general theorem to graphs, the paper proves that the C⊆{0,1}m0 of the spectral code of any C⊆{0,1}m1-vertex graph is C⊆{0,1}m2, via an elementary combinatorial argument: complete shattering forbids short cycles within each block (an odd cycle contradicts the telescoping sum C⊆{0,1}m3, and even cycles are handled by designating one edge at the low level), while girth–density tradeoffs force short cycles across blocks once the witness exceeds C⊆{0,1}m4 edges, violating the entrywise off-block condition. This yields:
- Higher-power graph spectra: for all C⊆{0,1}m5, sparsifiers of size C⊆{0,1}m6 preserving C⊆{0,1}m7, improving on the prior best C⊆{0,1}m8 of Jambulapati–Lee–Liu–Sidford.
- Hypergraph C⊆{0,1}m9-spectral sparsification: near-linear-size sparsifiers for c(u,v)(x)=(xu−xv)20 with no assumptions on hyperedge sizes; for c(u,v)(x)=(xu−xv)21 this gives the first chaining-free proof of spectral hypergraph sparsification, and for c(u,v)(x)=(xu−xv)22 the first near-linear-size sparsifiers of this kind.
The hypergraph reduction uses the graph-theoretic Sauer–Shelah lemma of Cesa-Bianchi and Haussler to fix consistent "representative pairs" witnessing high energy, converting hyperedges into graph edges at a cost of c(u,v)(x)=(xu−xv)23.
Technique: Sauer-Shelah as the common core
Methodologically, the paper replaces the toolkit of Gilmer's entropy method, matrix Chernoff concentration, and Talagrand chaining with the Sauer–Shelah lemma and its relatives. The Boolean argument proceeds by iteratively peeling dense subcodes: if removing a set c(u,v)(x)=(xu−xv)24 is necessary to shrink the code below c(u,v)(x)=(xu−xv)25 codewords, then repeatedly invoking Sauer–Shelah extracts disjoint complete blocks; subsampling reduces off-block weight; pigeonhole arguments align off-diagonal patterns; greedy processing makes the collection block upper-triangular and then truly block-diagonal; and a final Sauer–Shelah invocation converts each surviving block into a complete subcode, yielding an identity matrix of size c(u,v)(x)=(xu−xv)26 and hence c(u,v)(x)=(xu−xv)27. The continuous analogue replaces codeword counts with sizes of c(u,v)(x)=(xu−xv)28-covers, using the Alon–Ben-David–Cesa-Bianchi–Haussler bound relating cover size to fat-shattering dimension. For unbounded aspect ratios, the proof analyzes auxiliary codes restricted to narrow value ranges c(u,v)(x)=(xu−xv)29, since cover-size bounds and Chernoff concentration both degrade with the ratio t0 — a dependence strong enough that when t1, even single-codeword preservation fails under naive sampling.
Conclusion
The paper establishes that a single block-diagonal obstruction parameter controls sparsification across Boolean, discrete, bounded-aspect-ratio, and fully general real-valued codes, with matching upper bounds and randomized-scheme lower bounds, and derives improved or first-of-their-kind sparsifiers for higher-order spectral energies, hypergraphs, bounded submodular sums, and valued CSPs. The principal open questions left by the work are whether t2 lower-bounds all sparsifiers (not just unbiased randomized ones) in the unbounded-aspect-ratio regime, and whether the submodular bound extends beyond bounded-range functions to arbitrary nonnegative submodular functions.