- The paper presents a nonasymptotic minimax risk characterization for robust mean estimation under star-shaped constraints with heavy-tailed noise and adversarial contamination.
- It develops matching lower and upper bounds using information-theoretic and robust tournament arguments that hinge on the local metric entropy of the constraint set.
- The study illustrates applications in sparse (ℓ0) and ℓ1-constrained frameworks, offering benchmarks for robust high-dimensional inference in challenging noise regimes.
Robust Mean Estimation under Star-Shaped Constraints with Heavy-Tailed Noise
The paper investigates the robust estimation of the mean parameter μ for a distribution on Rn where the observations are subjected to both adversarial contamination (fraction ε<c for some constant c<1/2) and heavy-tailed additive noise with only a finite second moment σ2. The estimation is further constrained by the requirement that μ resides within a known star-shaped subset K⊆Rn, generalizing convex constraints. The statistical risk of interest is the minimax expected squared ℓ2 loss over all estimators, contamination schemes, and admissible distributions as
M=μ^infμ∈Ksupξ∈Ξσ2supCsupEμ[∥μ^(C(X~))−μ∥2],
where C denotes the adversary and Rn0 denotes the family of mean-zero distributions with maximal covariance eigenvalue Rn1.
The core technical contribution is the nonasymptotic characterization of Rn2 as a function of: the geometry of Rn3 (encoded by local metric entropy), contamination rate Rn4, noise variance Rn5, and sample size Rn6. The main result is that, for sufficiently large Rn7 (relative to the metric entropy), the minimax risk is
Rn8
where Rn9 and ε<c0 is a critical radius determined by the local metric entropy: ε<c1
For sign-symmetric noise (and, equivalently, for known noise laws via symmetrization), the dependence on contamination improves to ε<c2.
Lower and Upper Bounds: Packing, Contamination, and Algorithmic Rates
The minimax lower bound consists of two sharp contributions:
- Information-theoretic bound: Via Fano's lemma and local metric entropy, the best possible risk absent contamination is bounded below by the ε<c3 above.
- Adversarial contamination bound: The contamination imposes an irreducible error of order ε<c4 (or ε<c5 under symmetry).
Matching upper bounds are achieved by an information-theoretically optimal, albeit computationally intractable, estimator. The estimator operates by constructing a multiscale, local-packing-based directed tree on ε<c6 and performing robust tournaments at progressively finer resolutions. At each node/layer, the estimator selects the "best" candidate using robust hypothesis tests involving tournament-based trimmed means and medians (in the heavy-tailed regime) or Huber-type procedures (for symmetric noise).
The minimaxity is formally established by matching lower and upper bounds, managing both unbounded and bounded constraint sets, and precisely controlling error propagation through multiple layers of the estimation tree. For all ε<c7, the sample complexity is ε<c8.
Analysis in Heavy-Tailed and Symmetric Regimes
For general (potentially asymmetric) heavy-tailed noise, only finite variance is assumed, and the robust tests at the heart of the procedure combine trimming and median-based decision rules. The analysis leverages Cantelli inequalities to control one-sided deviations and establishes that, as long as the packing scales are chosen relative to the intrinsic local metric entropy of ε<c9, error control is feasible.
In the symmetric or known (via symmetrization) noise regime, Huber-type robust mean estimators are shown to achieve polynomial- or exponential-rate concentration, permitting improvement in the contamination term. The precise improvement is c<1/20.
Examples and Implications for Structured Constraints
The paper systematically analyzes two canonical cases:
- c<1/21-constraint ("sparse" mean estimation): c<1/22, the set of c<1/23-sparse vectors. The minimax rate is
c<1/24
provided c<1/25 is at least of order c<1/26. This rate differs from results in the one-sample setting and emphasizes the need for large enough c<1/27 in heavy-tailed scenarios.
- c<1/28-constraint:
c<1/29
with the precise transition point mapped via entropy calculations.
The results also clarify the role of local geometry (packing/covering structure) for nonconvex constraint sets, such as star-shaped but highly nonconvex sets, establishing that convexity is not essential for nonasymptotic optimality.
Theoretical and Practical Implications
The main theoretical implication is a tight, nonasymptotic minimax characterization for high-dimensional robust mean estimation under mild moment assumptions and very general geometric constraints (star-shaped, including nonconvex sets). The local metric entropy regime is critical, and the contamination-imposed error rate is sharp.
Practically, while the estimators described are not computationally efficient, the analysis suggests benchmarks for polynomial-time procedures and highlights the fundamental statistical price for treating contamination and heavy tails simultaneously. For special cases (e.g., symmetry), efficient procedures based on robust M-estimation may be available.
Potential directions for future work include:
- Extension to other functionals, including regression and density estimation under contamination and structure.
- Algorithmically efficient procedures achieving (nearly) optimal rates for more general classes of constraints (beyond convex/Type-2 bodies).
- Refinements in adaptivity to unknown σ20 and σ21, possibly via Lepskiĭ-type methods.
- Investigation of robust covariance (rather than only mean) estimation under analogous conditions.
Conclusion
The paper rigorously determines minimax rates for constrained robust mean estimation under star-shaped geometry and heavy-tailed, only-finite-variance noise, in the presence of adversarial contamination. The rates are expressed precisely in terms of local metric entropy and reveal a nuanced dependence on symmetry assumptions, constraint set geometry, and contamination fraction. The techniques advance understanding of robust estimation beyond convex or sub-Gaussian settings and set statistical benchmarks for both theory and practice (2604.05063).