arXiv · 2605.02183
Manifold-Constrained Adversarial Training for Long-Tailed Robustness via Geometric Alignment
Abstract
Adversarial training is effective on balanced datasets, but its robustness degrades under longtailed class distributions, where tail classes suffer high robust error and unstable decision boundaries. We propose Manifold-Constrained Adversarial Training (MCAT), a unified framework that enforces the semantic validity of adversarial examples by penalizing deviations from class-conditional manifolds in feature space, while promoting balanced geometric separation across classes via an ETF-inspired regularization. We provide theoretical results that link geometric separation to lower bounds on adversarially robust margins, and show that manifold-constrained adversarial risk upperbounds robust risk on high-density semantic regions. Extensive experiments on standard longtailed benchmarks demonstrate consistent improvements in overall, balanced, and tail-class adversarial robustness.
Explore related subjects
Keep this discovery
Guanmeng Xian, Ning Yang, Philip S. Yu. 2026-05-04. Manifold-Constrained Adversarial Training for Long-Tailed Robustness via Geometric Alignment. https://arxiv.org/abs/2605.02183
Cite the original work for its findings. Save a collection to share your selection of sources.