vix.ing · top · new · best · stats · spec

Robustness Emerges Early in Training Dynamics, but Is Not Preserved

2026/08/05 by Jiangang Yang, Wenhui Shi, Lu Hu +2
Computer Science · #cs.LG #cs.CV

paper · pdf

Accepted by ECCV2026

arxiv created 2026/08/05 · arxiv updated 2026/08/06

Abstract

Robustness to natural corruptions remains a fundamental challenge for deep neural networks. In this paper, we identify a robustness fading phenomenon where shallow layers spontaneously develop robust representations and flat loss landscapes in early training, yet these properties are not preserved during standard convergence. To address this, we propose a framework that performs strategic interventions on training dynamics to stabilize the empirically identified early-emergent robust priors. Our approach includes two parameter-free strategies: Early-Phase Stabilization~(EPS) and Asymmetric Weight Reversion~(AWR), which stabilize or recover robust shallow configurations without modifying the model architecture or introducing learnable parameters. Extensive experiments demonstrate the efficacy of our framework across various benchmarks and architectures, yielding significant gains in downstream transfer, dynamic adaptation, and diverse computer vision applications.

Citations