🤖 AI Summary
Training deep neural networks often suffers from catastrophic loss explosions, leading to costly failures. Conventional monitoring metrics—such as weight or gradient norms—are lagging and lack discriminative power for early detection. This paper proposes a spectral alignment–based early-warning mechanism: it quantifies the alignment between layer-wise input distributions and the dominant left singular vectors of weight matrices to detect incipient representational collapse. Theoretically, we show that the collapse of sign diversity in spectral alignment serves as an interpretable, pre-divergence indicator of training instability. Our method requires only lightweight SVD computation and statistical tracking, entailing minimal overhead and straightforward deployment. Empirical evaluation on language models demonstrates that our approach issues warnings significantly earlier than conventional metrics, with clearer signals and stronger generalization across architectures. This work establishes a novel paradigm for stabilizing large-model training through interpretable, spectrum-aware monitoring.
📝 Abstract
Loss explosions in training deep neural networks can nullify multi-million dollar training runs. Conventional monitoring metrics like weight and gradient norms are often lagging and ambiguous predictors, as their values vary dramatically across different models and even between layers of the same model, making it difficult to establish a unified standard for detecting impending failure. We introduce Spectral Alignment (SA), a novel, theoretically-grounded metric that monitors the distributional alignment between layer inputs and the principal singular vectors of weight matrices. We show that a collapse in the sign diversity of this alignment is a powerful early predictor of representational collapse and training divergence. Empirical results on language models demonstrate that monitoring the SA distribution provides a significantly earlier and clearer warning of loss explosions than traditional scalar metrics. SA's low computational overhead makes it a practical tool for safeguarding model training.