Why Can Accurate Models Be Learned from Inaccurate Annotations?

📅 2025-05-22
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work investigates the intrinsic mechanism enabling deep models to generalize well under label noise. Theoretically, we show that label noise primarily perturbs low-order singular components of weight matrices, while the dominant subspace governing generalization—the principal subspace spanned by top singular vectors—remains stably aligned. This yields the first rigorous subspace-level characterization of generalization robustness under label corruption. Building on this insight, we propose LIP (Low-rank Invariant Projection), a lightweight plug-in module that explicitly regularizes the principal subspace structure of weight matrices during training. Extensive experiments across diverse synthetic and real-world noisy-label benchmarks demonstrate that LIP consistently improves classification accuracy of mainstream architectures—including ResNet and ViT—without architectural modification. Our results empirically and theoretically establish principal subspace stability as the fundamental geometric principle underlying label-noise robustness.

Technology Category

Machine Learning: Adversarial Learning & RobustnessComputer Vision: Adversarial Attacks & RobustnessNatural Language Processing: Safety and Robustness

Application Category

Web Mining and Content Analysis: Robustness and generalizability of Web mining methodsGraph Algorithms and Modeling for the Web: Foundation models and LLMs for Web-related graphsSecurity and Privacy: Security and privacy of machine learning and AI applications
📝 Abstract
Learning from inaccurate annotations has gained significant attention due to the high cost of precise labeling. However, despite the presence of erroneous labels, models trained on noisy data often retain the ability to make accurate predictions. This intriguing phenomenon raises a fundamental yet largely unexplored question: why models can still extract correct label information from inaccurate annotations remains unexplored. In this paper, we conduct a comprehensive investigation into this issue. By analyzing weight matrices from both empirical and theoretical perspectives, we find that label inaccuracy primarily accumulates noise in lower singular components and subtly perturbs the principal subspace. Within a certain range, the principal subspaces of weights trained on inaccurate labels remain largely aligned with those learned from clean labels, preserving essential task-relevant information. We formally prove that the angles of principal subspaces exhibit minimal deviation under moderate label inaccuracy, explaining why models can still generalize effectively. Building on these insights, we propose LIP, a lightweight plug-in designed to help classifiers retain principal subspace information while mitigating noise induced by label inaccuracy. Extensive experiments on tasks with various inaccuracy conditions demonstrate that LIP consistently enhances the performance of existing algorithms. We hope our findings can offer valuable theoretical and practical insights to understand of model robustness under inaccurate supervision.
Problem

Research questions and friction points this paper is trying to address.

Understanding why models learn accurately from inaccurate labels
Analyzing noise impact on weight matrices and principal subspaces
Proposing LIP to preserve principal subspace and reduce noise
Innovation

Methods, ideas, or system contributions that make the work stand out.

Analyzes weight matrices to identify noise accumulation
Proves minimal deviation in principal subspace angles
Introduces LIP to retain principal subspace information
💼 Related Jobs
No related jobs found.
C
Chongjie Si
Shanghai Jiao Tong University
Y
Yidan Cui
Shanghai Jiao Tong University
F
Fuchao Yang
Southeast University
X
Xiaokang Yang
Shanghai Jiao Tong University
W
Wei Shen
Shanghai Jiao Tong University