Subspace Optimization for Backpropagation-Free Continual Test-Time Adaptation

๐Ÿ“… 2026-03-30
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This work addresses the challenge of achieving both efficiency and performance in backpropagation-free test-time adaptation under continual distribution shifts. Existing methods often suffer from redundant computation after domain stabilization or are constrained by input prompt updates. To overcome these limitations, the authors propose PACE, a novel framework that enables efficient adaptation by directly optimizing the affine parameters of normalization layers within a low-dimensional subspace. Key innovations include a subspace optimization strategy combining CMA-ES with Fastfood random projections, an adaptive stopping mechanism, and a cache of domain-specific vectors. Evaluated across multiple continual distribution shift benchmarks, PACE significantly outperforms current backpropagation-free approaches, reducing runtime by over 50% while achieving state-of-the-art accuracy.

Technology Category

Machine Learning: Transfer, Domain Adaptation, Multi-Task LearningSearch and Optimization: Learning to SearchComputer Vision: Learning & Optimization for CV

Application Category

Graph Algorithms and Modeling for the Web: Efficient manipulation of static and dynamic Web-related graphsSearch and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for rankingUser Modeling, Personalization and Recommendation: Fairness-aware retrieval and ranking
๐Ÿ“ Abstract
We introduce PACE, a backpropagation-free continual test-time adaptation system that directly optimizes the affine parameters of normalization layers. Existing derivative-free approaches struggle to balance runtime efficiency with learning capacity, as they either restrict updates to input prompts or require continuous, resource-intensive adaptation regardless of domain stability. To address these limitations, PACE leverages the Covariance Matrix Adaptation Evolution Strategy with the Fastfood projection to optimize high-dimensional affine parameters within a low-dimensional subspace, leading to superior adaptive performance. Furthermore, we enhance the runtime efficiency by incorporating an adaptation stopping criterion and a domain-specialized vector bank to eliminate redundant computation. Our framework achieves state-of-the-art accuracy across multiple benchmarks under continual distribution shifts, reducing runtime by over 50% compared to existing backpropagation-free methods.
Problem

Research questions and friction points this paper is trying to address.

continual test-time adaptation
backpropagation-free
runtime efficiency
learning capacity
distribution shifts
Innovation

Methods, ideas, or system contributions that make the work stand out.

backpropagation-free
test-time adaptation
subspace optimization
CMA-ES
Fastfood projection
๐Ÿ”Ž Similar Papers
No similar papers found.