Theoretical Guarantees for One-Shot Magnitude Pruning and Compute-Adaptive Early Exit

๐Ÿ“… 2026-09-10
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
็ ”็ฉถ้€š่ฟ‡ไธ€ๆฌกๆ€งๅน…ๅบฆๅ‰ชๆžๅ’Œ่‡ช้€‚ๅบ”ๆๅ‰้€€ๅ‡บๆ–นๆณ•ๅ‡ๅฐ‘็ฅž็ป็ฝ‘็ปœ่ฎก็ฎ—้‡๏ผŒๅนถ่ฏๆ˜Žไบ†่ฟ™ไบ›ๆ–นๆณ•็š„ๆœ‰ๆ•ˆๆ€งๅŠ่ฏฏๅทฎ้š่ฎก็ฎ—ๅทฎ่ท็š„ๅ˜ๅŒ–่ง„ๅพ‹ใ€‚
๐Ÿ“ Abstract
We study compute reduction in neural networks through a unified partial versus full computation view, captured by one-shot magnitude pruning in the static regime and early exit in the adaptive regime. In an asymptotic single-neuron model, we prove a concentration theorem for one-shot magnitude pruning with explicit rates. We also introduce the conditional perceptron for early exit and show that its excess generalization error decays as a power of the compute gap, with an exponent that grows to infinity as the alignment between partial and full computations tends to one. We then extend the analysis to deep networks, characterizing how pruning-induced distortions accumulate with depth and deriving a corresponding compute-accuracy tradeoff for frozen-backbone early exit under a neural network Gaussian process model. Numerical simulations corroborate the predicted scaling laws.
Problem

Research questions and friction points this paper is trying to address.

one-shot magnitude pruning
compute-adaptive early exit
neural networks
computation reduction
generalization error
Innovation

Methods, ideas, or system contributions that make the work stand out.

one-shot magnitude pruning
compute-adaptive early exit
conditional perceptron
excess generalization error
neural network Gaussian process
๐Ÿ”Ž Similar Papers
No similar papers found.
๐Ÿ’ผ Related Jobs
No related jobs found.
E
Erdem Koyuncu
Department of Electrical and Computer Engineering, University of Illinois Chicago, IL, USA