Is Monotonic Sampling Necessary in Diffusion Models?

📅 2026-05-12
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This study investigates whether monotonic denoising schedules are necessary in diffusion models. By designing four classes of structured non-monotonic noise schedules, the authors conduct a systematic evaluation across 90 configurations and 42 hyperparameter ablation settings (involving 10–200 function evaluations) on CIFAR-10 using three prominent frameworks: DDPM, EDM, and Flow Matching. They introduce a novel diagnostic metric—the “schedule sensitivity coefficient”—to quantify architectural dependence on schedule monotonicity, revealing that DDPM suffers significant performance degradation, Flow Matching exhibits moderate sensitivity, and EDM remains largely unaffected. The results demonstrate that non-monotonic schedules do not outperform monotonic baselines, thereby affirming the validity and architecture-specific nature of monotonic sampling in diffusion-based generative modeling.
📝 Abstract
Diffusion models generate samples by iteratively denoising a Gaussian prior, traversing a sequence of noise levels that, in every published sampler, decreases monotonically. Six years of intensive work has refined nearly every aspect of this recipe, including the corruption operator, the training objective, the schedule shape, the architecture, and the ODE solver. Yet the assumption of monotonicity itself has never been systematically tested. Here we ask whether monotonic sampling is load-bearing or merely conventional. We design four families of structured nonmonotonic schedules and apply them to three architecturally distinct generative models, DDPM, EDM, and Flow Matching, across NFE budgets ranging from 10 to 200 function evaluations, plus a 42-cell hyperparameter ablation, on CIFAR-10. Across all 90 tested configurations, no tested nonmonotonic schedule improves on the monotonic baseline. The magnitude of the penalty, however, spans nearly three orders of magnitude: persistent and substantial in DDPM, intermediate in Flow Matching, and indistinguishable from zero in EDM. We show that this variation is not noise but a structural property of each trained denoiser, and we formalize it as the Schedule Sensitivity Coefficient, a cheap, architecture-agnostic diagnostic that provides evidence of non-convergence to the Bayes-optimal denoiser at the critical noise level. Our findings justify the field's tacit reliance on monotonic schedules and supply a new probe of diffusion model quality complementary to sample-quality metrics such as Frechet Inception Distance.
Problem

Research questions and friction points this paper is trying to address.

diffusion models
monotonic sampling
noise schedule
denoising
schedule sensitivity
Innovation

Methods, ideas, or system contributions that make the work stand out.

nonmonotonic sampling
diffusion models
schedule sensitivity
Bayes-optimal denoiser
model diagnostics
🔎 Similar Papers
2024-04-19Neural Information Processing SystemsCitations: 14
2022-09-02ACM Computing SurveysCitations: 1628
M
Muhammad Haris Khan
Department of Computer Science, University of Copenhagen, Denmark