arXiv · 2610.11288
Memorization and Malign Generalization in Conditional Diffusion Models with Random Features
Abstract
Conditional diffusion models generate diverse, novel, and high-quality samples under prescribed conditions. However, theoretical understanding of their memorization and generalization remains limited, while recent works have characterized these behaviors primarily in unconditional settings. In this work, we analyze a random-feature conditional score model in the high-dimensional proportional limit, deriving asymptotic expressions for training and test losses. By decomposing the test loss, we show that in the overparameterized regime, increasing model width improves prediction of the condition-dependent mean while reducing within-condition prediction variance, a phenomenon we term "malign generalization." Furthermore, analyzing the training loss reveals that more informative conditions lead to memorization of training samples at smaller widths. These theoretical findings are supported by experiments with U-Net architectures on realistic data.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Gwangho Kim, Sungyoon Lee. 2026-10-08. Memorization and Malign Generalization in Conditional Diffusion Models with Random Features. https://arxiv.org/abs/2610.11288
Cite the original work for its findings. Save a collection to share your selection of sources.