Skip to content
Conference

CoRe-Diffusion: Bridging the Generative-Discriminative Gap in Dataset Distillation via Manifold-Aligned Contrastive Guidance

2026 · Poster Volume 0007 The 2026 Twenty-Second International Conference on Intelligent Computing July 23-26, 2026 Toronto, Canada · 0 citations

TL;DR

CoRe-Diffusion is proposed as a unified framework, which introduces Contrastive Negative Guidance, which utilizes inter-class mode centers—computed but largely unused during synthesis by existing methods—to exert zero-overhead repulsive gradients against hard negatives.

Abstract

Dataset Distillation aims to condense large-scale datasets into a tiny but highly informative subset, enabling efficient training while achieving performance comparable to the full dataset. To overcome the scalability bottlenecks of traditional optimization, Generative Dataset Distillation (GDD) leverages diffusion priors to reformulate the dataset condensation process as an efficient generative sampling paradigm. However, existing GDD paradigms primarily optimize intra-class likelihood while largely overlooking explicit inter-class separation, which may cause generated features to concentrate near decision boundaries. Furthermore, relying on manually designed guidance targets to steer synthesis often pushes the diffusion trajectory off the natural data manifold, inducing severe structural artifacts. To address these issues, CoRe-Diffusion is proposed as a unified framework. Specifically, it introduces Contrastive Negative Guidance, which utilizes inter-class mode centers—computed but largely unused during synthesis by existing methods—to exert zero-overhead repulsive gradients against hard negatives. To preserve synthesis fidelity, Manifold-Aligned Real-Anchor Discovery strictly constrains the guidance targets to the exact latent representations of real images, preventing the trajectory from drifting. This intrinsically preserves complex semantic structures, bypassing the prohibitive computational overhead of relying on external generative priors or auxiliary modules. Alongside an Annealed Sampling Schedule for smooth trajectory evolution, CoRe-Diffusion achieves state-of-the-art performance, yielding a 3.8\% absolute accuracy improvement on ImageNet-1K while introducing negligible computational overhead.

View source

Similar papers

Preprint Aug 2026

DIFFCZSL: Compositional Zero-Shot Learning Regularized by Diffusion Representations

DIFFCZSL is proposed, a diffusion-augmented framework that injects generative priors from pre-trained diffusion models into CLIP-based CZSL pipelines and highlights the complementary strengths of generative diffusion representations and discriminative vision-language models for compositional generalization.

Hangyu Tian, Zhen-Qi He, Yanghao Wang et al. · 0 citations
Preprint Aug 2026

DeCO: Discriminative Evidence Composition for Fine-Grained Dataset Distillation

DeCO uses attention rollout from a pretrained TransFG teacher to identify informative patches, applies spatial diversification to reduce redundant coverage, and organizes the resulting regions into class-wise evidence banks and consistently outperforms representative coreset and dataset-distillation baselines under dif...

Chuixuan Fan, Guang Li, Shi-Jie Wang et al. · 0 citations
Preprint Sep 2026

Joint Alignment and Distillation for Video Generation via Sample-Guided Distribution Matching

DM-Align is introduced, which derives a complementary gradient direction to guide the model toward human-preferred samples, and eliminates the need for multi-step reward evaluation and complex ODE-SDE conversions inherent in traditional RL.

Jiu-Zhou Lin, Jun-Long Wu, Feilong Zuo et al. · 1 citation
Preprint Sep 2026

Isotropic Embedding Perturbations for Robust Vision Language Encoders

Aether is introduced, a simple plug-in method that applies diffusion-style random perturbations in the embedding space via controlled alpha-mixing, specifically designed to provide isotropic regularization that remains semantically consistent.

Hyesong Choi, Daeun Kim, Song Park et al. · 0 citations
Conference Aug 2026

Lightweight Stable Diffusion via StableKOT: Knowledge Distillation Meets Optimal Transport

Stable Diffusion models deliver photorealistic image synthesis but face computational bottlenecks that hinder deployment in resource-constrained environments. However, most existing distillation methods for diffusion models rely on point-wise feature matching, which often fails to preserve global semantic structure, or...

Bich-Nga Pham, Quoc-Truong Truong, Anh-Khoa Nguyen Vu et al. · 0 citations
Sep 2026

Information-Bottlenecked Variational Autoencoder for Top-N Recommendation via Gating Mechanism

Generative models for top- \(N\) recommendation have garnered significant attention, with Variational Autoencoder (VAE) emerging as a promising approach for modeling user preferences. Yet, traditional VAE-based models encounter two major challenges: simplistic priors may cause posterior collapse, resulting in ineffecti...

Xiao-Bo Guo, Shaoshuai Li, You-Ru Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.