Dataset Condensation Atlas

Method · Optimization and training recipes

DiRe

DiRe: Diversity-promoting Regularization for Dataset Condensation

Saumyaranjan Mohanty, Aravind Reddy, Konda Reddy Mopuri

WACV 2026 · first public 2025-12-15 · arXiv 2512.13083

paper ↗code ↗catalogued✓ abstract read

In one paragraph

Proposes DiRe, a diversity regularizer combining cosine similarity and Euclidean distance terms that plugs into existing condensation methods off the shelf to reduce redundancy among synthesized samples; reports consistent generalization and diversity-metric improvements when added to state-of-the-art condensation methods from CIFAR-10 to ImageNet-1K.

Where it sits

Abstract (verbatim from arXiv)

In Dataset Condensation, the goal is to synthesize a small dataset that replicates the training utility of a large original dataset. Existing condensation methods synthesize datasets with significant redundancy, so there is a dire need to reduce redundancy and improve the diversity of the synthesized datasets. To tackle this, we propose an intuitive Diversity Regularizer (DiRe) composed of cosine similarity and Euclidean distance, which can be applied off-the-shelf to various state-of-the-art condensation methods. Through extensive experiments, we demonstrate that the addition of our regularizer improves state-of-the-art condensation methods on various benchmark datasets from CIFAR-10 to ImageNet-1K with respect to generalization and diversity metrics.

BibTeX (generated; prefer the venue's official entry)
@article{mohanty2025dire,
  title   = {DiRe: Diversity-promoting Regularization for Dataset Condensation},
  author  = {Saumyaranjan Mohanty and Aravind Reddy and Konda Reddy Mopuri},
  journal = {WACV 2026},
  year    = {2025}
}

Nearby in Optimization and training recipes

2026-05

C^2R — Mind Your Margin and Boundary: Are Your Distilled Datasets Truly Robust?

Muquan Li, Yingyi Ma, Yihong Huang et al. · ICML 2026notablepaper ↗

2026-04

COBRA — Fair Dataset Distillation via Cross-Group Barycenter Alignment

Mohammad Hossein Moslemi, Nima Hosseini Dashtbayaz, Zhimin Mei et al. · ICML 2026notablepaper ↗code ↗

2026-03

PTM-ST — Multimodal Dataset Distillation via Phased Teacher Models

Shengbin Guo, Hang Zhao, Senqiao Yang et al. · ICLR 2026notableVision–languagepaper ↗code ↗

2026-03

FD2 — FD$^2$: A Dedicated Framework for Fine-Grained Dataset Distillation

Hongxu Ma, Guang Li, Shijie Wang et al. · ECCV 2026notablepaper ↗

2025-05

PRISM — PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion

Jaehyun Choi, Jiwan Hur, Gyojin Han et al. · CVPR 2026notableVideopaper ↗