Dataset Condensation Atlas

Method · Gradient matching

CGM

Gradient Matching for Categorical Data Distillation in CTR Prediction

Cheng Wang, Jiacheng Sun, Zhenhua Dong, Ruixuan Li, Rui Zhang

RecSys 2023 · first public 2023-01

paper ↗catalogued✓ abstract read

In one paragraph

Proposes CGM (Categorical data distillation with Gradient Matching), which extends gradient-matching dataset distillation to the high-dimensional, sparse categorical features of click-through-rate prediction data, addressing the blocked gradient flow through categorical embeddings and the cost of the resulting bi-level optimization; distills a small synthetic dataset that trains CTR models from scratch toward performance close to training on the full data.

Where it sits

BibTeX (generated; prefer the venue's official entry)
@article{wang2023gradient,
  title   = {Gradient Matching for Categorical Data Distillation in CTR Prediction},
  author  = {Cheng Wang and Jiacheng Sun and Zhenhua Dong and Ruixuan Li and Rui Zhang},
  journal = {RecSys 2023},
  year    = {2023}
}

Nearby in Gradient matching

2025-11

Linear Gradient Matching — Dataset Distillation for Pre-Trained Self-Supervised Vision Models

George Cazenavette, Antonio Torralba, Vincent Sitzmann · NeurIPS 2025notablePre-training & transferpaper ↗code ↗

2025-05

PRISM — PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion

Jaehyun Choi, Jiwan Hur, Gyojin Han et al. · CVPR 2026notableVideopaper ↗

2025-02

GRADMM — Synthetic Text Generation for Training Large Language Models via Gradient Matching

Dang Nguyen, Zeman Li, Mohammadhossein Bateni et al. · ICML 2025notableTextpaper ↗code ↗

2024-04

Distilled Datamodel with Reverse Gradient Matching

Jingwen Ye, Ruonan Yu, Songhua Liu et al. · CVPR 2024notablepaper ↗

2023-12

Static-dynamic video DD — Dancing with Still Images: Video Distillation via Static-Dynamic Disentanglement

Ziyu Wang, Yue Xu, Cewu Lu et al. · CVPR 2024coreVideopaper ↗code ↗