arrow
Return

Efficient Dataset Distillation via Generative Pruning

delete2026-02-27
delete0
PRE
AI
Y
Yingyi Ma
M
Muquan Li
G
Guiduo Duan
K
Ke Qin
S
Shuang Liang
D
Dongyang Zhang
DOI:10.1109/tbdata.2026.3668524delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Dataset distillation (DD) has demonstrated the promise of synthesizing smaller datasets that enable competitive performance. While most of DD methods operate in pixel space, they often suffer from poor scalability on high-resolution datasets. Recent works have thus shifted towards parameterizing synthetic data using deep generative priors. However, these approaches apply latent updates to all generator layers, leading to substantial computational overhead. To address this limitation, we propose Generative Lightweight Distillation (GLiD), a unified framework that jointly compresses the generator and accelerates latent optimization for efficient distillation. GLiD introduces two key components: (1) We introduce a Classification–Diversity Sensitivity Pruning mechanism that quantifies both discriminative utility and semantic diversity of each output channel to guide structural pruning and selective layer-wise optimization. (2) We also present a Layer-Adaptive Scheduling strategy that dynamically allocates latent update steps across generator stages based on convergence behavior. Extensive experiments on CIFAR-10 and ImageNet-1K and its subsets demonstrate that our GLiD achieves up to 10× acceleration across different datasets, while maintaining performance competitive with state-of-the-art methods.
Keywords:
Dataset distillation
pruning
generative model

Journal

I
IEEE Transactions on Big Data
IF:
5.7
Papers:
887
Citations:
3.0K

Organization

U
university of electronic science and technology of china
Scholars:
1.3W
Papers: 4.8K
Citations: 4
Cited Papers

Cited Papers

Visualization of Big Spatial Data Using Coresets for Kernel Density Estimates
err2021-07-01
err6
errOAAI
errZheng, Yan; Ou, Yi; Lex, Alexander; Phillips, Jeff M.
errShare
errSave
High-Ratio Lossy Compression: Exploring the Autoencoder to Compress Scientific Data
err2023-02-01
err21
PREAI
errLiu, Tong; Wang, Jinzhen; Liu, Qing; Alibhai, Shakeel; Lu, Tao; He, Xubin
errShare
errSave
DC-BENCH: Dataset Condensation Benchmark
err2022-01-01
err0
PREAI
errCui,Justin; Wang,Ruochen; Si,Si; Hsieh,Cho-Jui
errShare
errSave
Mind the Gap in Distilling StyleGANs
err2022-01-01
err0
PREAI
errXu,Guodong; Hou,Yuenan; Liu,Ziwei; Loy,Chen Change
errShare
errSave
Content-Aware GAN Compression
err2021-06-01
err0
errOAAI
errYuchen Liu; Zhixin Shu; Yijun Li; Zhe Lin; Federico Perazzi; S.Y. Kung
errShare
errSave
researcher View more