arrow
Return

Drop2Sparse: Improving Dataset Distillation via Sparse Model

delete2025-08-01
delete0
PRE
AI
T
Ting-Feng Huang
Y
Yu‐Hsun Lin
DOI:10.1109/TCSVT.2025.3552047delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The success of modern deep learning algorithms requires large amounts of training data, which leads to high computational and storage costs. Dataset Distillation (DD) is a rising research field that resolves this issue by synthesizing a compact training dataset from a large one. Recent gradient matching DD methods have achieved remarkable results. However, these methods typically utilize weak models for DD performance improvement, while well-trained models are often considered inferior choices due to their lower performance. Conversely, our study provides new insights into the role of well-trained models in DD, particularly under high-storage budget scenarios. We identify a previously overlooked design principle—a positive correlation between model capability and storage budget. Based on this principle, we propose Drop2Sparse, an approach that randomly sparsifies well-trained models to create efficient models for various storage budget scenarios. Drop2Sparse concurrently infuses significant model diversity and regularization effects into DD, outperforming previous state-of-the-art methods by up to 3.8% on CIFAR and 3.6% on ImageNet-subset. Moreover, our method exhibits remarkable cross-architecture generalization and achieves promising results even under challenging scenarios, such as using an extremely reduced model pool or highly accelerated training.
Keywords:
Dataset compression
dataset distillation
gradient matching
model sparsification

Journal

IEEE Transactions on Circuits and Systems for Video Technology cover
IEEE Transactions on Circuits and Systems for Video Technology
IF:
11.1
Papers:
845
Citations:
3.1W

Organization

N
National Tsing Hua University
Scholars:
1.6W
Papers: 1.4W
Citations: 1.7W
Cited Papers

Cited Papers

DataDAM: Efficient Dataset Distillation with Attention Matching
err
IF0
err2023-10-01
err0
PREAI
errAhmad Sajedi; Samir Khaki; Ehsan Amjadian; Lucy Z. Liu; Yuri A. Lawryshyn; Konstantinos N. Plataniotis
errShare
errSave
errShare
errSave
errShare
errSave
Kinect-Like Depth Data Compression
err2013-10-01
err42
PREAI
errFu, Jingjing; Miao, Dan; Yu, Weiren; Wang, Shiqi; Lu, Yan; Li, Shipeng
errShare
errSave
Accelerating Dataset Distillation via Model Augmentation
err2023-06-01
err0
errOAAI
errLei Zhang; Jie Zhang; Bowen Lei; Subhabrata Mukherjee; Xiang Pan; Bo Zhao; Caiwen Ding; Yao Li; Dongkuan Xu
errShare
errSave
CAFE: Learning to Condense Dataset by Aligning Features
err2022-06-01
err0
errOAAI
errKai Wang; Bo Zhao; Xiangyu Peng; Zheng Zhu; Shuo Yang; Shuo Wang; Guan Huang; Hakan Bilen; Xinchao Wang; Yang You
errShare
errSave
Densely Connected Convolutional Networks
err2017-07-01
err0
errOAAI
errGao Huang; Zhuang Liu; Laurens Van Der Maaten; Kilian Q. Weinberger
errShare
errSave
researcher View more