Frequency Domain-Based Dataset Distillation
DongHyeok Shin, Seungjae Shin, Il-Chul Moon
摘要
This paper presents FreD, a novel parameterization method for dataset distillation, which utilizes the frequency domain to distill a small-sized synthetic dataset from a large-sized original dataset. Unlike conventional approaches that focus on the spatial domain, FreD employs frequency-based transforms to optimize the frequency representations of each data instance. By leveraging the concentration of spatial domain information on specific frequency components, FreD intelligently selects a subset of frequency dimensions for optimization, leading to a significant reduction in the required budget for synthesizing an instance. Through the selection of frequency dimensions based on the explained variance, FreD demonstrates both theoretical and empirical evidence of its ability to operate efficiently within a limited budget, while better preserving the information of the original dataset compared to conventional parameterization methods. Furthermore, based on the orthogonal compatibility of FreD with existing methods, we confirm that FreD consistently improves the performances of existing distillation methods over the evaluation scenarios with different benchmark datasets. We release the code at https://github.com/sdh0818/FreD .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- What is Dataset Distillation Learning?William Yang, Ye Zhu, Zhiwei Deng, Olga RussakovskyICML 2024 · 被引用 13 次
- FADRM: Fast and Accurate Data Residual Matching for Dataset DistillationJiacheng Cui, Xinyue Bi, Yaxin Luo, Xiaohan Zhao 等NeurIPS 2025 · 被引用 12 次
- Ameliorate Spurious Correlations in Dataset CondensationJustin Cui, Ruochen Wang, Yuanhao Xiong, Cho-Jui HsiehICML 2024 · 被引用 7 次
- Color-Oriented Redundancy Reduction in Dataset DistillationBowen Yuan, Zijian Wang, Mahsa Baktashmotlagh, Yadan Luo 等NeurIPS 2024 · 被引用 7 次
- Multimodal Distribution Matching for Vision-Language Dataset DistillationJongoh Jeong, Hoyong Kwon, Minseok Kim, Kuk-Jin YoonCVPR 2026 · 被引用 3 次
它引用的顶会 Paper16
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Dataset Condensation with Gradient MatchingBo Zhao, Konda Reddy Mopuri, Hakan BilenICLR 2021 · 被引用 684 次
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 被引用 422 次
- Dataset Condensation with Differentiable Siamese AugmentationBo Zhao, Hakan BilenICML 2021 · 被引用 390 次
- Dataset Distillation with Infinitely Wide Convolutional NetworksTimothy Nguyen, Roman Novak, Lechao Xiao, Jaehoon LeeNeurIPS 2021 · 被引用 313 次
相关 Paper
- Distilling Dataset into Neural FieldDonghyeok Shin, HeeSun Bae, Gyuwon Sim, Wanmo Kang 等ICLR 2025
- Slimmable Dataset CondensationSonghua Liu, Jingwen Ye, Runpeng Yu, Xinchao WangCVPR 2023
- Hierarchical Features Matter: A Deep Exploration of Progressive Parameterization Method for Dataset DistillationXinhao Zhong, Hao Fang, Bin Chen, Xulin Gu 等CVPR 2025
- Boost Self-Supervised Dataset Distillation via Parameterization, Predefined Augmentation, and ApproximationSheng-Feng Yu, Jia-Jiun Yao, Wei-Chen ChiuICLR 2025
- Sparse Parameterization for Epitomic Dataset DistillationXing Wei, Anjia Cao, Funing Yang, Zhiheng MaNeurIPS 2023 · 被引用 26 次
