CAFE: Learning to Condense Dataset by Aligning Features

Wang, Kai; Zhao, Bo; Peng, Xiangyu; Zhu, Zheng; Yang, Shuo; Wang, Shuo; Huang, Guan; Bilen, Hakan; Wang, Xinchao; You, Yang

Computer Science > Computer Vision and Pattern Recognition

arXiv:2203.01531 (cs)

[Submitted on 3 Mar 2022 (v1), last revised 27 Mar 2022 (this version, v2)]

Title:CAFE: Learning to Condense Dataset by Aligning Features

Authors:Kai Wang, Bo Zhao, Xiangyu Peng, Zheng Zhu, Shuo Yang, Shuo Wang, Guan Huang, Hakan Bilen, Xinchao Wang, Yang You

View PDF

Abstract:Dataset condensation aims at reducing the network training effort through condensing a cumbersome training set into a compact synthetic one. State-of-the-art approaches largely rely on learning the synthetic data by matching the gradients between the real and synthetic data batches. Despite the intuitive motivation and promising results, such gradient-based methods, by nature, easily overfit to a biased set of samples that produce dominant gradients, and thus lack global supervision of data distribution. In this paper, we propose a novel scheme to Condense dataset by Aligning FEatures (CAFE), which explicitly attempts to preserve the real-feature distribution as well as the discriminant power of the resulting synthetic set, lending itself to strong generalization capability to various architectures. At the heart of our approach is an effective strategy to align features from the real and synthetic data across various scales, while accounting for the classification of real samples. Our scheme is further backed up by a novel dynamic bi-level optimization, which adaptively adjusts parameter updates to prevent over-/under-fitting. We validate the proposed CAFE across various datasets, and demonstrate that it generally outperforms the state of the art: on the SVHN dataset, for example, the performance gain is up to 11%. Extensive experiments and analyses verify the effectiveness and necessity of proposed designs.

Comments:	The manuscript has been accepted by CVPR-2022!
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2203.01531 [cs.CV]
	(or arXiv:2203.01531v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2203.01531

Submission history

From: Kai Wang [view email]
[v1] Thu, 3 Mar 2022 05:58:49 UTC (2,468 KB)
[v2] Sun, 27 Mar 2022 17:13:08 UTC (2,494 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:CAFE: Learning to Condense Dataset by Aligning Features

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:CAFE: Learning to Condense Dataset by Aligning Features

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators