Dataset Distillation with Convexified Implicit Gradients
Loo, Noel, Hasani, Ramin, Lechner, Mathias, Rus, Daniela
We propose a new dataset distillation algorithm using reparameterization and convexification of implicit gradients (RCIG), that substantially improves the state-of-the-art. To this end, we first formulate dataset distillation as a bi-level optimization problem. Then, we show how implicit gradients can be effectively used to compute meta-gradient updates. We further equip the algorithm with a convexified approximation that corresponds to learning on top of a frozen finite-width neural tangent kernel. Finally, we improve bias in implicit gradients by parameterizing the neural network to enable analytical computation of final-layer parameters given the body parameters. RCIG establishes the new state-of-the-art on a diverse series of dataset distillation tasks. Notably, with one image per class, on resized ImageNet, RCIG sees on average a 108\% improvement over the previous state-of-the-art distillation algorithm. Similarly, we observed a 66\% gain over SOTA on Tiny-ImageNet and 37\% on CIFAR-100.
Nov-9-2023
- Country:
- Pacific Ocean > South Pacific Ocean
- Coral Sea (0.04)
- North America
- United States
- Maryland > Baltimore (0.04)
- Arizona (0.04)
- Wisconsin > Dane County
- Madison (0.04)
- New York > New York County
- New York City (0.04)
- Massachusetts > Middlesex County
- Cambridge (0.04)
- Hawaii > Honolulu County
- Honolulu (0.04)
- California > San Diego County
- San Diego (0.04)
- Canada > Ontario
- Toronto (0.04)
- United States
- Europe > Spain
- Canary Islands (0.04)
- Asia > Middle East
- Jordan (0.04)
- Pacific Ocean > South Pacific Ocean
- Genre:
- Research Report (1.00)
- Industry:
- Technology: