Structured Pruning Adapters
Hedegaard, Lukas, Alok, Aman, Jose, Juby, Iosifidis, Alexandros
–arXiv.org Artificial Intelligence
Adapters are a parameter-efficient alternative to fine-tuning, which augment a frozen base network to learn new tasks. Yet, the inference of the adapted model is often slower than the corresponding fine-tuned model. To improve on this, we propose Structured Pruning Adapters (SPAs), a family of compressing, task-switching network adapters, that accelerate and specialize networks using tiny parameter sets and structured pruning. Specifically, we propose a channel-based SPA and evaluate it with a suite of pruning methods on multiple computer vision benchmarks. Compared to regular structured pruning with fine-tuning, our channel-SPAs improve accuracy by 6.9% on average while using half the parameters at 90% pruned weights. Alternatively, they can learn adaptations with 17x fewer parameters at 70% pruning with 1.6% lower accuracy. Similarly, our block-SPA requires far fewer parameters than pruning with fine-tuning. Our experimental code and Python library of adapters are available at github.com/lukashedegaard/structured-pruning-adapters.
arXiv.org Artificial Intelligence
Feb-2-2023
- Country:
- Oceania > Australia
- New South Wales > Sydney (0.04)
- North America
- Dominican Republic (0.04)
- Canada > Ontario
- Toronto (0.14)
- Europe
- Romania > Sud - Muntenia Development Region
- Giurgiu County > Giurgiu (0.04)
- Denmark > Central Jutland
- Aarhus (0.04)
- Romania > Sud - Muntenia Development Region
- Oceania > Australia
- Genre:
- Research Report (0.40)
- Technology: