Scaleable input gradient regularization for adversarial robustness
Finlay, Chris, Oberman, Adam M
Input gradient regularization is not thought to be an effective means for promoting adversarial robustness. In this work we revisit this regularization scheme with some new ingredients. First, we derive new per-image theoretical robustness bounds based on local gradient information, and curvature information when available. These bounds strongly motivate input gradient regularization. Second, we implement a scaleable version of input gradient regularization which avoids double backpropagation: adversarially robust ImageNet models are trained in 33 hours on four consumer grade GPUs. Finally, we show experimentally that input gradient regularization is competitive with adversarial training.
May-27-2019
- Country:
- North America
- United States
- Washington > King County
- Bellevue (0.04)
- Virginia > Fairfax County
- McLean (0.04)
- Texas > Dallas County
- Dallas (0.04)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Hawaii > Honolulu County
- Honolulu (0.04)
- Florida > Miami-Dade County
- Miami (0.04)
- California
- Los Angeles County > Long Beach (0.04)
- Santa Clara County > San Jose (0.04)
- Washington > King County
- Canada
- Quebec > Montreal (0.14)
- Ontario > Toronto (0.14)
- British Columbia > Metro Vancouver Regional District
- Vancouver (0.04)
- United States
- Europe
- Sweden > Stockholm
- Stockholm (0.04)
- Netherlands > North Holland
- Amsterdam (0.04)
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Germany
- Saarland > Saarbrücken (0.04)
- Bavaria > Upper Bavaria
- Munich (0.04)
- Sweden > Stockholm
- Asia > Middle East
- Jordan (0.04)
- UAE > Abu Dhabi Emirate
- Abu Dhabi (0.04)
- North America
- Genre:
- Research Report (0.84)
- Industry:
- Information Technology > Security & Privacy (0.48)
- Government > Military (0.30)
- Technology: