Generalization Error Analysis of Neural networks with Gradient Based Regularization
Li, Lingfeng, Tai, Xue-Cheng, Yang, Jiang
–arXiv.org Artificial Intelligence
We study gradient-based regularization methods for neural networks. We mainly focus on two regularization methods: the total variation and the Tikhonov regularization. Applying these methods is equivalent to using neural networks to solve some partial differential equations, mostly in high dimensions in practical applications. In this work, we introduce a general framework to analyze the generalization error of regularized networks. The error estimate relies on two assumptions on the approximation error and the quadrature error. Moreover, we conduct some experiments on the image classification tasks to show that gradient-based methods can significantly improve the generalization ability and adversarial robustness of neural networks. A graphical extension of the gradient-based methods are also considered in the experiments.
arXiv.org Artificial Intelligence
Jul-6-2021
- Country:
- North America > United States
- Rhode Island > Providence County
- Providence (0.04)
- California > San Diego County
- San Diego (0.04)
- Rhode Island > Providence County
- Europe > United Kingdom
- England > Cambridgeshire > Cambridge (0.04)
- Asia > China
- Hong Kong (0.04)
- Guangdong Province > Shenzhen (0.04)
- North America > United States
- Genre:
- Research Report > New Finding (0.46)
- Industry:
- Government (0.47)
- Technology: