What does a deep neural network confidently perceive? The effective dimension of high certainty class manifolds and their low confidence boundaries
Fort, Stanislav, Cubuk, Ekin Dogus, Ganguli, Surya, Schoenholz, Samuel S.
–arXiv.org Artificial Intelligence
The geometry of these class manifolds (CMs) is widely studied and intimately related to model performance; for example, the margin depends on CM boundaries. We exploit the notions of Gaussian width and Gordon's escape theorem to tractably estimate the effective dimension of CMs and their boundaries through tomographic intersections with random affine subspaces of varying dimension. We show several connections between the dimension of CMs, generalization, and robustness. In particular we investigate how CM dimension depends on 1) the dataset, 2) architecture (including ResNet, WideResNet & Vision Transformer), 3) initialization, 4) stage of training, 5) class, 6) network width, 7) ensemble size, 8) label randomization, 9) training set size, and 10) robustness to data corruption. Together a picture emerges that higher performing and more robust models have higher dimensional CMs. Moreover, we offer a new perspective on ensembling via intersections of CMs. Our code is on Github.
arXiv.org Artificial Intelligence
Oct-11-2022
- Country:
- North America
- United States
- Ohio (0.04)
- California > Santa Clara County
- Stanford (0.04)
- Mountain View (0.04)
- Canada > Ontario
- Toronto (0.04)
- United States
- Europe
- United Kingdom > England
- Cambridgeshire > Cambridge (0.04)
- Switzerland > Zürich
- Zürich (0.04)
- United Kingdom > England
- North America
- Genre:
- Research Report > New Finding (0.46)
- Technology: