Exploring Alignment of Representations with Human Perception
Nanda, Vedant, Majumdar, Ayan, Kolling, Camila, Dickerson, John P., Gummadi, Krishna P., Love, Bradley C., Weller, Adrian
–arXiv.org Artificial Intelligence
We argue that a valuable perspective on when a model learns \textit{good} representations is that inputs that are mapped to similar representations by the model should be perceived similarly by humans. We use \textit{representation inversion} to generate multiple inputs that map to the same model representation, then quantify the perceptual similarity of these inputs via human surveys. Our approach yields a measure of the extent to which a model is aligned with human perception. Using this measure of alignment, we evaluate models trained with various learning paradigms (\eg~supervised and self-supervised learning) and different training losses (standard and robust training). Our results suggest that the alignment of representations with human perception provides useful additional insights into the qualities of a model. For example, we find that alignment with human perception can be used as a measure of trust in a model's prediction on inputs where different models have conflicting outputs. We also find that various properties of a model like its architecture, training paradigm, training loss, and data augmentation play a significant role in learning representations that are aligned with human perception.
arXiv.org Artificial Intelligence
Nov-29-2021
- Country:
- Asia (0.04)
- North America
- United States > Maryland (0.04)
- Canada > Quebec
- Montreal (0.04)
- Europe > United Kingdom
- England > Cambridgeshire > Cambridge (0.04)
- Genre:
- Research Report > New Finding (1.00)
- Industry:
- Technology: