Adversarial Self-Attention for Language Understanding
Wu, Hongqiu, Ding, Ruixue, Zhao, Hai, Xie, Pengjun, Huang, Fei, Zhang, Min
–arXiv.org Artificial Intelligence
Deep neural models (e.g. Transformer) naturally learn spurious features, which create a ``shortcut'' between the labels and inputs, thus impairing the generalization and robustness. This paper advances the self-attention mechanism to its robust variant for Transformer-based pre-trained language models (e.g. BERT). We propose \textit{Adversarial Self-Attention} mechanism (ASA), which adversarially biases the attentions to effectively suppress the model reliance on features (e.g. specific keywords) and encourage its exploration of broader semantics. We conduct a comprehensive evaluation across a wide range of tasks for both pre-training and fine-tuning stages. For pre-training, ASA unfolds remarkable performance gains compared to naive training for longer steps. For fine-tuning, ASA-empowered models outweigh naive models by a large margin considering both generalization and robustness.
arXiv.org Artificial Intelligence
Feb-8-2023
- Country:
- Oceania > Australia
- New South Wales > Sydney (0.04)
- North America
- United States
- Washington > King County
- Seattle (0.04)
- New York > New York County
- New York City (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- California
- San Diego County > San Diego (0.04)
- Los Angeles County > Long Beach (0.04)
- Washington > King County
- Canada > British Columbia
- United States
- Europe
- Austria (0.04)
- Sweden > Stockholm
- Stockholm (0.04)
- Italy > Tuscany
- Florence (0.04)
- France > Hauts-de-France
- Denmark > Capital Region
- Copenhagen (0.04)
- Asia > China
- Africa > Ethiopia
- Addis Ababa > Addis Ababa (0.04)
- Oceania > Australia
- Genre:
- Research Report (0.50)
- Technology: