Post-hoc analysis of Arabic transformer models
Abdelali, Ahmed, Durrani, Nadir, Dalvi, Fahim, Sajjad, Hassan
–arXiv.org Artificial Intelligence
Arabic is a Semitic language which is widely spoken with many dialects. Given the success of pre-trained language models, many transformer models trained on Arabic and its dialects have surfaced. While there have been an extrinsic evaluation of these models with respect to downstream NLP tasks, no work has been carried out to analyze and compare their internal representations. We probe how linguistic information is encoded in the transformer models, trained on different Arabic dialects. Figure 1: Data regimes of various pre-trained We perform a layer and neuron analysis Transformer models of Arabic on the models using morphological tagging tasks for different dialects of Arabic and a dialectal identification task.
arXiv.org Artificial Intelligence
Oct-18-2022
- Country:
- North America
- Dominican Republic (0.04)
- United States
- Louisiana (0.04)
- California (0.04)
- Washington > King County
- Seattle (0.04)
- Texas > Travis County
- Austin (0.04)
- New York > New York County
- New York City (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Hawaii > Honolulu County
- Honolulu (0.04)
- Canada > British Columbia
- Europe
- Italy (0.04)
- Slovenia (0.04)
- United Kingdom > England
- Cambridgeshire > Cambridge (0.04)
- Ukraine > Kyiv Oblast
- Kyiv (0.04)
- France > Provence-Alpes-Côte d'Azur
- Bouches-du-Rhône > Marseille (0.04)
- Belgium > Brussels-Capital Region
- Brussels (0.04)
- Asia
- China > Hong Kong (0.04)
- Nepal > Bagmati Province
- Kathmandu District > Kathmandu (0.04)
- Middle East > Qatar
- Japan > Kyūshū & Okinawa
- Kyūshū > Miyazaki Prefecture > Miyazaki (0.04)
- North America
- Genre:
- Research Report (0.82)
- Technology: