How Toxicity Classifiers and Large Language Models Respond to Ableism
Phutane, Mahika, Seelam, Ananya, Vashistha, Aditya
–arXiv.org Artificial Intelligence
People with disabilities (PwD) regularly encounter ableist hate and microaggressions online. While online platforms use machine learning models to moderate online harm, there is little research investigating how these models interact with ableism. In this paper, we curated a dataset of 100 social media comments targeted towards PwD, and recruited 160 participants to rate and explain how toxic and ableist these comments were. We then prompted state-of-the art toxicity classifiers (TCs) and large language models (LLMs) to rate and explain the harm. Our analysis revealed that TCs and LLMs rated toxicity significantly lower than PwD, but LLMs rated ableism generally on par with PwD. However, ableism explanations by LLMs overlooked emotional harm, and lacked specificity and acknowledgement of context, important facets of PwD explanations. Going forward, we discuss challenges in designing disability-aware toxicity classifiers, and advocate for the shift from ableism detection to ableism interpretation and explanation.
arXiv.org Artificial Intelligence
Oct-4-2024
- Country:
- Africa > Uganda (0.04)
- Oceania > Australia (0.04)
- South America > Paraguay
- North America
- United States
- Virginia (0.04)
- Texas > Travis County
- Austin (0.04)
- Nevada > Clark County
- Las Vegas (0.04)
- Colorado > Denver County
- Denver (0.04)
- Massachusetts > Suffolk County
- Boston (0.04)
- Hawaii > Honolulu County
- Honolulu (0.04)
- Illinois > Cook County
- Chicago (0.04)
- New Jersey > Hudson County
- Hoboken (0.04)
- New York > New York County
- New York City (0.04)
- Canada
- United States
- Europe
- Germany > Hamburg (0.04)
- Czechia > Prague (0.04)
- United Kingdom
- Scotland > City of Glasgow
- Glasgow (0.04)
- England > Cambridgeshire
- Cambridge (0.04)
- Scotland > City of Glasgow
- Spain > Valencian Community
- Valencia Province > Valencia (0.04)
- Italy
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Greece > Attica
- Athens (0.04)
- France
- Provence-Alpes-Côte d'Azur > Bouches-du-Rhône
- Marseille (0.04)
- Auvergne-Rhône-Alpes > Lyon
- Lyon (0.04)
- Provence-Alpes-Côte d'Azur > Bouches-du-Rhône
- Asia
- India (0.04)
- Indonesia > Bali (0.04)
- Taiwan > Taiwan Province
- Taipei (0.04)
- Singapore > Central Region
- Singapore (0.04)
- Middle East > UAE
- Abu Dhabi Emirate > Abu Dhabi (0.04)
- Genre:
- Overview (0.93)
- Research Report
- New Finding (1.00)
- Experimental Study (0.93)
- Industry:
- Law (1.00)
- Education (1.00)
- Media (0.93)
- Information Technology (0.67)
- Health & Medicine > Therapeutic Area
- Neurology (1.00)
- Government > Regional Government
- Technology: