A Taxonomy of Challenges to Curating Fair Datasets
Zhao, Dora, Scheuerman, Morgan Klaus, Chitre, Pooja, Andrews, Jerone T. A., Panagiotidou, Georgia, Walker, Shawn, Pine, Kathleen H., Xiang, Alice
–arXiv.org Artificial Intelligence
Despite extensive efforts to create fairer machine learning (ML) datasets, there remains a limited understanding of the practical aspects of dataset curation. Drawing from interviews with 30 ML dataset curators, we present a comprehensive taxonomy of the challenges and trade-offs encountered throughout the dataset curation lifecycle. Our findings underscore overarching issues within the broader fairness landscape that impact data curation. We conclude with recommendations aimed at fostering systemic changes to better facilitate fair dataset curation practices.
arXiv.org Artificial Intelligence
Jun-10-2024
- Country:
- Africa > West Africa (0.04)
- Asia
- India (0.04)
- Japan (0.04)
- Middle East > Jordan (0.04)
- Europe
- France (0.04)
- Northern Europe (0.04)
- United Kingdom > England
- Oxfordshire > Oxford (0.04)
- Western Europe (0.04)
- North America > United States
- Arizona (0.04)
- California > Santa Clara County
- Palo Alto (0.04)
- Illinois (0.04)
- Genre:
- Personal > Interview (1.00)
- Questionnaire & Opinion Survey (1.00)
- Research Report > New Finding (1.00)
- Industry:
- Education (0.67)
- Government (1.00)
- Health & Medicine (1.00)
- Information Technology > Security & Privacy (1.00)
- Law (1.00)
- Technology:
- Information Technology
- Artificial Intelligence
- Issues > Social & Ethical Issues (1.00)
- Machine Learning (1.00)
- Natural Language (1.00)
- Vision (1.00)
- Communications > Social Media (1.00)
- Data Science > Data Quality (1.00)
- Security & Privacy (1.00)
- Artificial Intelligence
- Information Technology