Solving Label Variation in Scientific Information Extraction via Multi-Task Learning
Pham, Dong, Ho, Xanh, Ha, Quang-Thuy, Aizawa, Akiko
–arXiv.org Artificial Intelligence
Scientific Information Extraction (ScientificIE) is a critical task that involves the identification of scientific entities and their relationships. The complexity of this task is compounded by the necessity for domain-specific knowledge and the limited availability of annotated data. Two of the most popular datasets for ScientificIE are SemEval-2018 Task-7 and SciERC. They have overlapping samples and differ in their annotation schemes, which leads to conflicts. In this study, we first introduced a novel approach based on multi-task learning to address label variations. We then proposed a soft labeling technique that converts inconsistent labels into probabilistic distributions. The experimental results demonstrated that the proposed method can enhance the model robustness to label noise and improve the end-to-end performance in both ScientificIE tasks. The analysis revealed that label variations can be particularly effective in handling ambiguous instances. Furthermore, the richness of the information captured by label variations can potentially reduce data size requirements. The findings highlight the importance of releasing variation labels and promote future research on other tasks in other domains. Overall, this study demonstrates the effectiveness of multi-task learning and the potential of label variations to enhance the performance of ScientificIE.
arXiv.org Artificial Intelligence
Dec-25-2023
- Country:
- North America
- Dominican Republic (0.04)
- United States
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Minnesota > Hennepin County
- Europe
- Portugal > Lisbon
- Lisbon (0.04)
- Middle East > Malta
- Port Region > Southern Harbour District > Valletta (0.04)
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Belgium > Brussels-Capital Region
- Brussels (0.04)
- Portugal > Lisbon
- Asia
- China > Hong Kong (0.04)
- Vietnam > Hanoi
- Hanoi (0.04)
- Middle East
- UAE > Abu Dhabi Emirate
- Abu Dhabi (0.04)
- Israel > Haifa District
- Haifa (0.04)
- UAE > Abu Dhabi Emirate
- Japan > Honshū
- Kantō > Tokyo Metropolis Prefecture > Tokyo (0.14)
- North America
- Genre:
- Overview > Innovation (0.34)
- Research Report
- Promising Solution (0.48)
- New Finding (0.34)
- Technology: