Uncovering Misattributed Suicide Causes through Annotation Inconsistency Detection in Death Investigation Notes
Wang, Song, Zhou, Yiliang, Han, Ziqiang, Tao, Cui, Xiao, Yunyu, Ding, Ying, Ghosh, Joydeep, Peng, Yifan
–arXiv.org Artificial Intelligence
Data accuracy is essential for scientific research and policy development. The National Violent Death Reporting System (NVDRS) data is widely used for discovering the patterns and causes of death. Recent studies suggested the annotation inconsistencies within the NVDRS and the potential impact on erroneous suicide-cause attributions. We present an empirical Natural Language Processing (NLP) approach to detect annotation inconsistencies and adopt a cross-validation-like paradigm to identify problematic instances. We analyzed 267,804 suicide death incidents between 2003 and 2020 from the NVDRS. Our results showed that incorporating the target state's data into training the suicide-crisis classifier brought an increase of 5.4% to the F-1 score on the target state's test set and a decrease of 1.1% on other states' test set. To conclude, we demonstrated the annotation inconsistencies in NVDRS's death investigation notes, identified problematic instances, evaluated the effectiveness of correcting problematic instances, and eventually proposed an NLP improvement solution.
arXiv.org Artificial Intelligence
Mar-29-2024
- Country:
- North America
- Puerto Rico (0.05)
- Dominican Republic (0.04)
- United States
- Ohio (0.08)
- Colorado (0.06)
- District of Columbia (0.05)
- South Carolina (0.05)
- Maine (0.05)
- Utah (0.05)
- Nevada (0.05)
- Indiana (0.05)
- Missouri (0.05)
- Wisconsin (0.05)
- Arizona (0.05)
- Maryland (0.05)
- Kansas (0.05)
- North Carolina (0.05)
- Michigan (0.05)
- Tennessee (0.05)
- North Dakota (0.05)
- Rhode Island (0.05)
- Iowa (0.05)
- New Mexico (0.05)
- Oklahoma (0.05)
- New Jersey (0.05)
- Vermont (0.05)
- Nebraska (0.05)
- Idaho (0.05)
- Virginia (0.05)
- Illinois (0.05)
- Pennsylvania (0.05)
- Kentucky (0.05)
- Massachusetts (0.05)
- California (0.05)
- Alaska (0.05)
- New Hampshire (0.05)
- West Virginia (0.04)
- Arkansas (0.04)
- Alabama (0.04)
- South Dakota (0.04)
- Connecticut (0.04)
- Hawaii (0.04)
- Wyoming (0.04)
- Mississippi (0.04)
- Montana (0.04)
- Texas > Travis County
- Austin (0.14)
- Minnesota
- Hennepin County > Minneapolis (0.14)
- Olmsted County > Rochester (0.04)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Oregon > Multnomah County
- Portland (0.04)
- New York > New York County
- New York City (0.14)
- Canada > British Columbia
- Europe
- Asia > China
- Shandong Province > Qingdao (0.04)
- North America
- Genre:
- Research Report
- New Finding (1.00)
- Experimental Study (1.00)
- Research Report
- Industry:
- Technology: