Adverse Event Extraction from Discharge Summaries: A New Dataset, Annotation Scheme, and Initial Findings
Guellil, Imane, Andres, Salomé, Anand, Atul, Guthrie, Bruce, Zhang, Huayu, Hasan, Abul, Wu, Honghan, Alex, Beatrice
–arXiv.org Artificial Intelligence
In this work, we present a manually annotated corpus for Adverse Event (AE) extraction from discharge summaries of elderly patients, a population often underrepresented in clinical NLP resources. The dataset includes 14 clinically significant AEs-such as falls, delirium, and intracranial haemorrhage, along with contextual attributes like negation, diagnosis type, and in-hospital occurrence. Uniquely, the annotation schema supports both discontinuous and overlapping entities, addressing challenges rarely tackled in prior work. We evaluate multiple models using FlairNLP across three annotation granularities: fine-grained, coarse-grained, and coarse-grained with negation. While transformer-based models (e.g., BERT-cased) achieve strong performance on document-level coarse-grained extraction (F1 = 0.943), performance drops notably for fine-grained entity-level tasks (e.g., F1 = 0.675), particularly for rare events and complex attributes. These results demonstrate that despite high-level scores, significant challenges remain in detecting underrepresented AEs and capturing nuanced clinical language. Developed within a Trusted Research Environment (TRE), the dataset is available upon request via DataLoch and serves as a robust benchmark for evaluating AE extraction methods and supporting future cross-dataset generalisation.
arXiv.org Artificial Intelligence
Jun-19-2025
- Country:
- Asia
- China > Jiangsu Province (0.04)
- Middle East > Israel (0.04)
- Europe > United Kingdom
- Scotland (0.04)
- North America
- Canada (0.04)
- United States
- Massachusetts (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Virginia (0.04)
- Asia
- Genre:
- Research Report
- Experimental Study (1.00)
- New Finding (1.00)
- Research Report
- Industry:
- Health & Medicine
- Diagnostic Medicine > Imaging (0.68)
- Health Care Providers & Services (1.00)
- Pharmaceuticals & Biotechnology (1.00)
- Therapeutic Area
- Gastroenterology (1.00)
- Immunology (1.00)
- Infections and Infectious Diseases (1.00)
- Internal Medicine (0.94)
- Neurology (1.00)
- Oncology (0.92)
- Orthopedics/Orthopedic Surgery (0.96)
- Health & Medicine
- Technology: