Entity Framing and Role Portrayal in the News
Mahmoud, Tarek, Xie, Zhuohan, Dimitrov, Dimitar, Nikolaidis, Nikolaos, Silvano, Purificação, Yangarber, Roman, Sharma, Shivam, Sartori, Elisa, Stefanovitch, Nicolas, Martino, Giovanni Da San, Piskorski, Jakub, Nakov, Preslav
–arXiv.org Artificial Intelligence
We introduce a novel multilingual hierarchical corpus annotated for entity framing and role portrayal in news articles. The dataset uses a unique taxonomy inspired by storytelling elements, comprising 22 fine-grained roles, or archetypes, nested within three main categories: protagonist, antagonist, and innocent. Each archetype is carefully defined, capturing nuanced portrayals of entities such as guardian, martyr, and underdog for protagonists; tyrant, deceiver, and bigot for antagonists; and victim, scapegoat, and exploited for innocents. The dataset includes 1,378 recent news articles in five languages (Bulgarian, English, Hindi, European Portuguese, and Russian) focusing on two critical domains of global significance: the Ukraine-Russia War and Climate Change. Over 5,800 entity mentions have been annotated with role labels. This dataset serves as a valuable resource for research into role portrayal and has broader implications for news analysis. We describe the characteristics of the dataset and the annotation process, and we report evaluation results on fine-tuned state-of-the-art multilingual transformers and hierarchical zero-shot learning using LLMs at the level of a document, a paragraph, and a sentence.
arXiv.org Artificial Intelligence
Feb-20-2025
- Country:
- Africa > South Africa (0.14)
- Asia
- China
- India (0.04)
- Japan > Honshū
- Kansai > Osaka Prefecture > Osaka (0.04)
- Middle East > Israel (0.04)
- North Korea (0.14)
- Russia > Far Eastern Federal District
- Republic of Buryatia > Ulan-Ude (0.04)
- Thailand > Bangkok
- Bangkok (0.04)
- Europe
- Croatia > Dubrovnik-Neretva County
- Dubrovnik (0.04)
- Finland > Uusimaa
- Helsinki (0.04)
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Russia > Central Federal District
- Moscow Oblast > Moscow (0.04)
- Sweden > Östergötland County
- Linköping (0.04)
- Ukraine (0.26)
- Croatia > Dubrovnik-Neretva County
- North America
- Canada > Ontario
- Toronto (0.04)
- Dominican Republic (0.04)
- Mexico > Mexico City
- Mexico City (0.04)
- United States
- Illinois > Cook County
- Chicago (0.04)
- Texas > Travis County
- Austin (0.04)
- Illinois > Cook County
- Canada > Ontario
- Genre:
- Research Report (0.50)
- Industry:
- Government
- Military (0.93)
- Regional Government > Europe Government (0.68)
- Law > Civil Rights & Constitutional Law (0.67)
- Law Enforcement & Public Safety > Crime Prevention & Enforcement (1.00)
- Media > News (1.00)
- Government
- Technology: