QuanTemp: A real-world open-domain benchmark for fact-checking numerical claims
V, Venktesh, Anand, Abhijit, Anand, Avishek, Setty, Vinay
–arXiv.org Artificial Intelligence
Automated fact checking has gained immense interest to tackle the growing misinformation in the digital era. Existing systems primarily focus on synthetic claims on Wikipedia, and noteworthy progress has also been made on real-world claims. In this work, we release QuanTemp, a diverse, multi-domain dataset focused exclusively on numerical claims, encompassing temporal, statistical and diverse aspects with fine-grained metadata and an evidence collection without leakage. This addresses the challenge of verifying real-world numerical claims, which are complex and often lack precise information, not addressed by existing works that mainly focus on synthetic claims. We evaluate and quantify the limitations of existing solutions for the task of verifying numerical claims. We also evaluate claim decomposition based methods, numerical understanding based models and our best baselines achieves a macro-F1 of 58.32. This demonstrates that QuanTemp serves as a challenging evaluation set for numerical claim verification.
arXiv.org Artificial Intelligence
May-1-2024
- Country:
- Oceania > Australia (0.04)
- South America
- Paraguay > Asunción
- Asunción (0.04)
- Colombia > Meta Department
- Villavicencio (0.04)
- Paraguay > Asunción
- North America
- United States
- District of Columbia > Washington (0.05)
- Ohio (0.04)
- Texas > Travis County
- Austin (0.04)
- Indiana > Marion County
- Indianapolis (0.04)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Oregon > Multnomah County
- Portland (0.04)
- New Mexico > Santa Fe County
- Santa Fe (0.04)
- Washington > King County
- Seattle (0.04)
- New York > New York County
- New York City (0.04)
- Canada
- Ontario > Toronto (0.04)
- Quebec > Montreal (0.04)
- British Columbia > Metro Vancouver Regional District
- Vancouver (0.14)
- United States
- Europe
- Ukraine (0.04)
- Ireland (0.04)
- Spain > Valencian Community
- Valencia Province > Valencia (0.04)
- Portugal > Lisbon
- Lisbon (0.04)
- France
- Provence-Alpes-Côte d'Azur > Bouches-du-Rhône
- Marseille (0.04)
- Auvergne-Rhône-Alpes > Lyon
- Lyon (0.04)
- Provence-Alpes-Côte d'Azur > Bouches-du-Rhône
- Germany > Lower Saxony
- Hanover (0.04)
- Netherlands > South Holland
- Delft (0.04)
- Norway > Western Norway
- Belgium > Brussels-Capital Region
- Brussels (0.04)
- Asia
- China > Hong Kong (0.04)
- Singapore (0.04)
- Philippines (0.04)
- Indonesia > Bali (0.04)
- India (0.04)
- Middle East
- Jordan (0.04)
- UAE > Abu Dhabi Emirate
- Abu Dhabi (0.04)
- Africa
- South Africa (0.04)
- Nigeria (0.04)
- Genre:
- Research Report > Experimental Study (0.68)
- Industry:
- Government (1.00)
- Education (0.67)
- Media > News (0.66)
- Health & Medicine > Therapeutic Area (0.46)
- Technology: