Embrace Divergence for Richer Insights: A Multi-document Summarization Benchmark and a Case Study on Summarizing Diverse Information from News Articles
Huang, Kung-Hsiang, Laban, Philippe, Fabbri, Alexander R., Choubey, Prafulla Kumar, Joty, Shafiq, Xiong, Caiming, Wu, Chien-Sheng
–arXiv.org Artificial Intelligence
Previous research in multi-document news summarization has typically concentrated on collating information that all sources agree upon. However, to our knowledge, the summarization of diverse information dispersed across multiple articles about an event has not been previously investigated. The latter imposes a different set of challenges for a summarization model. In this paper, we propose a new task of summarizing diverse information encountered in multiple news articles encompassing the same event. To facilitate this task, we outlined a data collection schema for identifying diverse information and curated a dataset named DiverseSumm. The dataset includes 245 news stories, with each story comprising 10 news articles and paired with a human-validated reference. Moreover, we conducted a comprehensive analysis to pinpoint the position and verbosity biases when utilizing Large Language Model (LLM)-based metrics for evaluating the coverage and faithfulness of the summaries, as well as their correlation with human assessments. We applied our findings to study how LLMs summarize multiple news articles by analyzing which type of diverse information LLMs are capable of identifying. Our analyses suggest that despite the extraordinary capabilities of LLMs in single-document summarization, the proposed task remains a complex challenge for them mainly due to their limited coverage, with GPT-4 only able to cover less than 40% of the diverse information on average.
arXiv.org Artificial Intelligence
Sep-17-2023
- Country:
- South America > Argentina
- Pampas > Buenos Aires F.D. > Buenos Aires (0.04)
- North America
- Dominican Republic (0.04)
- United States
- Colorado (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Illinois > Champaign County
- Urbana (0.04)
- Maryland > Montgomery County
- Gaithersburg (0.04)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Oregon > Multnomah County
- Portland (0.04)
- New Mexico > Santa Fe County
- Santa Fe (0.04)
- Washington > King County
- Seattle (0.04)
- New York > New York County
- New York City (0.04)
- Canada > Ontario
- Toronto (0.04)
- Europe
- Poland (0.46)
- Russia (0.04)
- Ukraine (0.04)
- Germany > Berlin (0.04)
- France (0.04)
- United Kingdom (0.04)
- Italy > Tuscany
- Florence (0.04)
- Sweden > Vaestra Goetaland
- Gothenburg (0.04)
- Croatia > Dubrovnik-Neretva County
- Dubrovnik (0.04)
- Asia
- Russia (0.46)
- Taiwan > Taiwan Province
- Taipei (0.04)
- Myanmar > Tanintharyi Region
- Dawei (0.04)
- Middle East > UAE
- Abu Dhabi Emirate > Abu Dhabi (0.04)
- Japan > Honshū
- Kansai > Osaka Prefecture > Osaka (0.04)
- China > Beijing
- Beijing (0.04)
- South America > Argentina
- Genre:
- Research Report > New Finding (1.00)
- Industry:
- Health & Medicine > Pharmaceuticals & Biotechnology (1.00)
- Media > News (0.67)
- Government
- Technology: