Identifying Factual Inconsistencies in Summaries: Grounding Model Inference via Task Taxonomy
Xu, Liyan, Su, Zhenlin, Yu, Mo, Xu, Jin, Choi, Jinho D., Zhou, Jie, Liu, Fei
–arXiv.org Artificial Intelligence
Factual inconsistencies pose a significant hurdle for the faithful summarization by generative models. While a major direction to enhance inconsistency detection is to derive stronger Natural Language Inference (NLI) models, we propose an orthogonal aspect that underscores the importance of incorporating task-specific taxonomy into the inference. To this end, we consolidate key error types of inconsistent facts in summaries, and incorporate them to facilitate both the zero-shot and supervised paradigms of LLMs. Extensive experiments on ten datasets of five distinct domains suggest that, zero-shot LLM inference could benefit from the explicit solution space depicted by the error type taxonomy, and achieves state-of-the-art performance overall, surpassing specialized non-LLM baselines, as well as recent LLM baselines. We further distill models that fuse the taxonomy into parameters through our designed prompt completions and supervised training strategies, efficiently substituting state-of-the-art zero-shot inference with much larger LLMs.
arXiv.org Artificial Intelligence
Jun-19-2024
- Country:
- North America
- Dominican Republic (0.04)
- United States
- Washington > King County
- Seattle (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Washington > King County
- Mexico > Mexico City
- Mexico City (0.04)
- Canada > Ontario
- Toronto (0.05)
- Europe
- Asia
- Singapore (0.04)
- Middle East > UAE
- Abu Dhabi Emirate > Abu Dhabi (0.04)
- China > Guangdong Province
- Guangzhou (0.04)
- North America
- Genre:
- Research Report (0.82)
- Technology: