LingBench++: A Linguistically-Informed Benchmark and Reasoning Framework for Multi-Step and Cross-Cultural Inference with LLMs
Lian, Da-Chen, Huang, Ri-Sheng, Chen, Pin-Er, Lim, Chunki, Lin, You-Kuan, Tseng, Guan-Yu, Yang, Zi-Cheng, Lin, Zhen-Yu, Chen, Pin-Cheng, Hsieh, Shu-Kai
–arXiv.org Artificial Intelligence
We propose LingBench++, a linguistically-informed benchmark and reasoning framework designed to evaluate large language models (LLMs) on complex linguistic tasks inspired by the International Linguistics Olympiad (IOL). Unlike prior benchmarks that focus solely on final answer accuracy, LingBench++ provides structured reasoning traces, stepwise evaluation protocols, and rich typological metadata across over 90 low-resource and cross-cultural languages. We further develop a multi-agent architecture integrating grammatical knowledge retrieval, tool-augmented reasoning, and deliberate hypothesis testing. Through systematic comparisons of baseline and our proposed agentic models, we demonstrate that models equipped with external knowledge sources and iterative reasoning outperform single-pass approaches in both accuracy and interpretability. LingBench++ offers a comprehensive foundation for advancing linguistically grounded, culturally informed, and cognitively plausible reasoning in LLMs.
arXiv.org Artificial Intelligence
Jul-25-2025
- Country:
- Africa
- Democratic Republic of the Congo > Kinshasa Province
- Kinshasa (0.04)
- Kenya (0.04)
- Niger (0.04)
- Tanzania (0.04)
- Zambia (0.04)
- Democratic Republic of the Congo > Kinshasa Province
- Asia
- Indonesia > Bali (0.04)
- Myanmar (0.04)
- Philippines > Luzon
- Ilocos Region > Province of Pangasinan (0.04)
- Taiwan (0.04)
- Thailand > Bangkok
- Bangkok (0.04)
- Europe
- Belgium (0.04)
- Bulgaria > Sofia City Province
- Sofia (0.04)
- Italy > Tuscany
- Florence (0.04)
- Middle East > Malta
- Eastern Region > Northern Harbour District > St. Julian's (0.04)
- Sweden (0.04)
- North America
- Mexico > Mexico City
- Mexico City (0.04)
- United States (0.04)
- Mexico > Mexico City
- Africa
- Genre:
- Research Report > New Finding (0.93)
- Technology: