Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Open in new window