Asking Again and Again: Exploring LLM Robustness to Repeated Questions
–arXiv.org Artificial Intelligence
This study examines whether large language models (LLMs), such as ChatGPT, specifically the latest GPT-4o-mini, exhibit sensitivity to repeated prompts and whether repeating a question can improve response accuracy. We hypothesize that reiterating a question within a single prompt might enhance the model's focus on key elements of the query. To test this, we evaluate ChatGPT's performance on a large sample of two reading comprehension datasets under both open-book and closed-book settings, varying the repetition of each question to 1, 3, or 5 times per prompt. Our findings indicate that the model does not demonstrate sensitivity to repeated questions, highlighting its robustness and consistency in this context.
arXiv.org Artificial Intelligence
Dec-10-2024
- Country:
- North America
- United States
- Texas > Travis County
- Austin (0.04)
- Florida > Miami-Dade County
- Miami (0.05)
- Colorado > Boulder County
- Boulder (0.04)
- Texas > Travis County
- Canada > Ontario
- Toronto (0.04)
- United States
- Europe
- Middle East > Malta
- Eastern Region > Northern Harbour District > St. Julian's (0.04)
- Denmark > Capital Region
- Copenhagen (0.04)
- Belgium > Brussels-Capital Region
- Brussels (0.04)
- Middle East > Malta
- Asia
- Indonesia > Bali (0.04)
- Thailand > Bangkok
- Bangkok (0.05)
- Middle East > UAE
- Abu Dhabi Emirate > Abu Dhabi (0.04)
- North America
- Genre:
- Research Report > New Finding (1.00)
- Industry:
- Government (0.49)
- Technology: