On Language Models' Sensitivity to Suspicious Coincidences
Padmanabhan, Sriram, Misra, Kanishka, Mahowald, Kyle, Choi, Eunsol
–arXiv.org Artificial Intelligence
Humans are sensitive to suspicious coincidences when generalizing inductively over data, as they make assumptions as to how the data was sampled. This results in smaller, more specific hypotheses being favored over more general ones. For instance, when provided the set {Austin, Dallas, Houston}, one is more likely to think that this is sampled from "Texas Cities" over "US Cities" even though both are compatible. Suspicious coincidence is strongly connected to pragmatic reasoning, and can serve as a testbed to analyze systems on their sensitivity towards the communicative goals of the task (i.e., figuring out the true category underlying the data). In this paper, we analyze whether suspicious coincidence effects are reflected in language models' (LMs) behavior. We do so in the context of two domains: 1) the number game, where humans made judgments of whether a number (e.g., 4) fits a list of given numbers (e.g., 16, 32, 2); and 2) by extending the number game setup to prominent cities. For both domains, the data is compatible with multiple hypotheses and we study which hypothesis is most consistent with the models' behavior. On analyzing five models, we do not find strong evidence for suspicious coincidences in LMs' zero-shot behavior. However, when provided access to the hypotheses space via chain-of-thought or explicit prompting, LMs start to show an effect resembling suspicious coincidences, sometimes even showing effects consistent with humans. Our study suggests that inductive reasoning behavior in LMs can be enhanced with explicit access to the hypothesis landscape.
arXiv.org Artificial Intelligence
Apr-15-2025
- Country:
- Africa
- Democratic Republic of the Congo > Kinshasa Province
- Kinshasa (0.04)
- Kenya > Nairobi City County
- Nairobi (0.04)
- Middle East > Egypt
- Cairo Governorate > Cairo (0.04)
- Sudan
- Khartoum (0.04)
- Khartoum State > Khartoum (0.04)
- Democratic Republic of the Congo > Kinshasa Province
- Asia
- Japan > Honshū
- Kantō > Tokyo Metropolis Prefecture > Tokyo (0.04)
- India > Maharashtra
- Mumbai (0.04)
- Pakistan > Sindh
- Karachi Division > Karachi (0.04)
- Cambodia > Phnom Penh Province
- Phnom Penh (0.04)
- Middle East
- Saudi Arabia > Riyadh Province
- Riyadh (0.04)
- UAE > Dubai Emirate
- Dubai (0.04)
- Saudi Arabia > Riyadh Province
- China
- Beijing > Beijing (0.04)
- Guangdong Province > Guangzhou (0.04)
- South Korea
- Thailand > Bangkok
- Bangkok (0.04)
- Bangladesh > Dhaka Division
- Dhaka District > Dhaka (0.04)
- Indonesia > Java
- Philippines > Luzon
- National Capital Region > City of Manila (0.04)
- Japan > Honshū
- Europe
- France > Provence-Alpes-Côte d'Azur
- Bouches-du-Rhône > Marseille (0.04)
- Russia > Central Federal District
- Moscow Oblast > Moscow (0.04)
- Spain > Galicia
- Madrid (0.04)
- France > Provence-Alpes-Côte d'Azur
- North America
- Canada > Quebec
- Montreal (0.04)
- Mexico > Mexico City
- Mexico City (0.04)
- United States
- California
- Los Angeles County > Los Angeles (0.04)
- San Francisco County > San Francisco (0.04)
- Illinois > Cook County
- Chicago (0.04)
- Massachusetts (0.04)
- Missouri > Jackson County
- Kansas City (0.04)
- Texas (0.24)
- Wisconsin > Milwaukee County
- Milwaukee (0.04)
- California
- Canada > Quebec
- Oceania
- South America
- Argentina > Pampas
- Buenos Aires F.D. > Buenos Aires (0.04)
- Brazil
- Rio de Janeiro > Rio de Janeiro (0.04)
- São Paulo (0.04)
- Argentina > Pampas
- Africa
- Genre:
- Research Report > New Finding (1.00)