DALL-E 2 Fails to Reliably Capture Common Syntactic Processes
Leivada, Evelina, Murphy, Elliot, Marcus, Gary
–arXiv.org Artificial Intelligence
Machine intelligence is increasingly being linked to claims about sentience, language processing, and an ability to comprehend and transform natural language into a range of stimuli. We systematically analyze the ability of DALL-E 2 to capture 8 grammatical phenomena pertaining to compositionality that are widely discussed in linguistics and pervasive in human language: binding principles and coreference, passives, word order, coordination, comparatives, negation, ellipsis, and structural ambiguity. Whereas young children routinely master these phenomena, learning systematic mappings between syntax and semantics, DALL-E 2 is unable to reliably infer meanings that are consistent with the syntax. These results challenge recent claims concerning the capacity of such systems to understand of human language. We make available the full set of test materials as a benchmark for future testing.
arXiv.org Artificial Intelligence
Oct-25-2022
- Country:
- North America > United States
- Texas > Harris County
- Houston (0.04)
- New York > New York County
- New York City (0.04)
- Texas > Harris County
- Europe
- United Kingdom > England
- Oxfordshire > Oxford (0.04)
- Spain > Catalonia
- Tarragona Province > Tarragona (0.04)
- Netherlands > South Holland
- Dordrecht (0.04)
- United Kingdom > England
- North America > United States
- Genre:
- Research Report (1.00)
- Technology: