A Pipeline For Discourse Circuits From CCG
Liu, Jonathon, Shaikh, Razin A., Rodatz, Benjamin, Yeung, Richie, Coecke, Bob
–arXiv.org Artificial Intelligence
There is a significant disconnect between linguistic theory and modern NLP practice, which relies heavily on inscrutable black-box architectures. DisCoCirc is a newly proposed model for meaning that aims to bridge this divide, by providing neuro-symbolic models that incorporate linguistic structure. DisCoCirc represents natural language text as a `circuit' that captures the core semantic information of the text. These circuits can then be interpreted as modular machine learning models. Additionally, DisCoCirc fulfils another major aim of providing an NLP model that can be implemented on near-term quantum computers. In this paper we describe a software pipeline that converts English text to its DisCoCirc representation. The pipeline achieves coverage over a large fragment of the English language. It relies on Combinatory Categorial Grammar (CCG) parses of the input text as well as coreference resolution information. This semantic and syntactic information is used in several steps to convert the text into a simply-typed $\lambda$-calculus term, and then into a circuit diagram. This pipeline will enable the application of the DisCoCirc framework to NLP tasks, using both classical and quantum approaches.
arXiv.org Artificial Intelligence
Nov-29-2023
- Country:
- Asia > Middle East
- Israel (0.04)
- Europe
- Ireland
- Connaught > County Galway
- Galway (0.04)
- Leinster > County Dublin
- Dublin (0.04)
- Connaught > County Galway
- Netherlands > South Holland
- Dordrecht (0.04)
- Spain > Valencian Community
- Valencia Province > Valencia (0.04)
- United Kingdom
- England
- Cambridgeshire > Cambridge (0.04)
- Oxfordshire > Oxford (0.14)
- Scotland > City of Edinburgh
- Edinburgh (0.04)
- England
- Ireland
- North America > United States
- Massachusetts > Middlesex County
- Cambridge (0.04)
- New York > New York County
- New York City (0.04)
- Washington > King County
- Seattle (0.04)
- Massachusetts > Middlesex County
- Asia > Middle East
- Genre:
- Research Report (0.40)
- Industry:
- Education (0.46)
- Law (0.46)
- Transportation (0.48)
- Technology: