DeepMind Study Resolves Delusions in Sequence Models for Interaction and Control
Large-scale language models such as transformers have become the de facto standard for a wide range of natural language processing (NLP) tasks. Despite their apparent linguistic savvy, such sequence models are known to lack a real understanding of the cause and effect of their actions, which can lead to false decisions due to auto-suggestive delusions. In the new paper Shaking the Foundations: Delusions in Sequence Models for Interaction and Control, a DeepMind research team explores the origin of these mismatches and addresses the problem by treating actions as causal interventions. The team shows that a system can learn to condition or intervene on data by training with the use of factual and counterfactual error signals respectively. Sequence models are updated based on collected data, and these updates will differ depending on whether the data was generated by the model itself (i.e.
Oct-29-2021, 14:43:47 GMT
- Technology: