Generalizing Goal-Conditioned Reinforcement Learning with Variational Causal Reasoning

Jan-18-2025, 11:46:43 GMT–Neural Information Processing Systems

As a pivotal component to attaining generalizable solutions in human intelligence, reasoning provides great potential for reinforcement learning (RL) agents' generalization towards varied goals by summarizing part-to-whole arguments and discovering cause-and-effect relations. However, how to discover and represent causalities remains a huge gap that hinders the development of causal RL. In this paper, we augment Goal-Conditioned RL (GCRL) with Causal Graph (CG), a structure built upon the relation between objects and events. We novelly formulate the GCRL problem into variational likelihood maximization with CG as latent variables. To optimize the derived objective, we propose a framework with theoretical performance guarantees that alternates between two steps: using interventional data to estimate the posterior of CG; using CG to learn generalizable models and interpretable policies.

artificial intelligence, generalizing goal-conditioned reinforcement learning, machine learning, (2 more...)

Neural Information Processing Systems

Jan-18-2025, 11:46:43 GMT

Conferences Web Page

Add feedback

Technology:
- Information Technology > Artificial Intelligence
  - Machine Learning > Reinforcement Learning (0.99)
  - Representation & Reasoning > Model-Based Reasoning (0.74)