Imitation with Neural Density Models
–Neural Information Processing Systems
We propose a new framework for Imitation Learning (IL) via density estimation of the expert's occupancy measure followed by Maximum Occupancy Entropy Reinforcement Learning (RL) using the density as a reward. Our approach maximizes a non-adversarial model-free RL objective that provably lower bounds reverse Kullback-Leibler divergence between occupancy measures of the expert and imitator. We present a practical IL algorithm, Neural Density Imitation (NDI), which obtains state-of-the-art demonstration efficiency on benchmark control tasks.
Neural Information Processing Systems
Apr-25-2026, 06:03:31 GMT
- Country:
- North America (0.28)
- Industry:
- Leisure & Entertainment > Games > Computer Games (0.46)
- Technology: