Model predictive control-based value estimation for efficient reinforcement learning

Open in new window