On the continuity and smoothness of the value function in reinforcement learning and optimal control

Open in new window