Explicable Reward Design for Reinforcement Learning Agents

Neural Information Processing Systems 

A reward function plays the central role during the learning/training process of a reinforcement learning (RL) agent. Given a "task" the agent is expected to perform (i.e., the desired learning outcome), there are typically many different reward specifications under which an optimal policy