An advantage based policy transfer algorithm for reinforcement learning with metrics of transferability

Open in new window