Goto

Collaborating Authors

 Country






EmergentComplexityandZero-shotTransfervia UnsupervisedEnvironmentDesign

Neural Information Processing Systems

Awide range ofreinforcement learning (RL) problems --including robustness, transfer learning, unsupervised RL, and emergent complexity -- require specifying a distribution of tasks or environments in which a policy will be trained.


79ec2a4246feb2126ecf43c4a4418002-Paper.pdf

Neural Information Processing Systems

Weformulate the decoding process asanoptimization problem which allows for multiple attributesweaimtocontrol tobeeasilyincorporated asdifferentiable constraints to the optimization. By relaxing this discrete optimization to a continuous one, we make use of Lagrangian multipliers and gradient-descent based techniques to generate the desired text.