Goto

Collaborating Authors

 Country



a59afb1b7d82ec353921a55c579ee26d-Paper.pdf

Neural Information Processing Systems

Towardsthis end,wedesigned apreconditioned gradient solverforkernelmethods exploiting both GPU acceleration and parallelization with multiple GPUs, implementing out-of-core variants of common linear algebra operations to guarantee optimal hardware utilization.



LifelongPolicyGradientLearning ofFactoredPolicies forFasterTrainingWithoutForgetting

Neural Information Processing Systems

We provide a novel method for lifelong policy gradient learning that trains lifelong function approximators directly via policygradients, allowing the agent to benefit from accumulated knowledge throughout the entire training process.