Goto

Collaborating Authors

 Technology










f2d887e01a80e813d9080038decbbabb-AuthorFeedback.pdf

Neural Information Processing Systems

Thank you for your detailed reviews and comments. Wehope our clarifications, which we will include in the final1 version of the paper, will strengthen your confidence in the novelty and significance of our results. Both speed up DPP sampling given a polynomial time pre-processing step.


Verifiable Reinforcement Learning via Policy Extraction

Neural Information Processing Systems

Trajectoriestakenby , left : s 7! left, and right : s 7! rightareshownas dashededges, rededges, andgreenedges, respectively. Let ={ left : s 7! left, right : s 7! right}, andletg( )= Es d( )[g(s, )]bethe 0-1 loss.