Goto

Collaborating Authors

 Country




Verifiable Reinforcement Learning via Policy Extraction

Neural Information Processing Systems

Trajectoriestakenby , left : s 7! left, and right : s 7! rightareshownas dashededges, rededges, andgreenedges, respectively. Let ={ left : s 7! left, right : s 7! right}, andletg( )= Es d( )[g(s, )]bethe 0-1 loss.







Neural Attribution for Semantic Bug-Localization in Student Programs

Neural Information Processing Systems

Most open online courses on programming makeuse ofautomated grading systems tosupport programming assignments and give real-time feedback. These systems usually rely on test results toquantify theprograms' functional correctness.