Goto

Collaborating Authors

 Country



ALawofIteratedLogarithmforMulti-Agent ReinforcementLearning

Neural Information Processing Systems

In contrast, the mathematics needed to analyze such schemes is what forms the focus in Stochastic Approximation (SA) theory [2, 4]. More generally, SA refers to an iterative scheme that helps find zeroes or optimal points of a function, for which only noisy evaluationsarepossible.





AUnifyingPost-Processing Frameworkfor Multi-ObjectiveLearn-to-DeferProblems

Neural Information Processing Systems

Inthisparadigm, wepermit thesystem to defer a subset of its tasks to the expert. Although there are currently systems that follow this paradigm and are designed to optimize the accuracy of the final human-AI team, the general methodology for developing such systems under a set of constraints (e.g., algorithmic fairness, expert intervention budget, defer of anomaly,etc.)