Goto

Collaborating Authors

 Country


MinimaxValueIntervalforOff-PolicyEvaluation andPolicyOptimization

Neural Information Processing Systems

FunctionApproximation Throughout thepaper,weassume access totwofunction classesQ (S A R)andW (S A R). Todevelop intuition, theyare supposed to modelQฯ€ and wฯ€/ยต, respectively, though most of our main results are stated without assuming any kind of realizability.


MinimaxValueIntervalforOff-PolicyEvaluation andPolicyOptimization

Neural Information Processing Systems

FunctionApproximation Throughout thepaper,weassume access totwofunction classesQ (S A R)andW (S A R). Todevelop intuition, theyare supposed to modelQฯ€ and wฯ€/ยต, respectively, though most of our main results are stated without assuming any kind of realizability.



InterventionalFew-ShotLearning

Neural Information Processing Systems

It is worth noting that the contribution of IFSL is orthogonal to existing fine-tuning and meta-learning based FSL methods, hence IFSL can improveallofthem,achievinganew1-/5-shot state-of-the-art onminiImageNet, tieredImageNet, andcross-domain CUB.