Bandit-Feedback Online Multiclass Classification: Variants and Tradeoffs

Neural Information Processing Systems 

Consider the domain of multiclass classification within the adversarial online setting. What is the price of relying on bandit feedback as opposed to full information? To what extent can an adaptive adversary amplify the loss compared to an oblivious one? To what extent can a randomized learner reduce the loss compared to a deterministic one?

Similar Docs  Excel Report  more

TitleSimilaritySource
None found