Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk

Open in new window