Contextual Bandits with Smooth Regret: Efficient Learning in Continuous Action Spaces

Open in new window