Incentivizing Combinatorial Bandit Exploration

Jan-19-2025, 06:26:43 GMT–Neural Information Processing Systems

Consider a bandit algorithm that recommends actions to self-interested users in a recommendation system. The users are free to choose other actions and need to be incentivized to follow the algorithm's recommendations. While the users prefer to exploit, the algorithm can incentivize them to explore by leveraging the information collected from the previous users. All published work on this problem, known as incentivized exploration, focuses on small, unstructured action sets and mainly targets the case when the users' beliefs are independent across actions. However, realistic exploration problems often feature large, structured action sets and highly correlated beliefs.

algorithm, combinatorial semi-bandit, incentivizing combinatorial bandit exploration

Neural Information Processing Systems

Jan-19-2025, 06:26:43 GMT

Conferences Web Page

Add feedback

Technology:
- Information Technology > Artificial Intelligence (0.44)