Rebounding Bandits for Modeling Satiation Effects

Oct-9-2024, 18:04:36 GMT–Neural Information Processing Systems

Psychological research shows that enjoyment of many goods is subject to satiation, with short-term satisfaction declining after repeated exposures to the same item. Nevertheless, proposed algorithms for powering recommender systems seldom model these dynamics, instead proceeding as though user preferences were fixed in time. In this work, we introduce rebounding bandits, a multi-armed bandit setup, where satiation dynamics are modeled as time-invariant linear dynamical systems. Expected rewards for each arm decline monotonically with consecutive exposures and rebound towards the initial reward whenever that arm is not pulled. Unlike classical bandit algorithms, methods for tackling rebounding bandits must plan ahead and model-based methods rely on estimating the parameters of the satiation dynamics.

algorithm, modeling satiation effect, satiation dynamic, (1 more...)

Neural Information Processing Systems

Oct-9-2024, 18:04:36 GMT

Conferences Web Page

Add feedback

Technology:
- Information Technology
  - Data Science > Data Mining
    - Big Data (0.64)
  - Artificial Intelligence > Representation & Reasoning
    - Personal Assistant Systems (0.64)