Preference Based Adaptation for Learning Objectives

Mar-27-2025, 01:08:00 GMT–Neural Information Processing Systems

In many real-world learning tasks, it is hard to directly optimize the true performance measures, meanwhile choosing the right surrogate objectives is also difficult. Under this situation, it is desirable to incorporate an optimization of objective process into the learning loop based on weak modeling of the relationship between the true measure and the objective. In this work, we discuss the task of objective adaptation, in which the learner iteratively adapts the learning objective to the underlying true objective based on the preference feedback from an oracle. We show that when the objective can be linearly parameterized, this preference based learning problem can be solved by utilizing the dueling bandit model.

artificial intelligence, machine learning, objective, (18 more...)

Neural Information Processing Systems

Mar-27-2025, 01:08:00 GMT

Conferences PDF

Add feedback

Industry:
- Education (0.34)

Technology:
- Information Technology > Artificial Intelligence
  - Machine Learning > Statistical Learning (0.46)
  - Representation & Reasoning > Optimization (0.68)