Unsupervised Behavior Extraction via Random Intent Priors

Oct-9-2025, 03:14:47 GMT–Neural Information Processing Systems

Reward-free data is abundant and contains rich prior knowledge of human behaviors, but it is not well exploited by offline reinforcement learning (RL) algorithms. In this paper, we propose UBER, an unsupervised approach to extract useful behaviors from offline reward-free datasets via diversified rewards.

artificial intelligence, machine learning, reinforcement learning, (15 more...)

Neural Information Processing Systems

Oct-9-2025, 03:14:47 GMT

Conferences PDF

Add feedback

Country:
- Asia
  - Middle East > Jordan (0.04)
  - China (0.04)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning
  - Reinforcement Learning (1.00)
  - Neural Networks > Deep Learning (0.46)

Duplicate Docs Excel Report

Title
a1c8a68e52499c9396854e3f967e37c0-Paper-Conference.pdf

Similar Docs Excel Report more

Title	Similarity	Source
None found