Probabilistic Matrix Factorization for Automated Machine Learning

Oct-8-2024, 01:18:07 GMT–Neural Information Processing Systems

In order to achieve state-of-the-art performance, modern machine learning techniques require careful data pre-processing and hyperparameter tuning. Moreover, given the ever increasing number of machine learning models being developed, model selection is becoming increasingly important. Automating the selection and tuning of machine learning pipelines, which can include different data preprocessing methods and machine learning models, has long been one of the goals of the machine learning community. In this paper, we propose to solve this meta-learning task by combining ideas from collaborative filtering and Bayesian optimization. Specifically, we use a probabilistic matrix factorization model to transfer knowledge across experiments performed in hundreds of different datasets and use an acquisition function to guide the exploration of the space of possible pipelines. In our experiments, we show that our approach quickly identifies highperforming pipelines across a wide range of datasets, significantly outperforming the current state-of-the-art.

artificial intelligence, machine learning, pipeline, (16 more...)

Neural Information Processing Systems

Oct-8-2024, 01:18:07 GMT

Conferences PDF

Add feedback

Country:
- North America > United States (0.28)

Genre:
- Research Report > New Finding (0.66)

Technology:
- Information Technology > Artificial Intelligence
  - Machine Learning
    - Neural Networks (1.00)
    - Statistical Learning (1.00)
  - Representation & Reasoning > Optimization (1.00)