Semi-supervised Predictive Clustering Trees for (Hierarchical) Multi-label Classification
Levatić, Jurica, Ceci, Michelangelo, Kocev, Dragi, Džeroski, Sašo
–arXiv.org Artificial Intelligence
Semi-supervised learning (SSL) is a common approach to learning predictive models using not only labeled examples, but also unlabeled examples. While SSL for the simple tasks of classification and regression has received a lot of attention from the research community, this is not properly investigated for complex prediction tasks with structurally dependent variables. This is the case of multi-label classification and hierarchical multi-label classification tasks, which may require additional information, possibly coming from the underlying distribution in the descriptive space provided by unlabeled examples, to better face the challenging task of predicting simultaneously multiple class labels. In this paper, we investigate this aspect and propose a (hierarchical) multi-label classification method based on semi-supervised learning of predictive clustering trees. We also extend the method towards ensemble learning and propose a method based on the random forest approach. Extensive experimental evaluation conducted on 23 datasets shows significant advantages of the proposed method and its extension with respect to their supervised counterparts. Moreover, the method preserves interpretability and reduces the time complexity of classical tree-based models.
arXiv.org Artificial Intelligence
Jul-19-2022
- Country:
- North America > United States
- Pennsylvania > Philadelphia County
- Philadelphia (0.04)
- Massachusetts > Middlesex County
- Cambridge (0.04)
- California
- San Francisco County > San Francisco (0.14)
- Santa Clara County > Palo Alto (0.04)
- Monterey County > Monterey (0.04)
- Pennsylvania > Philadelphia County
- Europe
- United Kingdom
- England > Bristol (0.04)
- Wales > Ceredigion
- Aberystwyth (0.04)
- Slovenia > Central Slovenia
- Municipality of Ljubljana > Ljubljana (0.04)
- Italy > Apulia
- Bari (0.04)
- United Kingdom
- North America > United States
- Genre:
- Research Report
- New Finding (1.00)
- Experimental Study (0.67)
- Research Report
- Industry:
- Health & Medicine (0.46)
- Technology: