Semi-supervised Predictive Clustering Trees for (Hierarchical) Multi-label Classification

Levatić, Jurica; Ceci, Michelangelo; Kocev, Dragi; Džeroski, Sašo

Computer Science > Machine Learning

arXiv:2207.09237v1 (cs)

[Submitted on 19 Jul 2022 (this version), latest version 30 Mar 2024 (v2)]

Title:Semi-supervised Predictive Clustering Trees for (Hierarchical) Multi-label Classification

Authors:Jurica Levatić, Michelangelo Ceci, Dragi Kocev, Sašo Džeroski

View PDF

Abstract:Semi-supervised learning (SSL) is a common approach to learning predictive models using not only labeled examples, but also unlabeled examples. While SSL for the simple tasks of classification and regression has received a lot of attention from the research community, this is not properly investigated for complex prediction tasks with structurally dependent variables. This is the case of multi-label classification and hierarchical multi-label classification tasks, which may require additional information, possibly coming from the underlying distribution in the descriptive space provided by unlabeled examples, to better face the challenging task of predicting simultaneously multiple class labels.
In this paper, we investigate this aspect and propose a (hierarchical) multi-label classification method based on semi-supervised learning of predictive clustering trees. We also extend the method towards ensemble learning and propose a method based on the random forest approach. Extensive experimental evaluation conducted on 23 datasets shows significant advantages of the proposed method and its extension with respect to their supervised counterparts. Moreover, the method preserves interpretability and reduces the time complexity of classical tree-based models.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2207.09237 [cs.LG]
	(or arXiv:2207.09237v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2207.09237

Submission history

From: Jurica Levatić [view email]
[v1] Tue, 19 Jul 2022 12:49:00 UTC (8,736 KB)
[v2] Sat, 30 Mar 2024 11:55:26 UTC (2,565 KB)

Computer Science > Machine Learning

Title:Semi-supervised Predictive Clustering Trees for (Hierarchical) Multi-label Classification

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Semi-supervised Predictive Clustering Trees for (Hierarchical) Multi-label Classification

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators