A Simple but Effective Closed-form Solution for Extreme Multi-label Learning

Onishi, Kazuma; Hayashi, Katsuhiko

Computer Science > Information Retrieval

arXiv:2501.10179 (cs)

[Submitted on 17 Jan 2025]

Title:A Simple but Effective Closed-form Solution for Extreme Multi-label Learning

Authors:Kazuma Onishi, Katsuhiko Hayashi

View PDF HTML (experimental)

Abstract:Extreme multi-label learning (XML) is a task of assigning multiple labels from an extremely large set of labels to each data instance. Many current high-performance XML models are composed of a lot of hyperparameters, which complicates the tuning process. Additionally, the models themselves are adapted specifically to XML, which complicates their reimplementation. To remedy this problem, we propose a simple method based on ridge regression for XML. The proposed method not only has a closed-form solution but also is composed of a single hyperparameter. Since there are no precedents on applying ridge regression to XML, this paper verified the performance of the method by using various XML benchmark datasets. Furthermore, we enhanced the prediction of low-frequency labels in XML, which hold informative content. This prediction is essential yet challenging because of the limited amount of data. Here, we employed a simple frequency-based weighting. This approach greatly simplifies the process compared with existing techniques. Experimental results revealed that it can achieve levels of performance comparable to, or even exceeding, those of models with numerous hyperparameters. Additionally, we found that the frequency-based weighting significantly improved the predictive performance for low-frequency labels, while requiring almost no changes in implementation. The source code for the proposed method is available on github at this https URL.

Comments:	10pages, Accepted at ECIR25
Subjects:	Information Retrieval (cs.IR); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2501.10179 [cs.IR]
	(or arXiv:2501.10179v1 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.2501.10179

Submission history

From: Katauhiko Hayashi [view email]
[v1] Fri, 17 Jan 2025 13:24:13 UTC (478 KB)

Computer Science > Information Retrieval

Title:A Simple but Effective Closed-form Solution for Extreme Multi-label Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Information Retrieval

Title:A Simple but Effective Closed-form Solution for Extreme Multi-label Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators