Confidence sets with expected sizes for Multiclass Classification

Denis, Christophe; Hebiri, Mohamed

Mathematics > Statistics Theory

arXiv:1608.08783v1 (math)

[Submitted on 31 Aug 2016 (this version), latest version 18 Dec 2017 (v2)]

Title:Confidence sets with expected sizes for Multiclass Classification

Authors:Christophe Denis (LAMA), Mohamed Hebiri (LAMA)

View PDF

Abstract:Challenging multiclass classification problems such as image annotation may involve a large number of classes. In this context, confusion between classes may occur, and single label classification may be misleading. We provide in the present paper a general device that, given a classification procedure and an unlabeled dataset, outputs a set of class labels, instead of a single one. Interestingly, this procedure does not require that the unlabeled dataset explores the whole classes. Even more, the method is calibrated to control the expected size of the output set while minimizing the classification risk. We show the statistical optimality of the procedure and establish rates of convergence under the Tsybakov margin condition. It turns out that these rates are linear on the number of labels. We illustrate the numerical performance of the procedure on simulated and on real data. In particular, we show that with moderate expected size, w.r.t. the number of labels, the procedure provides significant improvement of the classification risk.

Subjects:	Statistics Theory (math.ST)
Cite as:	arXiv:1608.08783 [math.ST]
	(or arXiv:1608.08783v1 [math.ST] for this version)
	https://doi.org/10.48550/arXiv.1608.08783

Submission history

From: Mohamed Hebiri [view email] [via CCSD proxy]
[v1] Wed, 31 Aug 2016 09:22:58 UTC (20 KB)
[v2] Mon, 18 Dec 2017 08:06:34 UTC (25 KB)

Mathematics > Statistics Theory

Title:Confidence sets with expected sizes for Multiclass Classification

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Mathematics > Statistics Theory

Title:Confidence sets with expected sizes for Multiclass Classification

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators