Confidence sets with expected sizes for Multiclass Classification

Christophe Denis; Mohamed Hebiri

Pré-Publication, Document De Travail Année : 2016

Confidence sets with expected sizes for Multiclass Classification

(1) , (1)

Christophe Denis

Fonction : Auteur
PersonId : 1036220
IdHAL : christophedenisuge

Laboratoire d'Analyse et de Mathématiques Appliquées

Mohamed Hebiri

Fonction : Auteur

Laboratoire d'Analyse et de Mathématiques Appliquées

Résumé

Challenging multiclass classification problems such as image annotation may involve a large number of classes. In this context, confusion between classes may occur, and single label classification may be misleading. We provide in the present paper a general device that, given a classification procedure and an unlabeled dataset, outputs a set of class labels, instead of a single one. Interestingly, this procedure does not require that the unlabeled dataset explores the whole classes. Even more, the method is calibrated to control the expected size of the output set while minimizing the classification risk. We show the statistical optimality of the procedure and establish rates of convergence under the Tsybakov margin condition. It turns out that these rates are linear on the number of labels. We illustrate the numerical performance of the procedure on simulated and on real data. In particular, we show that with moderate expected size, w.r.t. the number of labels, the procedure provides significant improvement of the classification risk.

Mots clés

Multiclass classification confidence sets plug-in confidence sets cumulative distribution functions

Domaines

Statistiques [math.ST]

Fichier principal

MultiClass_DH.pdf (219.5 Ko)

Origine : Fichiers produits par l'(les) auteur(s)

Mohamed Hebiri : Connectez-vous pour contacter le contributeur

https://hal.science/hal-01357850

Soumis le : mardi 30 août 2016-15:14:43

Dernière modification le : lundi 11 mars 2024-12:04:04

Dates et versions

hal-01357850 , version 1 (30-08-2016)

hal-01357850 , version 2 (28-11-2017)

Identifiants

HAL Id : hal-01357850 , version 1
ARXIV : 1608.08783

Citer

Christophe Denis, Mohamed Hebiri. Confidence sets with expected sizes for Multiclass Classification. 2016. ⟨hal-01357850v1⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

140 Consultations

81 Téléchargements

Confidence sets with expected sizes for Multiclass Classification

Résumé

Mots clés

Domaines

Dates et versions

Identifiants

Citer

Exporter

Altmetric

Partager