Linear regression through PAC-Bayesian truncation

Jean-Yves Audibert; Olivier Catoni

Pré-Publication, Document De Travail Année : 2011

Linear regression through PAC-Bayesian truncation

(1, 2) , (3, 4)

1
2
3
4

Jean-Yves Audibert

Fonction : Auteur
PersonId : 931557

imagine [Marne-la-Vallée]

Statistical Machine Learning and Parsimony

Olivier Catoni

Fonction : Auteur
PersonId : 858015

Département de Mathématiques et Applications - ENS Paris

Computational Learning, Aggregation, Supervised Statistical, Inference, and Classification

Résumé

We consider the problem of predicting as well as the best linear combination of d given functions in least squares regression under L^\infty constraints on the linear combination. When the input distribution is known, there already exists an algorithm having an expected excess risk of order d/n, where n is the size of the training data. Without this strong assumption, standard results often contain a multiplicative log(n) factor, complex constants involving the conditioning of the Gram matrix of the covariates, kurtosis coefficients or some geometric quantity characterizing the relation between L^2 and L^\infty-balls and require some additional assumptions like exponential moments of the output. This work provides a PAC-Bayesian shrinkage procedure with a simple excess risk bound of order d/n holding in expectation and in deviations, under various assumptions. The common surprising factor of these results is their simplicity and the absence of exponential moment condition on the output distribution while achieving exponential deviations. The risk bounds are obtained through a PAC-Bayesian analysis on truncated differences of losses. We also show that these results can be generalized to other strongly convex loss functions.

Mots clés

Linear regression Generalization error Shrinkage PAC-Bayesian theorems Risk bounds Robust statistics Resistant estimators Gibbs posterior distributions Randomized estimators Statistical learning theory

Domaines

Statistiques [math.ST] Théorie [stat.TH]

Fichier principal

dovern2long.pdf (291.04 Ko)

Origine : Fichiers produits par l'(les) auteur(s)

Jean-Yves Audibert : Connectez-vous pour contacter le contributeur

https://hal.science/hal-00522536

Soumis le : dimanche 11 septembre 2011-17:26:49

Dernière modification le : vendredi 19 avril 2024-16:18:54

Archivage à long terme le : lundi 12 décembre 2011-02:21:46

Dates et versions

hal-00522536 , version 1 (30-09-2010)

hal-00522536 , version 2 (11-09-2011)

Identifiants

HAL Id : hal-00522536 , version 2
ARXIV : 1010.0072

Citer

Jean-Yves Audibert, Olivier Catoni. Linear regression through PAC-Bayesian truncation. 2011. ⟨hal-00522536v2⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

ENS-PARIS ENPC CNRS INRIA PARISTECH LIGM IMAGINE INRIA2 PSL MATH_ENS_PARIS UNIV-EIFFEL JSE2024

594 Consultations

309 Téléchargements

Linear regression through PAC-Bayesian truncation

Résumé

Mots clés

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager