PAC-Bayesian Learning and Domain Adaptation

Pascal Germain; Amaury Habrard; François Laviolette; Emilie Morvant

Communication Dans Un Congrès Année : 2012

PAC-Bayesian Learning and Domain Adaptation

(1) , (2) , (1) , (3, 4, 2)

1
2
3
4

Pascal Germain

Fonction : Auteur
PersonId : 14639
IdHAL : pascal-germain
ORCID : 0000-0003-3998-9533
IdRef : 240406532

GRAAL

Amaury Habrard

Fonction : Auteur
PersonId : 439
IdHAL : amaury-habrard
ORCID : 0000-0003-3038-9347
IdRef : 084103655

Laboratoire Hubert Curien

François Laviolette

Fonction : Auteur

GRAAL

Emilie Morvant

Fonction : Auteur
PersonId : 410
IdHAL : emilie-morvant
ORCID : 0000-0002-8301-7240
IdRef : 179027468

éQuipe AppRentissage et MultimediA [Marseille]

Laboratoire d'informatique Fondamentale de Marseille

Laboratoire Hubert Curien

Résumé

In machine learning, Domain Adaptation (DA) arises when the distribution gen- erating the test (target) data differs from the one generating the learning (source) data. It is well known that DA is an hard task even under strong assumptions, among which the covariate-shift where the source and target distributions diverge only in their marginals, i.e. they have the same labeling function. Another popular approach is to consider an hypothesis class that moves closer the two distributions while implying a low-error for both tasks. This is a VC-dim approach that restricts the complexity of an hypothesis class in order to get good generalization. Instead, we propose a PAC-Bayesian approach that seeks for suitable weights to be given to each hypothesis in order to build a majority vote. We prove a new DA bound in the PAC-Bayesian context. This leads us to design the first DA-PAC-Bayesian algorithm based on the minimization of the proposed bound. Doing so, we seek for a ρ-weighted majority vote that takes into account a trade-off between three quantities. The first two quantities being, as usual in the PAC-Bayesian approach, (a) the complexity of the majority vote (measured by a Kullback-Leibler divergence) and (b) its empirical risk (measured by the ρ-average errors on the source sample). The third quantity is (c) the capacity of the majority vote to distinguish some structural difference between the source and target samples.

Mots clés

Domain Adaptation PAC-Bayes

Domaines

Machine Learning [stat.ML]

Fichier principal

dapacbayes.pdf (946.19 Ko)

poster_DAPACBAYES.pdf (2.23 Mo)

Origine : Fichiers produits par l'(les) auteur(s)

Format : Autre

Emilie Morvant : Connectez-vous pour contacter le contributeur

https://hal.science/hal-00749366

Soumis le : lundi 10 décembre 2012-17:25:47

Dernière modification le : vendredi 24 mars 2023-14:52:56

Archivage à long terme le : samedi 17 décembre 2016-08:27:59

Dates et versions

hal-00749366 , version 1 (10-12-2012)

Identifiants

HAL Id : hal-00749366 , version 1
ARXIV : 1212.2340

Citer

Pascal Germain, Amaury Habrard, François Laviolette, Emilie Morvant. PAC-Bayesian Learning and Domain Adaptation. Multi-Trade-offs in Machine Learning, NIPS 2012 Workshop, Dec 2012, Lake Tahoe, United States. ⟨hal-00749366⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

UNIV-ST-ETIENNE IOGS LIF CNRS UNIV-AMU LAHC EC-MARSEILLE PARISTECH LIS-LAB UDL

328 Consultations

207 Téléchargements

PAC-Bayesian Learning and Domain Adaptation

Résumé

Mots clés

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager