Skip to Main content Skip to Navigation
Preprints, Working Papers, ...

Robust Multi-Output Learning with Highly Incomplete Data via Restricted Boltzmann Machines

Giancarlo Fissore 1, 2 Aurelien Decelle 1, 2 Cyril Furtlehner 1, 2 Yufei Han
2 TAU - TAckling the Underspecified
Inria Saclay - Ile de France, LRI - Laboratoire de Recherche en Informatique
Abstract : In a standard multi-output classification scenario, both features and labels of training data are partially observed. This challenging issue is widely witnessed due to sensor or database failures, crowd-sourcing and noisy communication channels in industrial data analytic services. Classic methods for handling multi-output classification with incomplete supervision information usually decompose the problem into an imputation stage that reconstructs the missing training information, and a learning stage that builds a classifier based on the imputed training set. These methods fail to fully leverage the dependencies between features and labels. In order to take full advantage of these dependencies we consider a purely probabilistic setting in which the features imputation and multi-label classification problems are jointly solved. Indeed, we show that a simple Restricted Boltzmann Machine can be trained with an adapted algorithm based on mean-field equations to efficiently solve problems of inductive and transductive learning in which both features and labels are missing at random. The effectiveness of the approach is demonstrated empirically on various datasets, with particular focus on a real-world Internet-of-Things security dataset.
Complete list of metadatas

https://hal.archives-ouvertes.fr/hal-02420824
Contributor : Aurélien Decelle <>
Submitted on : Friday, December 20, 2019 - 10:23:16 AM
Last modification on : Wednesday, October 14, 2020 - 3:59:10 AM

Links full text

Identifiers

  • HAL Id : hal-02420824, version 1
  • ARXIV : 1912.09382

Collections

Citation

Giancarlo Fissore, Aurelien Decelle, Cyril Furtlehner, Yufei Han. Robust Multi-Output Learning with Highly Incomplete Data via Restricted Boltzmann Machines. 2019. ⟨hal-02420824⟩

Share

Metrics

Record views

25