Decoupled Greedy Learning of CNNs

Eugene Belilovsky; Michael Eickenberg; Edouard Oyallon

Communication Dans Un Congrès Proceedings of the 37th International Conference on Machine Learning Année : 2020

Decoupled Greedy Learning of CNNs

(1) , (2) , (3)

1
2
3

Eugene Belilovsky

Fonction : Auteur

Montreal Institute for Learning Algorithms [Montréal]

Michael Eickenberg

Fonction : Auteur

Flatiron Institute

Edouard Oyallon

Fonction : Auteur
PersonId : 179157
IdHAL : edouard-oyallon
ORCID : 0000-0002-4826-7527
IdRef : 228745500

Machine Learning and Information Access

Résumé

A commonly cited inefficiency of neural network training by back-propagation is the update locking problem: each layer must wait for the signal to propagate through the full network before updating. Several alternatives that can alleviate this issue have been proposed. In this context, we consider a simpler, but more effective, substitute that uses minimal feedback, which we call Decoupled Greedy Learning (DGL). It is based on a greedy relaxation of the joint training objective, recently shown to be effective in the context of Convolutional Neural Networks (CNNs) on large-scale image classification. We consider an optimization of this objective that permits us to decouple the layer training, allowing for layers or modules in networks to be trained with a potentially linear parallelization in layers. With the use of a replay buffer we show this approach can be extended to asynchronous settings, where modules can operate with possibly large communication delays. We show theoretically and empirically that this approach converges. Then, we empirically find that it can lead to better generalization than sequential greedy optimization. We demonstrate the effectiveness of DGL against alternative approaches on the CIFAR-10 dataset and on the large-scale ImageNet dataset.

Domaines

Intelligence artificielle [cs.AI]

Edouard Oyallon : Connectez-vous pour contacter le contributeur

https://hal.science/hal-02945327

Soumis le : mardi 22 septembre 2020-11:07:54

Dernière modification le : samedi 11 novembre 2023-20:50:03

Dates et versions

hal-02945327 , version 1 (22-09-2020)

Identifiants

HAL Id : hal-02945327 , version 1
ARXIV : 1901.08164

Citer

Eugene Belilovsky, Michael Eickenberg, Edouard Oyallon. Decoupled Greedy Learning of CNNs. International Conference on Machine Learning, Jul 2020, Vienna (virtual), Austria. pp.5368-5377. ⟨hal-02945327⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

CNRS LIP6 SORBONNE-UNIVERSITE SU-SCIENCES

37 Consultations

0 Téléchargements

Decoupled Greedy Learning of CNNs

Résumé

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager