Unbiasing Truncated Backpropagation Through Time

Yann Ollivier; Corentin Tallec

Pré-Publication, Document De Travail Année : 2017

Unbiasing Truncated Backpropagation Through Time

(1, 2, 3, 4) , (1, 2)

1
2
3
4

Yann Ollivier

Fonction : Auteur
PersonId : 883809

Laboratoire de Recherche en Informatique

TAckling the Underspecified

Centre National de la Recherche Scientifique

Facebook AI Research [Paris]

Corentin Tallec

Fonction : Auteur

Laboratoire de Recherche en Informatique

TAckling the Underspecified

Résumé

Truncated Backpropagation Through Time (truncated BPTT) is a widespread method for learning recurrent computational graphs. Truncated BPTT keeps the computational benefits of Backpropagation Through Time (BPTT) while relieving the need for a complete backtrack through the whole data sequence at every step. However, truncation favors short-term dependencies: the gradient estimate of truncated BPTT is biased, so that it does not benefit from the convergence guarantees from stochastic gradient theory. We introduce Anticipated Reweighted Truncated Backpropagation (ARTBP), an algorithm that keeps the computational benefits of truncated BPTT, while providing unbiasedness. ARTBP works by using variable truncation lengths together with carefully chosen compensation factors in the backpropagation equation. We check the viability of ARTBP on two tasks. First, a simple synthetic task where careful balancing of temporal dependencies at different scales is needed: truncated BPTT displays unreliable performance, and in worst case scenarios, divergence, while ARTBP converges reliably. Second, on Penn Treebank character-level language modelling, ARTBP slightly outperforms truncated BPTT.

Domaines

Apprentissage [cs.LG]

Yann Ollivier : Connectez-vous pour contacter le contributeur

https://inria.hal.science/hal-01660627

Soumis le : lundi 11 décembre 2017-10:53:18

Dernière modification le : lundi 12 février 2024-09:44:03

Dates et versions

hal-01660627 , version 1 (11-12-2017)

Identifiants

HAL Id : hal-01660627 , version 1
ARXIV : 1705.08209

Citer

Yann Ollivier, Corentin Tallec. Unbiasing Truncated Backpropagation Through Time. 2017. ⟨hal-01660627⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

CNRS INRIA UMR8623 CENTRALESUPELEC INRIA2 LRI-AO UNIV-PARIS-SACLAY LISN GS-ENGINEERING GS-COMPUTER-SCIENCE LISN-AO

306 Consultations

0 Téléchargements

Unbiasing Truncated Backpropagation Through Time

Résumé

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Altmetric

Partager