arXiv · 1702.04770
Training Language Models Using Target-Propagation
Abstract
While Truncated Back-Propagation through Time (BPTT) is the most popular approach to training Recurrent Neural Networks (RNNs), it suffers from being inherently sequential (making parallelization difficult) and from truncating gradient flow between distant time-steps. We investigate whether Target Propagation (TPROP) style approaches can address these shortcomings. Unfortunately, extensive experiments suggest that TPROP generally underperforms BPTT, and we end with an analysis of this phenomenon, and suggestions for future work.
Explore related subjects
Keep this discovery
Sam Wiseman, Sumit Chopra, Marc'Aurelio Ranzato, Arthur Szlam, Ruoyu Sun, Soumith Chintala, Nicolas Vasilache. 2017-02-15. Training Language Models Using Target-Propagation. https://arxiv.org/abs/1702.04770
Cite the original work for its findings. Save a collection to share your selection of sources.