arXiv · 1711.06788
MinimalRNN: Toward More Interpretable and Trainable Recurrent Neural Networks
Abstract
We introduce MinimalRNN, a new recurrent neural network architecture that achieves comparable performance as the popular gated RNNs with a simplified structure. It employs minimal updates within RNN, which not only leads to efficient learning and testing but more importantly better interpretability and trainability. We demonstrate that by endorsing the more restrictive update rule, MinimalRNN learns disentangled RNN states. We further examine the learning dynamics of different RNN structures using input-output Jacobians, and show that MinimalRNN is able to capture longer range dependencies than existing RNN architectures.
Explore related subjects
Keep this discovery
Minmin Chen. 2017-11-18. MinimalRNN: Toward More Interpretable and Trainable Recurrent Neural Networks. https://arxiv.org/abs/1711.06788
Cite the original work for its findings. Save a collection to share your selection of sources.