arXiv · 1809.02836
Context-Free Transductions with Neural Stacks
Abstract
This paper analyzes the behavior of stack-augmented recurrent neural network (RNN) models. Due to the architectural similarity between stack RNNs and pushdown transducers, we train stack RNN models on a number of tasks, including string reversal, context-free language modelling, and cumulative XOR evaluation. Examining the behavior of our networks, we show that stack-augmented RNNs can discover intuitive stack-based strategies for solving our tasks. However, stack RNNs are more difficult to train than classical architectures such as LSTMs. Rather than employ stack-based strategies, more complex networks often find approximate solutions by using the stack as unstructured memory.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yiding Hao, William Merrill, Dana Angluin, Robert Frank, Noah Amsel, Andrew Benz, Simon Mendelsohn. 2018-09-08. Context-Free Transductions with Neural Stacks. https://arxiv.org/abs/1809.02836
Cite the original work for its findings. Save a collection to share your selection of sources.