English

What do RNN Language Models Learn about Filler-Gap Dependencies?

Computation and Language 2018-09-05 v1

Abstract

RNN language models have achieved state-of-the-art perplexity results and have proven useful in a suite of NLP tasks, but it is as yet unclear what syntactic generalizations they learn. Here we investigate whether state-of-the-art RNN language models represent long-distance filler-gap dependencies and constraints on them. Examining RNN behavior on experimentally controlled sentences designed to expose filler-gap dependencies, we show that RNNs can represent the relationship in multiple syntactic positions and over large spans of text. Furthermore, we show that RNNs learn a subset of the known restrictions on filler-gap dependencies, known as island constraints: RNNs show evidence for wh-islands, adjunct islands, and complex NP islands. These studies demonstrates that state-of-the-art RNN models are able to learn and generalize about empty syntactic positions.

Cite

@article{arxiv.1809.00042,
  title  = {What do RNN Language Models Learn about Filler-Gap Dependencies?},
  author = {Ethan Wilcox and Roger Levy and Takashi Morita and Richard Futrell},
  journal= {arXiv preprint arXiv:1809.00042},
  year   = {2018}
}

Comments

9 pages, to appear in Proceedings of BlackboxNLP 2018

R2 v1 2026-06-23T03:51:08.059Z