Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study

Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study. An, A., Qian, P., Wilcox, E., & Levy, R. In Empirical Methods in Natural Language Processing (EMNLP), 2019.

Paper abstract bibtex

Neural language models have achieved state-of-the-art performances on many NLP tasks, and recently have been shown to learn a number of hierarchically-sensitive syntactic dependencies between individual words. However, equally important for language processing is the ability to combine words into phrasal constituents, and use constituent-level features to drive downstream expectations. Here we investigate neural models' ability to represent constituent-level features, using coordinated noun phrases as a case study. We assess whether different neural language models trained on English and French represent phrase-level number and gender features, and use those features to drive downstream expectations. Our results suggest that models use a linear combination of NP constituent number to drive CoordNP/verb number agreement. This behavior is highly regular and even sensitive to local syntactic context, however it differs crucially from observed human behavior. Models have less success with gender agreement. Models trained on large corpora perform best, and there is no obvious advantage for models trained using explicit syntactic supervision.

@inproceedings{An2019a,
abstract = {Neural language models have achieved state-of-the-art performances on many NLP tasks, and recently have been shown to learn a number of hierarchically-sensitive syntactic dependencies between individual words. However, equally important for language processing is the ability to combine words into phrasal constituents, and use constituent-level features to drive downstream expectations. Here we investigate neural models' ability to represent constituent-level features, using coordinated noun phrases as a case study. We assess whether different neural language models trained on English and French represent phrase-level number and gender features, and use those features to drive downstream expectations. Our results suggest that models use a linear combination of NP constituent number to drive CoordNP/verb number agreement. This behavior is highly regular and even sensitive to local syntactic context, however it differs crucially from observed human behavior. Models have less success with gender agreement. Models trained on large corpora perform best, and there is no obvious advantage for models trained using explicit syntactic supervision.},
archivePrefix = {arXiv},
arxivId = {1909.04625},
author = {An, Aixiu and Qian, Peng and Wilcox, Ethan and Levy, Roger},
booktitle = {Empirical Methods in Natural Language Processing (EMNLP)},
eprint = {1909.04625},
file = {:Users/shanest/Documents/Library/An et al/Empirical Methods in Natural Language Processing (EMNLP)/An et al. - 2019 - Representation of Constituents in Neural Language Models Coordination Phrase as a Case Study.pdf:pdf},
keywords = {method: psycholinguistic,phenomenon: coordinated NPs},
title = {{Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study}},
url = {http://arxiv.org/abs/1909.04625},
year = {2019}
}

Downloads: 0

{"_id":"ZvQzxgenvMQGFRNck","bibbaseid":"an-qian-wilcox-levy-representationofconstituentsinneurallanguagemodelscoordinationphraseasacasestudy-2019","authorIDs":[],"author_short":["An, A.","Qian, P.","Wilcox, E.","Levy, R."],"bibdata":{"bibtype":"inproceedings","type":"inproceedings","abstract":"Neural language models have achieved state-of-the-art performances on many NLP tasks, and recently have been shown to learn a number of hierarchically-sensitive syntactic dependencies between individual words. However, equally important for language processing is the ability to combine words into phrasal constituents, and use constituent-level features to drive downstream expectations. Here we investigate neural models' ability to represent constituent-level features, using coordinated noun phrases as a case study. We assess whether different neural language models trained on English and French represent phrase-level number and gender features, and use those features to drive downstream expectations. Our results suggest that models use a linear combination of NP constituent number to drive CoordNP/verb number agreement. This behavior is highly regular and even sensitive to local syntactic context, however it differs crucially from observed human behavior. Models have less success with gender agreement. Models trained on large corpora perform best, and there is no obvious advantage for models trained using explicit syntactic supervision.","archiveprefix":"arXiv","arxivid":"1909.04625","author":[{"propositions":[],"lastnames":["An"],"firstnames":["Aixiu"],"suffixes":[]},{"propositions":[],"lastnames":["Qian"],"firstnames":["Peng"],"suffixes":[]},{"propositions":[],"lastnames":["Wilcox"],"firstnames":["Ethan"],"suffixes":[]},{"propositions":[],"lastnames":["Levy"],"firstnames":["Roger"],"suffixes":[]}],"booktitle":"Empirical Methods in Natural Language Processing (EMNLP)","eprint":"1909.04625","file":":Users/shanest/Documents/Library/An et al/Empirical Methods in Natural Language Processing (EMNLP)/An et al. - 2019 - Representation of Constituents in Neural Language Models Coordination Phrase as a Case Study.pdf:pdf","keywords":"method: psycholinguistic,phenomenon: coordinated NPs","title":"Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study","url":"http://arxiv.org/abs/1909.04625","year":"2019","bibtex":"@inproceedings{An2019a,\nabstract = {Neural language models have achieved state-of-the-art performances on many NLP tasks, and recently have been shown to learn a number of hierarchically-sensitive syntactic dependencies between individual words. However, equally important for language processing is the ability to combine words into phrasal constituents, and use constituent-level features to drive downstream expectations. Here we investigate neural models' ability to represent constituent-level features, using coordinated noun phrases as a case study. We assess whether different neural language models trained on English and French represent phrase-level number and gender features, and use those features to drive downstream expectations. Our results suggest that models use a linear combination of NP constituent number to drive CoordNP/verb number agreement. This behavior is highly regular and even sensitive to local syntactic context, however it differs crucially from observed human behavior. Models have less success with gender agreement. Models trained on large corpora perform best, and there is no obvious advantage for models trained using explicit syntactic supervision.},\narchivePrefix = {arXiv},\narxivId = {1909.04625},\nauthor = {An, Aixiu and Qian, Peng and Wilcox, Ethan and Levy, Roger},\nbooktitle = {Empirical Methods in Natural Language Processing (EMNLP)},\neprint = {1909.04625},\nfile = {:Users/shanest/Documents/Library/An et al/Empirical Methods in Natural Language Processing (EMNLP)/An et al. - 2019 - Representation of Constituents in Neural Language Models Coordination Phrase as a Case Study.pdf:pdf},\nkeywords = {method: psycholinguistic,phenomenon: coordinated NPs},\ntitle = {{Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study}},\nurl = {http://arxiv.org/abs/1909.04625},\nyear = {2019}\n}\n","author_short":["An, A.","Qian, P.","Wilcox, E.","Levy, R."],"key":"An2019a","id":"An2019a","bibbaseid":"an-qian-wilcox-levy-representationofconstituentsinneurallanguagemodelscoordinationphraseasacasestudy-2019","role":"author","urls":{"Paper":"http://arxiv.org/abs/1909.04625"},"keyword":["method: psycholinguistic","phenomenon: coordinated NPs"],"metadata":{"authorlinks":{}},"downloads":0},"bibtype":"inproceedings","biburl":"https://www.shane.st/teaching/575/win20/MachineLearning-interpretability.bib","creationDate":"2020-01-05T04:04:02.888Z","downloads":0,"keywords":["method: psycholinguistic","phenomenon: coordinated nps"],"search_terms":["representation","constituents","neural","language","models","coordination","phrase","case","study","an","qian","wilcox","levy"],"title":"Representation of Constituents in Neural Language Models: Coordination Phrase as a Case Study","year":2019,"dataSources":["okYcdTpf4JJ2zkj7A","znj7izS5PeehdLR3G","aGtG992oMsrqA3Aas"]}