Fetching the paper…
Reading the bibliography…
By introducing a small set of additional parameters, a probe learns to solve specific linguistic tasks (e.g., dependency parsing) in a supervised manner using feature representations (e.g., contextualized embeddings).
An imitation learning approach to unsupervised parsing
Bowen Li, Lili Mou, and Frank Keller. 2019a · 1906
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime G. Carbonell, Ruslan Salakhutdinov, and Quoc V. Le. 2019 · 1906
Earlier work this paper cites.
A critical analysis of biased parsers in unsupervised parsing
Chris Dyer, Gábor Melis, and Phil Blunsom. 2019 · 1909
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
On the shortest arborescence of a directed graph
Yoeng-Jin Chu. 1965 · 1965
Earlier work this paper cites.
Optimum branchings
Jack Edmonds. 1967 · 1967
Earlier work this paper cites.
A formal model of the structure of discourse
Livia Polanyi. 1988 · 1988
Earlier work this paper cites.
Assessing the Ability of LSTMs to Learn Syntax-Sensitive Dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 1990
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Mitchell Marcus, Beatrice Santorini, and Mary Ann Marcinkiewicz. 1993 · 1993
Earlier work this paper cites.
Three new probabilistic models for dependency parsing: An exploration
Jason M Eisner. 1996 · 1996
Earlier work this paper cites.
HEAD-DRIVEN STATISTICAL MODELS FOR NATURAL LANGUAGE PARSING
Michael Collins. 1999 · 1999
Earlier work this paper cites.
Building a discourse-tagged corpus in the framework of rhetorical structure theory
Lynn Carlson, Daniel Marcu, and Mary Ellen Okurowski. 2003 · 2003
Earlier work this paper cites.
Statistical dependency analysis with support vector machines
Hiroyasu Yamada and Yuji Matsumoto. 2003 · 2003
Earlier work this paper cites.
Corpus-based induction of syntactic structure: Models of dependency and constituency
Dan Klein and Christopher D Manning. 2004 · 2004
Earlier work this paper cites.
Dependency parsing
Sandra Kübler, Ryan McDonald, and Joakim Nivre. 2009 · 2009
Earlier work this paper cites.
Favor short dependencies: Parsing with soft and hard constraints on dependency length
Jason Eisner and Noah A Smith. 2010 · 2010
Earlier work this paper cites.
Neutralizing linguistically problematic annotations in unsupervised dependency parsing evaluation
Roy Schwartz, Omri Abend, Roi Reichart, and Ari Rappoport. 2011 · 2011
Earlier work this paper cites.
Evaluating dependency parsing: robust and heuristics-free cross-nnotation evaluation
Reut Tsarfaty, Joakim Nivre, and Evelina Ndersson. 2011 · 2011
Earlier work this paper cites.
Text-level discourse dependency parsing
Sujian Li, Liang Wang, Ziqiang Cao, and Wenjie Li. 2014 · 2014
Cited alongside, same era.
The Stanford CoreNLP natural language processing toolkit
Christopher D. Manning, Mihai Surdeanu, John Bauer, Jenny Finkel, Steven J. Bethard, and David McClosky. 2014 · 2014
Cited alongside, same era.
SemEval-2014 task 4: Aspect based sentiment analysis
Maria Pontiki, Dimitris Galanis, John Pavlopoulos, Harris Papageorgiou, Ion Androutsopoulos, and Suresh Manandhar. 2014 · 2014
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Does String-Based Neural MT Learn Source Syntax?
Xing Shi, Inkit Padhi, and Kevin Knight. 2016 · 2016
Cited alongside, same era.
Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2017 · 2017
Scidtb: Discourse dependency treebank for scientific abstracts
An Yang and Sujian Li. 2018 · 2018
Later among the works it cites.
Does BERT agree? Evaluating knowledge of structure dependence through agreement relations
Geoff Bacon and Terry Regier. 2019 · 2019
Later among the works it cites.
What Does BERT Look at? An Analysis of BERT’s Attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning. 2019 · 2019
Later among the works it cites.
Assessing BERT’s Syntactic Abilities
Yoav Goldberg. 2019 · 2019
Later among the works it cites.
Designing and interpreting probes with control tasks
John Hewitt and Percy Liang. 2019 · 2019
Later among the works it cites.
A Structural Probe for Finding Syntax in Word Representations
John Hewitt and Christopher D. Manning. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Conll 2017 shared task: Multilingual parsing from raw text to universal dependencies
Daniel Zeman, Martin Popel, Milan Straka, Jan Hajic, Joakim Nivre, Filip Ginter, Juhani Luotolahti, Sampo Pyysalo, Slav Petrov, Martin Potthast, et al. 2017 · 2017
Cited alongside, same era.
What you can cram into a single vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, Germán Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Under the hood: Using diagnostic classifiers to investigate and improve how language models track agreement information
Mario Giulianelli, Jack Harding, Florian Mohnert, Dieuwke Hupkes, and Willem Zuidema. 2018 · 2018
Cited alongside, same era.
Colorless Green Recurrent Networks Dream Hierarchically
Kristina Gulordava, Piotr Bojanowski, Edouard Grave, Tal Linzen, and Marco Baroni. 2018 · 2018
Cited alongside, same era.
Extracting Syntactic Trees from Transformer Encoder Self-Attentions
David Mareček and Rudolf Rosa. 2018 · 2018
Cited alongside, same era.
Do attention heads in bert track syntactic dependencies?
Phu Mon Htut, Jason Phang, Shikha Bordia, and Samuel R. Bowman. 2019 · 2019
Later among the works it cites.
Attention is not explanation
Sarthak Jain and Byron C Wallace. 2019 · 2019
Later among the works it cites.
What Does BERT Learn about the Structure of Language?
Ganesh Jawahar, Benoît Sagot, and Djamé Seddah. 2019 · 2019
Later among the works it cites.
Linguistic knowledge and transferability of contextual representations
Nelson F Liu, Matt Gardner, Yonatan Belinkov, Matthew E Peters, and Noah A Smith. 2019 · 2019
Later among the works it cites.
From Balustrades to Pierre Vinken: Looking for Syntax in Transformer Self-Attentions
David Mareček and Rudolf Rosa. 2019 · 2019
Later among the works it cites.
Inducing Syntactic Trees from BERT Representations
Rudolf Rosa and David Mareček. 2019 · 2019
Later among the works it cites.
Is attention interpretable?
Sofia Serrano and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Ordered neurons: Integrating tree structures into recurrent neural networks
Yikang Shen, Shawn Tan, Alessandro Sordoni, and Aaron Courville. 2019 · 2019
Later among the works it cites.
Next sentence prediction helps implicit discourse relation classification within and across domains
Wei Shi and Vera Demberg. 2019 · 2019
Later among the works it cites.
Attention is not not explanation
Sarah Wiegreffe and Yuval Pinter. 2019 · 2019
Later among the works it cites.
Syntax-aware aspect-level sentiment classification with proximity-weighted convolution network
Chen Zhang, Qiuchi Li, and Dawei Song. 2019 · 2019
Later among the works it cites.