Fetching the paper…
Reading the bibliography…
In this paper, we focus on learning structure-aware document representations from data without recourse to a discourse parser or additional annotations.
Éléments de Syntaxe Structurale
Louis Tesniére. 1959 · 1959
Earlier work this paper cites.
On shortest arborescence of a directed graph
Yoeng-Jin Chu and Tseng-Hong Liu. 1965 · 1965
Earlier work this paper cites.
Optimum branchings
Jack Edmonds. 1967 · 1967
Earlier work this paper cites.
Trainable grammars for speech recognition
James K. Baker. 1979 · 1979
Earlier work this paper cites.
Graph theory
William Thomas Tutte. 1984 · 1984
Earlier work this paper cites.
Rhetorical structure theory: Toward a functional theory of text organization
William C. Mann and Sandra A. Thompson. 1988 · 1988
Earlier work this paper cites.
Dependency Syntax: Theory and Practice
Igor A. Melc̆uk. 1988 · 1988
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Building a discourse-tagged corpus in the framework of rhetorical structure theory
Lynn Carlson, Daniel Marcu, and Mary Ellen Okurowski. 2001 · 2001
Earlier work this paper cites.
Complexity of dependencies in discourse: Are dependencies in discourse more complex than in syntax
Alan Lee, Rashmi Prasad, Aravind Joshi, Nikhil Dinesh, and Bonnie Webber. 2006 · 2006
Earlier work this paper cites.
Get out the vote: Determining support or opposition from congressional floor-debate transcripts
Matt Thomas, Bo Pang, and Lillian Lee. 2006 · 2006
Earlier work this paper cites.
Coherence in Natural Language: Data Structures and Applications
Florian Wolf and Edward Gibson. 2006 · 2006
Earlier work this paper cites.
Structured prediction models via the matrix-tree theorem
Terry Koo, Amir Globerson, Xavier Carreras Pérez, and Michael Collins. 2007 · 2007
Earlier work this paper cites.
On the complexity of non-projective data-driven dependency parsing
Ryan McDonald and Giorgio Satta. 2007 · 2007
Earlier work this paper cites.
Discourse-based answering of why-questions
Suzan Verberne, Lou Boves, Nelleke Oostdijk, and Peter-Arno Coppen. 2007 · 2007
Earlier work this paper cites.
Modeling local coherence: An entity-based approach
Regina Barzilay and Mirella Lapata. 2008 · 2008
Earlier work this paper cites.
The Penn discourse TreeBank 2.0
Rashmi Prasad, Nikhil Dinesh, Alan Lee, Eleni Miltsakaki, Livio Robaldo, Aravind K. Joshi, and Bonnie L. Webber. 2008 · 2008
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer. 2011 · 2011
Cited alongside, same era.
Automatically evaluating text coherence using discourse relations
Ziheng Lin, Hwee Tou Ng, and Min-Yen Kan. 2011 · 2011
Cited alongside, same era.
Text-level discourse parsing with rich linguistic features
Vanessa Wei Feng and Graeme Hirst. 2012 · 2012
Cited alongside, same era.
Unsupervised improving of sentiment analysis using global target context
Tomáš Brychcın and Ivan Habernal. 2013 · 2013
Cited alongside, same era.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-Rahman Mohamed, and Geoffrey Hinton. 2013 · 2013
Cited alongside, same era.
Single-document summarization as a tree knapsack problem
Tsutomu Hirao, Yasuhisa Yoshida, Masaaki Nishino, Norihito Yasuda, and Masaaki Nagata. 2013 · 2013
Graph-based coherence modeling for assessing readability
Mohsen Mesgar and Michael Strube. 2015 · 2015
Later among the works it cites.
A fast unified model for parsing and sentence understanding
Samuel R. Bowman, Jon Gauthier, Abhinav Rastogi, Raghav Gupta, Christopher D. Manning, and Christopher Potts. 2016 · 2016
Later among the works it cites.
Distraction-based neural networks for modeling documents
Qian Chen, Xiaodan Zhu, Zhenhua Ling, Si Wei, and Hui Jiang. 2016 · 2016
Later among the works it cites.
Long short-term memory-networks for machine reading
Jianpeng Cheng, Li Dong, and Mirella Lapata. 2016 · 2016
Later among the works it cites.
Empirical comparison of dependency conversions for rst discourse trees
Katsuhiko Hayashi, Tsutomu Hirao, and Masaaki Nagata. 2016 · 2016
Later among the works it cites.
Assessing the ability of LSTMs to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Implicitation of discourse connectives in (machine) translation
Thomas Meyer and Bonnie Webber. 2013 · 2013
Cited alongside, same era.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S. Corrado, and Jeff Dean. 2013 · 2013
Cited alongside, same era.
Integrating document clustering and topic modeling
Pengtao Xie and Eric P. Xing. 2013 · 2013
Cited alongside, same era.
Jointly modeling aspects, ratings and sentiments for movie recommendation (JMARS)
Qiming Diao, Minghui Qiu, Chao-Yuan Wu, Alexander J Smola, Jing Jiang, and Chong Wang. 2014 · 2014
Cited alongside, same era.
Text-level discourse dependency parsing
Sujian Li, Liang Wang, Ziqiang Cao, and Wenjie Li. 2014 · 2014
Cited alongside, same era.
The Stanford CoreNLP Natural Language Processing Toolkit
Christopher D. Manning, Mihai Surdeanu, John Bauer, Jenny Rose Finkel, Steven Bethard, and David McClosky. 2014 · 2014
Cited alongside, same era.
Later among the works it cites.
Key-value memory networks for directly reading documents
Alexander Miller, Adam Fisch, Jesse Dodge, Amir-Hossein Karimi, Antoine Bordes, and Jason Weston. 2016 · 2016
Later among the works it cites.
A decomposable attention model for natural language inference
Ankur Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Later among the works it cites.
Reasoning about entailment with neural attention
Tim Rocktäschel, Edward Grefenstette, Karl Moritz Hermann, Tomáš Kočiskỳ, and Phil Blunsom. 2016 · 2016
Later among the works it cites.
Learning natural language inference with LSTM
Shuohang Wang and Jing Jiang. 2016 · 2016
Later among the works it cites.
Hierarchical attention networks for document classification
Zichao Yang, Diyi Yang, Chris Dyer, Xiaodong He, Alex Smola, and Eduard Hovy. 2016 · 2016
Later among the works it cites.
Enhanced LSTM for natural language inference
Qian Chen, Xiaodan Zhu, Zhen-Hua Ling, Si Wei, Hui Jiang, and Diana Inkpen. 2017 · 2017
Closest in time.
Frustratingly short attention spans in neural language modeling
Michał Daniluk, Tim Rocktäschel, Johannes Welbl, and Sebastian Riedel. 2017 · 2017
Closest in time.
Neural discourse structure for text categorization
Yangfeng Ji and Noah Smith. 2017 · 2017
Closest in time.
Structured attention networks
Yoon Kim, Carl Denton, Luong Hoang, and Alexander M. Rush. 2017 · 2017
Closest in time.
Learning contextually informed representations for linear-time discourse parsing
Yang Liu and Mirella Lapata. 2017 · 2017
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio. 2015 · 2057
Closest in time.