Fetching the paper…
Reading the bibliography…
Sequence-to-sequence learning with neural networks has become the de facto standard for sequence prediction tasks.
Syntax Directed Translations and the Pushdown Assembler
Alfred V. Aho and Jeffrey D. Ullman · 1969
Earlier work this paper cites.
Trainable Grammars for Speech Recognition
James K. Baker · 1979
Earlier work this paper cites.
Synchronous Tree-Adjoining Grammars
Stuart M. Shieber and Yves Schabes · 1990
Earlier work this paper cites.
An Efficient Probabilistic Context-Free Parsing Algorithm that Computes Prefix Probabilities
Andreas Stolcke · 1995
Earlier work this paper cites.
Computational Complexity of Probabilistic Disambiguation by means of Tree-Grammars
Khalil Sima’an · 1996
Earlier work this paper cites.
Stochastic Inversion Transduction Grammars and Bilingual Parsing of Parallel Corpora
Dekai Wu · 1997
Earlier work this paper cites.
Computational Complexity of Problems on Probabilistic Grammars and Transducers
F. Casacuberta and Colin De La Higuera · 2000
Earlier work this paper cites.
A Syntax-based Statistical Translation Model
Kenji Yamada and Kevin Knight · 2001
Earlier work this paper cites.
The consensus string problem and the complexity of comparing hidden Markov models
Rune B. Lyngsø and Christian N. S. Pedersen · 2002
Earlier work this paper cites.
Learning Non-Isomorphic Tree Mappings for Machine Translation
Jason Eisner · 2003
Earlier work this paper cites.
Loosely Tree-Based Alignment for Machine Translation
Daniel Gildea · 2003
Earlier work this paper cites.
Multitext Grammars and Synchronous Parsers
I. Dan Melamed · 2003
Earlier work this paper cites.
Generalized Multitext Grammars
I. Dan Melamed, Giorgio Satta, and Benjamin Wellington · 2004
Earlier work this paper cites.
A Hierarchical Phrase-Based Model for Statistical Machine Translation
David Chiang · 2005
Earlier work this paper cites.
Machine Translation Using Probabilistic Synchronous Dependency Insertion Grammars
Yuan Ding and Martha Palmer · 2005
Earlier work this paper cites.
Connectionist Temporal Classification: Labelling Unsegmented Sequence Data with Recurrent Neural Networks
Alex Graves · 2006
Earlier work this paper cites.
A Syntax-Directed Translator with Extended Domain of Locality
Liang Huang, Kevin Knight, and Aravind Joshi · 2006
Earlier work this paper cites.
Induction of Probabilistic Synchronous Tree-insertion Grammars for Machine Translation
Rebecca Nesson, Stuart Shieber, and Alexander Rush · 2006
Earlier work this paper cites.
Quasi-Synchronous Grammars: Alignment by Soft Projection of Syntactic Dependencies
David Smith and Jason Eisner · 2006
Earlier work this paper cites.
What is the Jeopardy Model? A Quasi-Synchronous Grammar for QA
Mengqiu Wang, Noah A. Smith, and Teruko Mitamura · 2007
Earlier work this paper cites.
Learning Synchronous Grammars for Semantic Parsing with Lambda Calculus
Yuk Wah Wong and Raymond Mooney · 2007
Earlier work this paper cites.
A Discriminative Latent Variable Model for Statistical Machine Translation
Phil Blunsom, Trevor Cohn, and Miles Osborne · 2008
Earlier work this paper cites.
Latent-Variable Modeling of String Transductions with Finite-State Methods
Markus Dreyer, Jason Smith, and Jason Eisner · 2008
Earlier work this paper cites.
Training Tree Transducers
Jonathan Graehl, Kevin Knight, and Jonathan May · 2008
Earlier work this paper cites.
Bayesian Synchronous Grammar Induction
Phil Blunsom, Trevor Cohn, and Miles Osborne · 2009
Earlier work this paper cites.
Sentence Compression as Tree Transduction
Trevor Cohn and Mirella Lapata · 2009
Earlier work this paper cites.
Paraphrase Identification as Probabilistic Quasi-Synchronous Recognition
Dipanjan Das and Noah A. Smith · 2009
Earlier work this paper cites.
Feature-Rich Translation by Quasi-Synchronous Lattice Parsing
Kevin Gimpel and Noah A. Smith · 2009
Earlier work this paper cites.
Parser Adaptation and Projection with Quasi-Synchronous Grammar Features
David A. Smith and Jason Eisner · 2009
Earlier work this paper cites.
Painless Unsupervised Learning with Features
Taylor Berg-Kirkpatrick, Alexandre Bouchard-Cote, John DeNero, and Dan Klein · 2010
Earlier work this paper cites.
Joint Parsing and Alignment with Weakly Synchronized Grammars
David Burkett, John Blitzer, and Dan Klein · 2010
Earlier work this paper cites.
Title Generation with Quasi-Synchronous Grammar
Kristian Woodsend, Yansong Feng, and Mirella Lapata · 2010
Earlier work this paper cites.
Quasi-Synchronous Phrase Dependency Grammars for Machine Translation
Kevin Gimpel and Noah A. Smith · 2011
Earlier work this paper cites.
Learning to Simplify Sentences with Quasi-Synchronous Grammar and Integer Programming
Kristian Woodsend and Mirella Lapata · 2011
Earlier work this paper cites.
Recurrent Continuous Translation Models
Nal Kalchbrenner and Phil Blunsom · 2013
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks
Ilya Sutskever, Oriol Vinyals, and Quoc Le · 2014
Earlier work this paper cites.
Multiple Object Recognition with Visual Attention
Jimmy Ba, Volodymyr Mnih, and Koray Kavukcuoglu · 2015
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Improved Semantic Representations From Tree-Structured Long Short-Term Memory Networks
Kai Sheng Tai, Richard Socher, and Christopher D. Manning · 2015
Earlier work this paper cites.
Show, Attend and Tell: Neural Image Caption Generation with Visual Attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Earlier work this paper cites.
Long Short-Term Memory Over Tree Structures
Xiaodan Zhu, Parinaz Sobhani, and Hongyu Guo · 2015
Earlier work this paper cites.
Incorporating Structural Alignment Biases into an Attentional Neural Translation Model
Trevor Cohn, Cong Duy Vu Hoang, Ekaterina Vymolova, Kaisheng Yao, Chris Dyer, and Gholamreza Haffari · 2016
Earlier work this paper cites.
Recurrent Neural Network Grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A. Smith · 2016
Earlier work this paper cites.
Incorporating Copying Mechanism in Sequence-to-Sequence Learning
Jiatao Gu, Zhengdong Lu, Hang Li, and Victor OK Li · 2016
Cited alongside, same era.
Pointing the Unknown Words
Caglar Gulcehre, Sungjin Ahn, Ramesh Nallapati, Bowen Zhou, and Yoshua Bengio · 2016
Cited alongside, same era.
Data Recombination for Neural Semantic Parsing
Robin Jia and Percy Liang · 2016
Cited alongside, same era.
Unsupervised Neural Dependency Parsing
Yong Jiang, Wenjuan Han, and Kewei Tu · 2016
Cited alongside, same era.
Segmental Recurrent Neural Networks
Lingpeng Kong, Chris Dyer, and Noah A. Smith · 2016
Cited alongside, same era.
Latent predictor networks for code generation
Wang Ling, Phil Blunsom, Edward Grefenstette, Karl Moritz Hermann, Tomáš Kočiský, Fumin Wang, and Andrew Senior · 2016
Cited alongside, same era.
Compositional Generalization in a Deep Seq2seq Model by Separating Syntax and Semantics
Jake Russin, Jason Jo, Randall C. O’Reilly, and Yoshua Bengio · 2019
Later among the works it cites.
Visually Grounded Neural Syntax Acquisition
Haoyue Shi, Jiayuan Mao, Kevin Gimpel, and Karen Livescu · 2019
Later among the works it cites.
Good-Enough Compositional Data Augmentation
Jacob Andreas · 2020
Later among the works it cites.
Empirical Study of the Benefits of Overparameterization in Learning Latent Variable Models
Rares-Darius Buhai, Yoni Halpern, Yoon Kim, Andrej Risteski, and David Sontag · 2020
Later among the works it cites.
Compositional Generalization via Neural-Symbolic Stack Machines
Xinyun Chen, Chen Liang, Adams Wei Yu, Dawn Song, and Denny Zhou · 2020
Later among the works it cites.
Scaling Hidden Markov Language Models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural Machine Translation with Supervised Attention
Lemao Liu, Masao Utiyama, Andrew Finch, and Eiichiro Sumita · 2016
Cited alongside, same era.
Weighting Finite-State Transductions With Neural Context
Pushpendre Rastogi, Ryan Cotterell, and Jason Eisner · 2016
Cited alongside, same era.
Unsupervised Neural Hidden Markov Models
Ke Tran, Yonatan Bisk, Ashish Vaswani, Daniel Marcu, and Kevin Knight · 2016
Cited alongside, same era.
Online Segment to Segment Neural Transduction
Lei Yu, Jan Buys, and Phil Blunsom · 2016
Cited alongside, same era.
Towards String-to-Tree Neural Machine Translation
Roee Aharoni and Yoav Goldberg · 2017
Cited alongside, same era.
Tree-structured Decoding with Doubly-Recurrent Neural Networks
David Alvarez-Melis and Tommi S. Jaakkola · 2017
Cited alongside, same era.
Justin T. Chiu and Alexander M. Rush · 2020
Later among the works it cites.
Topic Modeling in Embedding Spaces
Adji B. Dieng, Francisco J. R. Ruiz, and David M. Blei · 2020
Later among the works it cites.
Compositional Generalization in Semantic Parsing: Pre-training vs. Specialized Architectures
Daniel Furrer, Marc van Zee, Nathan Scales, and Nathanael Schärli · 2020
Later among the works it cites.
Permutation Equivariant Models for Compositional Generalization in Language
Jonathan Gordon, David Lopez-Paz, Marco Baroni, and Diane Bouchacourt · 2020
Later among the works it cites.
Sequence-Level Mixed Sample Data Augmentation
Demi Guo, Yoon Kim, and Alexander Rush · 2020
Later among the works it cites.
Span-based Semantic Parsing for Compositional Generalization
Jonathan Herzig and Jonathan Berant · 2020
Later among the works it cites.
Grounded PCFG Induction with Images
Lifeng Jin and William Schuler · 2020
Later among the works it cites.
COGS: A Compositional Generalization Challenge Based on Semantic Interpretation
Najoung Kim and Tal Linzen · 2020
Later among the works it cites.
Posterior Control of Blackbox Generation
Xiang Lisa Li and Alexander M. Rush · 2020
Later among the works it cites.
Compositional Generalization by Learning Analytical Expressions
Qian Liu, Shengnan An, Jian-Guang Lou, Bei Chen, Zeqi Lin, Yan Gao, Bin Zhou, Nanning Zheng, and Dongmei Zhang · 2020
Later among the works it cites.
Does Syntax Need to Grow on Trees? Sources of Hierarchical Inductive Bias in Sequence-to-Sequence Networks
R. Thomas McCoy, Robert Frank, and Tal Linzen · 2020
Later among the works it cites.
Learning Compositional Rules via Neural Program Synthesis
Maxwell I. Nye, Armando Solar-Lezama, Joshua B. Tenenbaum, and Brenden M. Lake · 2020
Later among the works it cites.
Improving Compositional Generalization in Semantic Parsing
Inbar Oren, Jonathan Herzig, Nitish Gupta, Matt Gardner, and Jonathan Berant · 2020
Later among the works it cites.
Making a Point: Pointer-Generator Transformers for Disjoint Vocabularies
Nikhil Prabhu and Katharina Kann · 2020
Later among the works it cites.
Torch-Struct: Deep Structured Prediction Library
Alexander M. Rush · 2020
Later among the works it cites.
Neural Data-to-Text Generation via Jointly Learning the Segmentation and Correspondence
Xiaoyu Shen, Ernie Chang, Hui Su, Cheng Niu, and Dietrich Klakow · 2020
Later among the works it cites.
Recursive Top-Down Production for Sentence Generation with Latent Trees
Shawn Tan, Yikang Shen, Alessandro Sordoni, Aaron Courville, and Timothy J. O’Donnell · 2020
Later among the works it cites.
Second-Order Unsupervised Neural Dependency Parsing
Songlin Yang, Yong Jiang, Wenjuan Han, and Kewei Tu · 2020
Later among the works it cites.
Visually Grounded Compound PCFGs
Yanpeng Zhao and Ivan Titov · 2020
Later among the works it cites.
The Return of Lexical Dependencies: Neural Lexicalized PCFGs
Hao Zhu, Yonatan Bisk, and Graham Neubig · 2020
Later among the works it cites.
Lexicon Learning for Few-Shot Neural Sequence Modeling
Ekin Akyürek and Jacob Andreas · 2021
Closest in time.
Learning to Recombine and Resample Data for Compositional Generalization
Ekin Akyürek, Afra Feyza Akyürek, and Jacob Andreas · 2021
Closest in time.
Can Transformers Jump Around Right in Natural Language? Assessing Performance Transfer from SCAN
Rahma Chaabouni, Roberto Dessi, and Eugene Kharitonov · 2021
Closest in time.
Meta-Learning to Compositionally Generalize
Henry Conklin, Bailin Wang, Kenny Smith, and Ivan Titov · 2021
Closest in time.
The Devil is in the Detail: Simple Tricks Improve Systematic Generalization of Transformers
Róbert Csordás, Kazuki Irie, and Jürgen Schmidhuber · 2021
Closest in time.
Revisiting Iterative Back-Translation from the Perspective of Compositional Generalization
Yinuo Guo, Hualei Zhu, Zeqi Lin, Bei Chen, Jian-Guang Lou, and Dongmei Zhang · 2021
Closest in time.
VLGrammar: Grounded Grammar Induction of Vision and Language
Yining Hong, Qing Li, Song-Chun Zhu, and Siyuan Huang · 2021
Closest in time.
Limitations of Autoregressive Models and Their Alternatives
Chu-Cheng Lin, Aaron Jaech, Xin Li, Matthew R. Gormley, and Jason Eisner · 2021
Closest in time.
Counterfactual Data Augmentation for Neural Machine Translation
Qi Liu, Matt Kusner, and Phil Blunsom · 2021
Closest in time.
StylePTB: A Compositional Benchmark for Fine-grained Controllable Text Style Transfer
Yiwei Lyu, Paul Pu Liang, Hai Pham, Eduard Hovy, Barnabás Póczos, Ruslan Salakhutdinov, and Louis-Philippe Morency · 2021
Closest in time.
Copy that! Editing Sequences by Copying Spans
Sheena Panthaplackel, Miltiadis Allamanis, and Marc Brockschmidt · 2021
Closest in time.
Compositional Generalization and Natural Language Variation: Can a Semantic Parsing Approach Handle Both?
Peter Shaw, Ming-Wei Chang, Panupong Pasupat, and Kristina Toutanova · 2021
Closest in time.
Structured Reordering for Modeling Latent Alignments in Sequence Transduction
Bailin Wang, Mirella Lapata, and Ivan Titov · 2021
Closest in time.
Generating (Formulaic) Text by Splicing Together Nearest Neighbors
Sam Wiseman, Arturs Backurs, and Karl Stratos · 2021
Closest in time.
Applying the Transformer to Character-level Transduction
Shijie Wu, Ryan Cotterell, and Mans Hulden · 2021
Closest in time.
Compositional Generalization for Neural Semantic Parsing via Span-level Supervised Attention
Pengcheng Yin, Hao Fang, Graham Neubig, Adam Pauls, Emmanouil Antonios Platanios, Yu Su, Sam Thomson, and Jacob Andreas · 2021
Closest in time.
Video-aided Unsupervised Grammar Induction
Songyang Zhang, Linfeng Song, Lifeng Jin, Kun Xu, Dong Yu, and Jiebo Luo · 2021
Closest in time.
An Empirical Study of Compound PCFGs
Yanpeng Zhao and Ivan Titov · 2021
Closest in time.