Fetching the paper…
Reading the bibliography…
Our goal is to learn a semantic parser that maps natural language utterances into executable programs when only indirect supervision is available: examples are labeled with the correct execution result, but not the program itself.
Maximum likelihood from incomplete data via the EM algorithm
A. P. Dempster, L. N. M., and R. D. B. 1977 · 1977
Earlier work this paper cites.
Function optimization using connectionist reinforcement learning algorithms
R. J. Williams and J. Peng. 1991 · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams. 1992 · 1992
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
R. Sutton, D. McAllester, S. Singh, and Y. Mansour. 1999 · 1999
Earlier work this paper cites.
Optimal Learning: Computational procedures for Bayes-adaptive Markov decision processes
M. O. Duff. 2002 · 2002
Earlier work this paper cites.
Near-optimal reinforcement learning in polynomial time
M. Kearns and S. Singh. 2002 · 2002
Earlier work this paper cites.
Minimum error rate training in statistical machine translation
F. J. Och. 2003 · 2003
Earlier work this paper cites.
Efficient selectivity and backup operators in Monte-Carlo tree search
R. Coulom. 2006 · 2006
Earlier work this paper cites.
Minimum risk annealing for training log-linear models
D. A. Smith and J. Eisner. 2006 · 2006
Earlier work this paper cites.
Reinforcement learning for mapping instructions to actions
S. Branavan, H. Chen, L. S. Zettlemoyer, and R. Barzilay. 2009 · 2009
Earlier work this paper cites.
Driving semantic parsing from the world’s response
J. Clarke, D. Goldwasser, M. Chang, and D. Roth. 2010 · 2010
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio. 2010 · 2010
Earlier work this paper cites.
Bootstrapping semantic parsers from conversations
Y. Artzi and L. Zettlemoyer. 2011 · 2011
Earlier work this paper cites.
Learning dependency-based compositional semantics
P. Liang, M. I. Jordan, and D. Klein. 2011 · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. Gordon, and A. Bagnell. 2011 · 2011
Earlier work this paper cites.
Weakly supervised training of semantic parsers
J. Krishnamurthy and T. Mitchell. 2012 · 2012
Cited alongside, same era.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Y. Artzi and L. Zettlemoyer. 2013 · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
D. Kingma and J. Ba. 2014 · 2014
Cited alongside, same era.
Motor Skill Learning with Local Trajectory Methods
S. Levine. 2014 · 2014
Cited alongside, same era.
Generalization and exploration via randomized value functions
I. Osband, B. V. Roy, and Z. Wen. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Neural enquirer: Learning to query tables
P. Yin, Z. Lu, H. Li, and B. Kao. 2015 · 2015
Later among the works it cites.
Unifying count-based exploration and intrinsic motivation
M. Bellemare, S. Srinivasan, G. Ostrovski, T. Schaul, D. Saxton, and R. Munos. 2016 · 2016
Later among the works it cites.
Deep reinforcement learning for mention-ranking coreference models
K. Clark and C. D. Manning. 2016 · 2016
Later among the works it cites.
Language to logical form with neural attention
L. Dong and M. Lapata. 2016 · 2016
Later among the works it cites.
Data recombination for neural semantic parsing
R. Jia and P. Liang. 2016 · 2016
Later among the works it cites.
Deep reinforcement learning for dialogue generation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Pennington, R. Socher, and C. D. Manning. 2014 · 2014
Cited alongside, same era.
Large-scale semantic parsing without question-answer pairs
S. Reddy, M. Lapata, and M. Steedman. 2014 · 2014
Cited alongside, same era.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. J. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Józefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. G. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. A. Tucker, V. Vanhoucke, V. Vasudevan, F. B. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng. 2015 · 2015
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio. 2015 · 2015
Cited alongside, same era.
Scheduled sampling for sequence prediction with recurrent neural networks
S. Bengio, O. Vinyals, N. Jaitly, and N. Shazeer. 2015 · 2015
Cited alongside, same era.
Language understanding for text-based games using deep reinforcement learning
K. Narasimhan, T. Kulkarni, and R. Barzilay. 2015 · 2015
Cited alongside, same era.
Compositional semantic parsing on semi-structured tables
P. Pasupat and P. Liang. 2015 · 2015
Cited alongside, same era.
J. Li, W. Monroe, A. Ritter, D. Jurafsky, M. Galley, and J. Gao. 2016 · 2016
Later among the works it cites.
Simpler context-dependent logical forms via model projections
R. Long, P. Pasupat, and P. Liang. 2016 · 2016
Later among the works it cites.
Improving policy gradient by exploring under-appreciated rewards
O. Nachum, M. Norouzi, and D. Schuurmans. 2016 · 2016
Later among the works it cites.
Neural programmer: Inducing latent programs with gradient descent
A. Neelakantan, Q. V. Le, and I. Sutskever. 2016 · 2016
Later among the works it cites.
Reward augmented maximum likelihood for neural structured prediction
M. Norouzi, S. Bengio, N. Jaitly, M. Schuster, Y. Wu, D. Schuurmans, et al. 2016 · 2016
Later among the works it cites.
Deep exploration via bootstrapped DQN
I. Osband, C. Blundell, A. Pritzel, and B. V. Roy. 2016 · 2016
Later among the works it cites.
Inferring logical forms from denotations
P. Pasupat and P. Liang. 2016 · 2016
Later among the works it cites.
Programming with a differentiable forth interpreter
S. Riedel, M. Bosnjak, and T. Rocktäschel. 2016 · 2016
Later among the works it cites.
Neural symbolic machines: Learning semantic parsers on Freebase with weak supervision
C. Liang, J. Berant, Q. Le, and K. D. F. N. Lao. 2017 · 2017
Closest in time.