Fetching the paper…
Reading the bibliography…
We explore unconstrained natural language feedback as a learning signal for artificial agents.
From Language to Goals: Inverse Reinforcement Learning for Vision-Based Instruction Following
Fu, J.; Korattikara, A.; Levine, S.; and Guadarrama, S. 2019 · 1902
Earlier work this paper cites.
MeetUp! A Corpus of Joint Activity Dialogues in a Visual Environment
Ilinykh, N.; Zarrieß, S.; and Schlangen, D. 2019 · 1907
Earlier work this paper cites.
Why Build an Assistant in Minecraft?
Szlam, A.; Gray, J.; Srinet, K.; Jernite, Y.; Joulin, A.; Synnaeve, G.; Kiela, D.; Yu, H.; Chen, Z.; Goyal, S.; Guo, D.; Rothermel, D.; Zitnick, C. L.; and Weston, J. 2019 · 1907
Earlier work this paper cites.
Thomason, J.; Murray, M.; Cakmak, M.; and Zettlemoyer, L. 2019 · 1907
Earlier work this paper cites.
Self-Educated Language Agent With Hindsight Experience Replay For Instruction Following
Cideron, G.; Seurin, M.; Strub, F.; and Pietquin, O. 2019 · 1910
Earlier work this paper cites.
Learning to Interpret Natural Language Commands through Human-Robot Dialog
Thomason, J.; Zhang, S.; Mooney, R.; and Stone, P. 2015 · 1929
Earlier work this paper cites.
Section of mathematics and engineering: Some selected quick and easy methods of statistical analysis
Tukey, J. W. 1953 · 1953
Earlier work this paper cites.
Conditional logit analysis of qualitative choice behavior
McFadden, D. 1974 · 1974
Earlier work this paper cites.
Logic and Conversation
Grice, H. P. 1975 · 1975
Earlier work this paper cites.
The symbol grounding problem
Harnad, S. 1990 · 1990
Earlier work this paper cites.
Incorporating advice into agents that learn from reinforcements
Maclin, R.; and Shavlik, J. W. 1994 · 1994
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Puterman, M. L. 1994 · 1994
Earlier work this paper cites.
Using Language
Clark, H. H. 1996 · 1996
Earlier work this paper cites.
PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards
Goyal, P.; Niekum, S.; and Mooney, R. J. 2020 · 2002
Earlier work this paper cites.
Reward-rational (implicit) choice: A unifying formalism for reward learning
Jeon, H. J.; Milli, S.; and Dragan, A. D. 2020 · 2002
Earlier work this paper cites.
Apprenticeship Learning via Inverse Reinforcement Learning
Abbeel, P.; and Ng, A. Y. 2004 · 2004
Earlier work this paper cites.
Mining and summarizing customer reviews
Hu, M.; and Liu, B. 2004 · 2004
Earlier work this paper cites.
Guiding a reinforcement learner with natural language advice: Initial results in RoboCup soccer
Kuhlmann, G.; Stone, P.; Mooney, R.; and Shavlik, J. 2004 · 2004
Earlier work this paper cites.
The Power of Feedback
Hattie, J.; and Timperley, H. 2007 · 2007
Earlier work this paper cites.
Bayesian Inverse Reinforcement Learning
Ramachandran, D.; and Amir, E. 2007 · 2007
Earlier work this paper cites.
Learning to Connect Language and Perception
Mooney, R. J. 2008 · 2008
Earlier work this paper cites.
Focus on Formative Feedback
Shute, V. J. 2008 · 2008
Earlier work this paper cites.
Teachable robots: Understanding human teaching behavior to build more effective robot learners
Thomaz, A. L.; and Breazeal, C. 2008 · 2008
Earlier work this paper cites.
Inverse Reinforcement Learning with Natural Language Goals
Zhou, L.; and Small, K. 2020 · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
Argall, B. D.; Chernova, S.; Veloso, M.; and Browning, B. 2009 · 2009
Cited alongside, same era.
How people talk when teaching a robot
Kim, E. S.; Leyzberg, D.; Tsui, K. M.; and Scassellati, B. 2009 · 2009
Cited alongside, same era.
Interactively Shaping Agents via Human Reinforcement: The TAMER Framework
Knox, W. B.; and Stone, P. 2009 · 2009
Cited alongside, same era.
Effects of differential feedback on students’ examination performance
Lipnevich, A.; and Smith, J. 2009 · 2009
Cited alongside, same era.
Reinforcement Learning via Practice and Critique Advice
Judah, K.; Roy, S.; Fern, A.; and Dietterich, T. G. 2010 · 2010
Cited alongside, same era.
Efficient reductions for imitation learning
Ross, S.; and Bagnell, D. 2010 · 2010
Cited alongside, same era.
Deep Reinforcement Learning from Human Preferences
Christiano, P. F.; Leike, J.; Brown, T.; Martic, M.; Legg, S.; and Amodei, D. 2017 · 2017
Later among the works it cites.
Learning Symmetric Collaborative Dialogue Agents with Dynamic Knowledge Graph Embeddings
He, H.; Balakrishnan, A.; Eric, M.; and Liang, P. 2017 · 2017
Later among the works it cites.
Beating Atari with Natural Language Guided Reinforcement Learning
Kaplan, R.; Sauer, C.; and Sosa, A. 2017 · 2017
Later among the works it cites.
lmerTest package: tests in linear mixed effects models
Kuznetsova, A.; Brockhoff, P. B.; and Christensen, R. 2017 · 2017
Later among the works it cites.
Teaching Machines to Describe Images with Natural Language Feedback
Ling, H.; and Fidler, S. 2017 · 2017
Later among the works it cites.
Interactive Learning from Policy-Dependent Human Feedback
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bootstrapping Semantic Parsers from Conversations
Artzi, Y.; and Zettlemoyer, L. 2011 · 2011
Cited alongside, same era.
Modeling Expert Effects and Common Ground Using Questions Under Discussion
Djalali, A.; Clausen, D.; Lauer, S.; Schultz, K.; and Potts, C. 2011 · 2011
Cited alongside, same era.
Target-dependent twitter sentiment classification
Jiang, L.; Yu, M.; Zhou, M.; Liu, X.; and Zhao, T. 2011 · 2011
Cited alongside, same era.
Understanding Natural Language Commands for Robotic Navigation and Mobile Manipulation
Tellex, S.; Kollar, T.; Dickerson, S.; Walter, M. R.; Banerjee, A. G.; Teller, S.; and Roy, N. 2011 · 2011
Cited alongside, same era.
Corpus Evidence for Preference-Driven Interpretation
Djalali, A.; Lauer, S.; and Potts, C. 2012 · 2012
Cited alongside, same era.
Goal-Driven Answers in the Cards Dialogue Corpus
Potts, C. 2012 · 2012
Cited alongside, same era.
MacGlashan, J.; Ho, M. K.; Loftin, R.; Peng, B.; Wang, G.; Roberts, D. L.; Taylor, M. E.; and Littman, M. L. 2017 · 2017
Later among the works it cites.
Joint Concept Learning and Semantic Parsing from Natural Language Explanations
Srivastava, S.; Labutov, I.; and Mitchell, T. 2017 · 2017
Later among the works it cites.
How Players Speak to an Intelligent Game Character Using Natural Language Messages
Allison, F.; Luger, E.; and Hofmann, K. 2018 · 2018
Later among the works it cites.
BabyAI: First Steps Towards Grounded Language Learning With a Human In the Loop
Chevalier-Boisvert, M.; Bahdanau, D.; Lahlou, S.; Willems, L.; Saharia, C.; Nguyen, T. H.; and Bengio, Y. 2018 · 2018
Later among the works it cites.
Training Classifiers with Natural Language Explanations
Hancock, B.; Varma, P.; Wang, S.; Bringmann, M.; Liang, P.; and Ré, C. 2018 · 2018
Later among the works it cites.
Grounding language for transfer in deep reinforcement learning
Narasimhan, K.; Barzilay, R.; and Jaakkola, T. 2018 · 2018
Later among the works it cites.
Learning to Parse Natural Language to Grounded Reward Functions with Weak Supervision
Williams, E. C.; Gopalan, N.; Rhee, M.; and Tellex, S. 2018 · 2018
Later among the works it cites.
Learning to Understand Goal Specifications by Modelling Reward
Bahdanau, D.; Hill, F.; Leike, J.; Hughes, E.; Hosseini, S. A.; Kohli, P.; and Grefenstette, E. 2019 · 2019
Later among the works it cites.
Using natural language for reward shaping in reinforcement learning
Goyal, P.; Niekum, S.; and Mooney, R. J. 2019 · 2019
Later among the works it cites.
A Survey of Reinforcement Learning Informed by Natural Language
Luketina, J.; Nardelli, N.; Farquhar, G.; Foerster, J.; Andreas, J.; Grefenstette, E.; Whiteson, S.; and Rocktäschel, T. 2019 · 2019
Later among the works it cites.
Executing Instructions in Situated Collaborative Interactions
Suhr, A.; Yan, C.; Schluger, J.; Yu, S.; Khader, H.; Mouallem, M.; Zhang, I.; and Artzi, Y. 2019 · 2019
Later among the works it cites.
A Natural Language Corpus of Common Grounding under Continuous and Partially-Observable Context
Udagawa, T.; and Aizawa, A. 2019 · 2019
Later among the works it cites.
Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language Navigation
Wang, X.; Huang, Q.; Çelikyilmaz, A.; Gao, J.; Shen, D.; Wang, Y.-F.; Wang, W. Y.; and Zhang, L. 2019 · 2019
Later among the works it cites.
BERT Post-Training for Review Reading Comprehension and Aspect-based Sentiment Analysis
Xu, H.; Liu, B.; Shu, L.; and Philip, S. Y. 2019 · 2019
Later among the works it cites.
Sentiment Analysis: Mining Opinions, Sentiments, and Emotions
Liu, B. 2020 · 2020
Closest in time.
Conjugate Bayesian analysis of the Gaussian distribution
Murphy, K. 2007 · 2020
Closest in time.
Robots That Use Language
Tellex, S.; Gopalan, N.; Kress-Gazit, H.; and Matuszek, C. 2020 · 2020
Closest in time.
Jointly Improving Parsing and Perception for Natural Language Commands through Human-Robot Dialog
Thomason, J.; Padmakumar, A.; Sinapov, J.; Walker, N.; Jiang, Y.; Yedidsion, H.; Hart, J.; Stone, P.; and Mooney, R. J. 2020 · 2020
Closest in time.