Fetching the paper…
Reading the bibliography…
Grounding natural language instructions on the web to perform previously unseen tasks enables accessibility and automation.
Complete counterbalancing of immediate sequential effects in a latin square design
James V. Bradley. 1958 · 1958
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
From natural language instructions to complex processes: Issues in chaining trigger action rules
Nobuhiro Ito, Yuya Suzuki, and Akiko Aizawa. 2020 · 2001
Earlier work this paper cites.
Mapping natural language instructions to mobile ui action sequences
Yang Li, Jiacong He, Xin Zhou, Yuan Zhang, and Jason Baldridge. 2020 · 2005
Earlier work this paper cites.
Plow: A collaborative task learning agent
James Allen, Nathanael Chambers, George Ferguson, Lucian Galescu, Hyuckchul Jung, Mary Swift, and William Taysom. 2007 · 2007
Earlier work this paper cites.
Multi-modal end-user programming of web-based virtual assistant skills
Michael H Fischer, Giovanni Campagna, Euirim Choi, and Monica S Lam. 2020 · 2008
Earlier work this paper cites.
Coscripter: Automating & sharing how-to knowledge in the enterprise
Gilly Leshed, Eben M. Haber, Tara Matthews, and Tessa Lau. 2008 · 2008
Earlier work this paper cites.
Learning adaptive language interfaces through decomposition
Siddharth Karamcheti, Dorsa Sadigh, and Percy Liang. 2020 · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
David L. Chen and Raymond J. Mooney. 2011 · 2011
Earlier work this paper cites.
Arramon: A joint navigation-assembly instruction interpretation task in dynamic environments
Hyounghun Kim, Abhay Zala, Graham Burri, Hao Tan, and Mohit Bansal. 2020 · 2011
Earlier work this paper cites.
Multi-modal conversational search and browse
Larry Heck, Dilek Hakkani-Tür, Madhu Chinthakunta, Gokhan Tur, Rukmini Iyer, Partha Parthasarathy, Lisa Stifelman, Elizabeth Shriberg, and Ashley Fidler. 2013 · 2013
Earlier work this paper cites.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Sequence to sequence – video to text
Subhashini Venugopalan, Marcus Rohrbach, Jeff Donahue, Raymond J. Mooney, Trevor Darrell, and Kate Saenko. 2015 · 2015
Cited alongside, same era.
Learning multi-modal grounded linguistic semantics by playing "i spy"
Jesse Thomason, Jivko Sinapov, Maxwell Svetlik, Peter Stone, and Raymond J Mooney. 2016 · 2016
Cited alongside, same era.
Sugilite: Creating multimodal smartphone automation by demonstration
Toby Jia-Jun Li, Amos Azaria, and Brad A. Myers. 2017 · 2017
Cited alongside, same era.
Mapping instructions and visual observations to actions with reinforcement learning
Dipendra Misra, John Langford, and Yoav Artzi. 2017 · 2017
Cited alongside, same era.
Mapping natural language commands to web elements
Panupong Pasupat, Tian-Shun Jiang, Evan Liu, Kelvin Guu, and Percy Liang. 2018 · 2018
Later among the works it cites.
A deep reinforced model for abstractive summarization
Romain Paulus, Caiming Xiong, and Richard Socher. 2018 · 2018
Later among the works it cites.
Situational impairments during mobile interaction
Zhanna Sarsenbayeva. 2018 · 2018
Later among the works it cites.
Share of customers in the united states who have contacted customer service for any reason in the past month from 2015 to 2018
Statista Research Department. 2019 · 2018
Later among the works it cites.
Genie: A generator of natural language semantic parsers for virtual assistant commands
Giovanni Campagna, Silei Xu, Mehrad Moradshahi, Richard Socher, and Monica S. Lam. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Abigail See, Peter J Liu, and Christopher D Manning. 2017 · 2017
Cited alongside, same era.
Grounding visual explanations
Lisa Anne Hendricks, Ronghang Hu, Trevor Darrell, and Zeynep Akata. 2018 · 2018
Cited alongside, same era.
Kite: Building conversational bots from mobile apps
Toby Jia-Jun Li and Oriana Riva. 2018 · 2018
Cited alongside, same era.
Reinforcement learning on web interfaces using workflow-guided exploration
Evan Zheran Liu, Kelvin Guu, Panupong Pasupat, Tianlin Shi, and Percy Liang. 2018 · 2018
Cited alongside, same era.
Edit me: A corpus and a framework for understanding natural language image editing
Ramesh Manuvinakurike, Jacqueline Brixey, Trung Bui, Walter Chang, Doo Soon Kim, Ron Artstein, and Kallirroi Georgila. 2018 · 2018
Cited alongside, same era.
Learning to navigate in cities without a map
Piotr Mirowski, Matt Grimes, Mateusz Malinowski, Karl Moritz Hermann, Keith Anderson, Denis Teplyashin, Karen Simonyan, Andrew Zisserman, Raia Hadsell, et al. 2018 · 2018
Cited alongside, same era.
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
Improving grounded natural language understanding through human-robot dialog
Jesse Thomason, Aishwarya Padmakumar, Jivko Sinapov, Nick Walker, Yuqian Jiang, Harel Yedidsion, Justin Hart, Peter Stone, and Raymond J Mooney. 2019 · 2019
Later among the works it cites.
Geno: A developer tool for authoring multimodal interaction on existing web applications
Ritam Jyoti Sarmah, Yunpeng Ding, Di Wang, Cheuk Yin Phipson Lee, Toby Jia-Jun Li, and Xiang ’Anthony’ Chen. 2020 · 2020
Later among the works it cites.
Schema2QA: High-quality and low-cost Q&A agents for the structured web
Silei Xu, Giovanni Campagna, Jian Li, and Monica S Lam. 2020 · 2020
Later among the works it cites.