Fetching the paper…
Reading the bibliography…
Human beings, even small children, quickly become adept at figuring out how to use applications on their mobile devices.
“Policy invariance under reward transformations: Theory and application to reward shaping”
Andrew Ng, Daishi Harada and Stuart Russell · 1999
Earlier work this paper cites.
“Continuous control with deep reinforcement learning”
Timothy Lillicrap et al · 2015
Earlier work this paper cites.
“Human-level control through deep reinforcement learning”
Volodymyr Mnih et al · 2015
Earlier work this paper cites.
“Trust Region Policy Optimization”
John Schulman et al · 2015
Earlier work this paper cites.
“Mastering the game of Go with deep neural networks and tree search”
David Silver et al · 2016
Earlier work this paper cites.
Greg Brockman et al · 2016
Earlier work this paper cites.
“Docker-Android”
Budi Utomo · 2016
Earlier work this paper cites.
“A deep hierarchical approach to lifelong learning in minecraft”
Chen Tessler et al · 2017
Earlier work this paper cites.
“Mastering the Game of Go without Human Knowledge”
David Silver et al · 2017
Earlier work this paper cites.
“World of Bits: An Open-domain Platform for Web-based Agents”
Tianlin Shi et al · 2017
Cited alongside, same era.
“Proximal policy optimization algorithms”
John Schulman et al · 2017
Cited alongside, same era.
“OpenAI Baselines”
Prafulla Dhariwal et al · 2017
Cited alongside, same era.
“Reinforcement Learning on Web Interfaces using Workflow-Guided Exploration”
Evan Liu et al · 2018
Cited alongside, same era.
“Learning to navigate the web”
Izzeddin Gur, Ulrich Rueckert, Aleksandra Faust and Dilek Hakkani-Tur · 2018
Cited alongside, same era.
“QBE: QLearning-based exploration of android applications”
Yavuz Koroglu et al · 2018
Yavuz Koroglu and Alper Sen · 2019
Later among the works it cites.
“BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2019
Later among the works it cites.
“Learning reward machines for partially observable reinforcement learning”
Rodrigo Toro et al · 2019
Later among the works it cites.
“An empirical investigation of the challenges of real-world reinforcement learning”
Gabriel Dulac-Arnold et al · 2020
Later among the works it cites.
“Mapping Natural Language Instructions to Mobile UI Action Sequences”
Yang Li et al · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Reinforcement learning: An introduction”
Richard Sutton and Andrew Barto · 2018
Cited alongside, same era.
“DOM-Q-NET: Grounded RL on Structured Language”
Sheng Jia, Jamie Kiros and Jimmy Ba · 2019
Cited alongside, same era.
“Learning user interface element interactions”
Christian Degott, Nataniel Borges and Andreas Zeller · 2019
Cited alongside, same era.
“Reinforcement learning based curiosity-driven testing of Android applications”
Minxue Pan et al · 2020
Later among the works it cites.
“Widget captioning: generating natural language description for mobile user interface elements”
Yang Li et al · 2020
Later among the works it cites.
“VASTA: a vision and language-assisted smartphone task automation system”
Alborz Sereshkeh et al · 2020
Later among the works it cites.
“AppBuddy: Learning to Accomplish Tasks in Mobile Apps via Reinforcement Learning”
Maayan Shvo et al · 2021
Closest in time.