Fetching the paper…
Reading the bibliography…
In this work, we introduce a self-supervised behavior cloning transformer for text games, which are challenging benchmarks for multi-step reasoning in virtual environments.
Deep reinforcement learning with a combinatorial action space for predicting popular Reddit threads
Ji He, Mari Ostendorf, Xiaodong He, Jianshu Chen, Jianfeng Gao, Lihong Li, and Li Deng. 2016 · 2016
Earlier work this paper cites.
Textworld: A learning environment for text-based games
Marc-Alexandre Côté, Akos Kádár, Xingdi Yuan, Ben Kybartas, Tavian Barnes, Emery Fine, James Moore, Matthew Hausknecht, Layla El Asri, Mahmoud Adada, et al. 2018 · 2018
Earlier work this paper cites.
Behavioral cloning from observation
Faraz Torabi, Garrett Warnell, and Peter Stone. 2018 · 2018
Earlier work this paper cites.
Counting to explore and generalize in text-based games
Xingdi Yuan, Marc-Alexandre Côté, Alessandro Sordoni, Romain Laroche, Rémi Tachet des Combes, Matthew J. Hausknecht, and Adam Trischler. 2018 · 2018
Earlier work this paper cites.
Graph constrained reinforcement learning for natural language action spaces
Prithviraj Ammanabrolu and Matthew Hausknecht. 2020 · 2020
Earlier work this paper cites.
Interactive fiction games: A colossal adventure
Matthew Hausknecht, Prithviraj Ammanabrolu, Marc-Alexandre Côté, and Xingdi Yuan. 2020 · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Earlier work this paper cites.
Deep reinforcement learning with stacked hierarchical attention for text-based games
Yunqiu Xu, Meng Fang, Ling Chen, Yali Du, Joey Tianyi Zhou, and Chengqi Zhang. 2020 · 2020
Earlier work this paper cites.
Keep CALM and explore: Language models for action generation in text-based games
Shunyu Yao, Rohan Rao, Matthew Hausknecht, and Karthik Narasimhan. 2020 · 2020
Earlier work this paper cites.
Decision transformer: Reinforcement learning via sequence modeling
Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Misha Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch. 2021 · 2021
Cited alongside, same era.
Offline reinforcement learning as one big sequence modeling problem
Michael Janner, Qiyang Li, and Sergey Levine. 2021 · 2021
Cited alongside, same era.
Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and Baselines
Keerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla, Sadhana Kumaravel, Gerald Tesauro, Kartik Talamadupula, Mrinmaya Sachan, and Murray Campbell. 2021 · 2021
Cited alongside, same era.
ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Mohit Shridhar, Xingdi Yuan, Marc-Alexandre Côté, Yonatan Bisk, Adam Trischler, and Matthew Hausknecht. 2021 · 2021
Cited alongside, same era.
Pre-trained language models as prior knowledge for playing text-based games
Pre-trained language models for interactive decision-making
Shuang Li, Xavier Puig, Chris Paxton, Yilun Du, Clinton Wang, Linxi Fan, Tao Chen, De-An Huang, Ekin Akyürek, Anima Anandkumar, Jacob Andreas, Igor Mordatch, Antonio Torralba, and Yuke Zhu. 2022 · 2022
Later among the works it cites.
Impact of pretraining term frequencies on few-shot reasoning
Yasaman Razeghi, Robert L Logan IV, Matt Gardner, and Sameer Singh. 2022 · 2022
Later among the works it cites.
Multi-stage episodic control for strategic exploration in text games
Jens Tuyls, Shunyu Yao, Sham M. Kakade, and Karthik Narasimhan. 2022 · 2022
Later among the works it cites.
ScienceWorld: Is your agent smarter than a 5th grader?
Ruoyao Wang, Peter Jansen, Marc-Alexandre Côté, and Prithviraj Ammanabrolu. 2022 · 2022
Later among the works it cites.
TextWorldExpress: Simulating text games at one million steps per second
Peter Jansen and Marc-alexandre Cote. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ishika Singh, Gargi Singh, and Ashutosh Modi. 2021 · 2021
Cited alongside, same era.
Process-level representation of scientific protocols with interactive annotation
Ronen Tamari, Fan Bai, Alan Ritter, and Gabriel Stanovsky. 2021 · 2021
Cited alongside, same era.
Affordance extraction with an external knowledge database for text-based simulated environments
Paul Gelhausen, Marion Fischer, and George Peters. 2022 · 2022
Cited alongside, same era.
A systematic survey of text worlds as embodied natural language environments
Peter Jansen. 2022 · 2022
Cited alongside, same era.
Bill Yuchen Lin, Yicheng Fu, Karina Yang, Prithviraj Ammanabrolu, Faeze Brahman, Shiyu Huang, Chandra Bhagavatula, Yejin Choi, and Xiang Ren. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Behavior cloned transformers are neurosymbolic reasoners
Ruoyao Wang, Peter Jansen, Marc-Alexandre Côté, and Prithviraj Ammanabrolu. 2023 · 2023
Closest in time.