Fetching the paper…
Reading the bibliography…
Recent studies have delved into constructing generalist agents for open-world environments like Minecraft.
Automatic decomposition of planned assembly sequences into skill primitives
Heiko Mosemann and Friedrich M Wahl · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, et al · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
Robot skills for manufacturing: From concept to industrial deployment
Mikkel Rath Pedersen, Lazaros Nalpantidis, et al · 2016
Earlier work this paper cites.
Zero-shot task generalization with multi-task deep reinforcement learning
Junhyuk Oh, Satinder Singh, et al · 2017
Earlier work this paper cites.
A deep hierarchical approach to lifelong learning in minecraft
Chen Tessler, Shahar Givony, et al · 2017
Earlier work this paper cites.
Babyai: A platform to study the sample efficiency of grounded language learning
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, et al · 2018
Earlier work this paper cites.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, et al · 2018
Earlier work this paper cites.
Minerl: a large-scale dataset of minecraft demonstrations
William H Guss, Brandon Houghton, et al · 2019
Earlier work this paper cites.
Obstacle tower: A generalization challenge in vision, control, and planning
Arthur Juliani, Ahmed Khalifa, et al · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych · 2019
Earlier work this paper cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, et al · 2019
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, et al · 2020
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, et al · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Phillip Wallis, et al · 2021
Earlier work this paper cites.
Action candidate based clipped double q-learning for discrete and continuous action tasks
Haobo Jiang, Jin Xie, and Jian Yang · 2021
Earlier work this paper cites.
igibson 1.0: A simulation environment for interactive tasks in large realistic scenes
Bokui Shen, Fei Xia, et al · 2021
Earlier work this paper cites.
Video pretraining (vpt): Learning to act by watching unlabeled online videos
Bowen Baker, Ilge Akkaya, et al · 2022
Earlier work this paper cites.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Linxi Fan, Guanzhi Wang, et al · 2022
Earlier work this paper cites.
Inner monologue: Embodied reasoning through planning with language models
Wenlong Huang, Fei Xia, et al · 2022
Earlier work this paper cites.
Action candidate driven clipped double q-learning for discrete and continuous action tasks
Haobo Jiang, Guangyu Li, et al · 2022
Earlier work this paper cites.
Juewu-mc: Playing minecraft with sample-efficient hierarchical reinforcement learning
Zichuan Lin, Junyou Li, et al · 2022
Earlier work this paper cites.
Seihai: A sample-efficient hierarchical ai for the minerl competition
Hangyu Mao, Chao Wang, et al · 2022
Earlier work this paper cites.
Scott E. Reed, Konrad Zolna, et al · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, et al · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, et al · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, et al · 2023
Cited alongside, same era.
Open-world multi-task control through goal-aware representation learning and adaptive horizon prediction
Shaofei Cai, Zihao Wang, et al · 2023
Cited alongside, same era.
Mind2web: Towards a generalist agent for the web
Xiang Deng, Yu Gu, et al · 2023
Cited alongside, same era.
Jarvis-1: Open-world multi-task agents with memory-augmented multimodal language models
Zihao Wang, Shaofei Cai, et al · 2023
Later among the works it cites.
Baichuan 2: Open large-scale language models
Aiyuan Yang, Bin Xiao, et al · 2023
Later among the works it cites.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, et al · 2023
Later among the works it cites.
Skill reinforcement learning and planning for open-world long-horizon tasks
Haoqi Yuan, Chi Zhang, et al · 2023
Later among the works it cites.
Creative agents: Empowering agents with imagination for creative tasks
Chi Zhang, Penglin Cai, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ziluo Ding, Hao Luo, et al · 2023
Cited alongside, same era.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, et al · 2023
Cited alongside, same era.
Guiding pretraining in reinforcement learning with large language models
Yuqing Du, Olivia Watkins, et al · 2023
Cited alongside, same era.
Llama rider: Spurring large language models to explore the open world
Yicheng Feng, Yuxuan Wang, et al · 2023
Cited alongside, same era.
Mastering diverse domains through world models
Danijar Hafner, Jurgis Pasukonis, et al · 2023
Cited alongside, same era.
Auto mc-reward: Automated dense reward design with large language models for minecraft
Hao Li, Xue Yang, et al · 2023
Cited alongside, same era.
Code as policies: Language model programs for embodied control
Jacky Liang, Wenlong Huang, et al · 2023
Cited alongside, same era.
Later among the works it cites.
Huatuogpt, towards taming language model to be a doctor
Hongbo Zhang, Junying Chen, et al · 2023
Later among the works it cites.
See and think: Embodied agent in virtual environment
Zhonghan Zhao, Wenhao Chai, et al · 2023
Later among the works it cites.
Steve-eye: Equipping llm-based embodied agents with visual perception in open worlds
Sipeng Zheng, Jiazheng Liu, et al · 2023
Later among the works it cites.
Xizhou Zhu, Yuntao Chen, et al · 2023
Later among the works it cites.
Rocket-1: Master open-world interaction with visual-temporal context prompting
Shaofei Cai, Zihao Wang, et al · 2024
Closest in time.
OpenWebAgent: An open toolkit to enable web agents on large language models
Iat Long Iong, Xiao Liu, et al · 2024
Closest in time.
Autowebglm: A large language model-based web navigating agent
Hanyu Lai, Xiao Liu, et al · 2024
Closest in time.
Optimus-1: Hybrid multimodal memory empowered agents excel in long-horizon tasks
Zaijing Li, Yuquan Xie, et al · 2024
Closest in time.
Interaction pattern disentangling for multi-agent reinforcement learning
Shunyu Liu, Jie Song, et al · 2024
Closest in time.
WebLINX: Real-world website navigation with multi-turn dialogue
Xing Han Lu, Zdeněk Kasner, et al · 2024
Closest in time.
A survey on large language model based autonomous agents
Lei Wang, Chen Ma, et al · 2024
Closest in time.
Omnijarvis: Unified vision-language-action tokenization enables open-world instruction following agents
Zihao Wang, Shaofei Cai, et al · 2024
Closest in time.
A survey on game playing agents and large models: Methods, applications, and challenges
Xinrun Xu, Yuxin Wang, et al · 2024
Closest in time.
An Yang, Baosong Yang, et al · 2024
Closest in time.
Zhongjing: Enhancing the chinese medical capabilities of large language model through expert feedback and real-world multi-turn dialogue
Songhua Yang, Hanjie Zhao, et al · 2024
Closest in time.
Adam: An embodied causal agent in open-world environments
Shu Yu and Chaochao Lu · 2024
Closest in time.
Minedreamer: Learning to follow instructions via chain-of-imagination for simulated-world control
Enshen Zhou, Yiran Qin, et al · 2024
Closest in time.
Navgpt: Explicit reasoning in vision-and-language navigation with large language models
Gengze Zhou, Yicong Hong, et al · 2024
Closest in time.