Fetching the paper…
Reading the bibliography…
We introduce BiGym, a new benchmark and learning environment for mobile bi-manual demo-driven robotic manipulation.
Planning and acting in partially observable stochastic domains
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra · 1998
Earlier work this paper cites.
Sarsop: Efficient point-based pomdp planning by approximating optimally reachable belief spaces
H. Kurniawati, D. Hsu, and W. S. Lee · 2009
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Earlier work this paper cites.
Openai gym, 2016
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Concrete problems in ai safety
D. Amodei, C. Olah, J. Steinhardt, P. Christiano, J. Schulman, and D. Mané · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
D. Hendrycks and K. Gimpel · 2016
Earlier work this paper cites.
J. L. Ba, J. R. Kiros, and G. E. Hinton · 2016
Earlier work this paper cites.
Transferring end-to-end visuomotor control from simulation to real world for a multi-stage task
S. James, A. J. Davison, and E. Johns · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Know what you don’t know: Unanswerable questions for squad
P. Rajpurkar, R. Jia, and P. Liang · 2018
Earlier work this paper cites.
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, D. d. L. Casas, D. Budden, A. Abdolmaleki, J. Merel, A. Lefrancq, et al · 2018
Earlier work this paper cites.
Deep q-learning from demonstrations
T. Hester, M. Vecerik, O. Pietquin, M. Lanctot, T. Schaul, B. Piot, D. Horgan, J. Quan, A. Sendonaris, I. Osband, et al · 2018
Earlier work this paper cites.
Sim-to-real reinforcement learning for deformable object manipulation
J. Matas, S. James, and A. J. Davison · 2018
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Cited alongside, same era.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine · 2020
Cited alongside, same era.
Rlbench: The robot learning benchmark & learning environment
S. James, Z. Ma, D. R. Arrojo, and A. J. Davison · 2020
Cited alongside, same era.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, R. Martín-Martín, A. Joshi, S. Nasiriany, and Y. Zhu · 2020
Cited alongside, same era.
Discriminative particle filter reinforcement learning for complex partial observations
X. Ma, P. Karkus, D. Hsu, W. S. Lee, and N. Ye · 2020
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
T. Z. Zhao, V. Kumar, S. Levine, and C. Finn · 2023
Later among the works it cites.
Perceiver-actor: A multi-task transformer for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2023
Later among the works it cites.
Act3d: Infinite resolution action detection transformer for robotic manipulation
T. Gervet, Z. Xian, N. Gkanatsios, and K. Fragkiadaki · 2023
Later among the works it cites.
Rvt: Robotic view transformer for 3d object manipulation
A. Goyal, J. Xu, Y. Guo, V. Blukis, Y.-W. Chao, and D. Fox · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, R. Martín-Martín, A. Joshi, S. Nasiriany, and Y. Zhu · 2020
Cited alongside, same era.
Awac: Accelerating online reinforcement learning with offline datasets
A. Nair, A. Gupta, M. Dalal, and S. Levine · 2020
Cited alongside, same era.
Ikea furniture assembly environment for long-horizon complex manipulation tasks
Y. Lee, E. S. Hu, and J. J. Lim · 2021
Cited alongside, same era.
Habitat 2.0: Training home assistants to rearrange their habitat
A. Szot, A. Clegg, E. Undersander, E. Wijmans, Y. Zhao, J. Turner, N. Maestre, M. Mukadam, D. S. Chaplot, O. Maksymets, et al · 2021
Cited alongside, same era.
Softgym: Benchmarking deep reinforcement learning for deformable object manipulation
X. Lin, Y. Wang, J. Olkin, and D. Held · 2021
Cited alongside, same era.
Maniskill: Generalizable manipulation skill benchmark with large-scale demonstrations
T. Mu, Z. Ling, F. Xiang, D. Yang, X. Li, S. Tao, Z. Huang, Z. Jia, and H. Su · 2021
Cited alongside, same era.
Mastering visual continuous control: Improved data-augmented reinforcement learning
D. Yarats, R. Fergus, A. Lazaric, and L. Pinto · 2021
Cited alongside, same era.
Later among the works it cites.
Locomujoco: A comprehensive imitation learning benchmark for locomotion
F. Al-Hafez, G. Zhao, J. Peters, and D. Tateo · 2023
Later among the works it cites.
Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation
C. Li, R. Zhang, J. Wong, C. Gokmen, S. Srivastava, R. Martín-Martín, C. Wang, G. Levine, M. Lingelbach, J. Sun, et al · 2023
Later among the works it cites.
Furniturebench: Reproducible real-world benchmark for long-horizon complex manipulation
M. Heo, Y. Lee, D. Lee, and J. J. Lim · 2023
Later among the works it cites.
Bi-dexhands: Towards human-level bimanual dexterous manipulation
Y. Chen, Y. Geng, F. Zhong, J. Ji, J. Jiang, Z. Lu, H. Dong, and Y. Yang · 2023
Later among the works it cites.
Robopianist: Dexterous piano playing with deep reinforcement learning
K. Zakka, P. Wu, L. Smith, N. Gileadi, T. Howell, X. B. Peng, S. Singh, Y. Tassa, P. Florence, A. Zeng, et al · 2023
Later among the works it cites.
Maniskill2: A unified benchmark for generalizable manipulation skills
J. Gu, F. Xiang, X. Li, Z. Ling, X. Liu, T. Mu, Y. Tang, S. Tao, X. Wei, Y. Yao, et al · 2023
Later among the works it cites.
Gymnasium, Mar. 2023
M. Towers, J. K. Terry, A. Kwiatkowski, J. U. Balis, G. d. Cola, T. Deleu, M. Goulão, A. Kallinteris, A. KG, M. Krimmel, R. Perez-Vicente, A. Pierré, S. Schulhoff, J. J. Tai, A. T. J. Shen, and O. G. Younis · 2023
Later among the works it cites.
Waypoint-based imitation learning for robotic manipulation
L. X. Shi, A. Sharma, T. Z. Zhao, and C. Finn · 2023
Later among the works it cites.
Hierarchical diffusion policy for kinematics-aware multi-task robotic manipulation
X. Ma, S. Patidar, I. Haughton, and S. James · 2024
Closest in time.
Render and diffuse: Aligning image and action spaces for diffusion-based behaviour cloning
V. Vosylius, Y. Seo, J. Uruç, and S. James · 2024
Closest in time.
Humanoidbench: Simulated humanoid benchmark for whole-body locomotion and manipulation
C. Sferrazza, D.-M. Huang, X. Lin, Y. Lee, and P. Abbeel · 2024
Closest in time.
Robocasa: Large-scale simulation of everyday tasks for generalist robots
S. Nasiriany, A. Maddukuri, L. Zhang, A. Parikh, A. Lo, A. Joshi, A. Mandlekar, and Y. Zhu · 2024
Closest in time.
Continuous control with coarse-to-fine reinforcement learning
Y. Seo, J. Uruç, and S. James · 2024
Closest in time.