Fetching the paper…
Reading the bibliography…
Assembly of multi-part physical structures is both a valuable end product for autonomous robotics, as well as a valuable diagnostic task for open-ended training of embodied intelligent agents.
Principles of object perception
Spelke, E. S · 1990
Earlier work this paper cites.
Rapidly-exploring random trees: Progress and prospects
LaValle, S. M., Kuffner, J. J., Donald, B., et al · 2001
Earlier work this paper cites.
Acme: A research framework for distributed reinforcement learning
Hoffman, M., Shahriari, B., Aslanides, J., Barth-Maron, G., Behbahani, F., Norman, T., Abdolmaleki, A., Cassirer, A., Yang, F., Baumli, K., Henderson, S., Novikov, A., Colmenarejo, S. G., Cabi, S., Gulcehre, C., Paine, T. L., Cowie, A., Wang, Z., Piot, B., and de Freitas, N · 2006
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Earlier work this paper cites.
High-dimensional continuous control using generalized advantage estimation
Schulman, J., Moritz, P., Levine, S., Jordan, M., and Abbeel, P · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
Veličković, P., Cucurull, G., Casanova, A., Romero, A., Lio, P., and Bengio, Y · 2017
Earlier work this paper cites.
Relational inductive biases, deep learning, and graph networks
Battaglia, P. W., Hamrick, J. B., Bapst, V., Sanchez-Gonzalez, A., Zambaldi, V., Malinowski, M., Tacchetti, A., Raposo, D., Santoro, A., Faulkner, R., et al · 2018
Earlier work this paper cites.
JAX: composable transformations of Python+NumPy programs, 2018
Bradbury, J., Frostig, R., Hawkins, P., Johnson, M. J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., and Zhang, Q · 2018
Earlier work this paper cites.
Hardware conditioned policies for multi-robot transfer learning
Chen, T., Murali, A., and Gupta, A · 2018
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
Kalashnikov, D., Irpan, A., Pastor, P., Ibarz, J., Herzog, A., Jang, E., Quillen, D., Holly, E., Kalakrishnan, M., Vanhoucke, V., et al · 2018
Earlier work this paper cites.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
Levine, S., Pastor, P., Krizhevsky, A., Ibarz, J., and Quillen, D · 2018
Earlier work this paper cites.
Gpu-accelerated robotic simulation for distributed reinforcement learning
Liang, J., Makoviychuk, V., Handa, A., Chentanez, N., Macklin, M., and Fox, D · 2018
Earlier work this paper cites.
Graph networks as learnable physics engines for inference and control
Sanchez-Gonzalez, A., Heess, N., Springenberg, J. T., Merel, J., Riedmiller, M., Hadsell, R., and Battaglia, P · 2018
Earlier work this paper cites.
Can robots assemble an ikea chair?
Suárez-Ruiz, F., Zhou, X., and Pham, Q.-C · 2018
Cited alongside, same era.
Nervenet: Learning structured policy with graph neural networks
Wang, T., Liao, R., Ba, J., and Fidler, S · 2018
Cited alongside, same era.
Emergent tool use from multi-agent autocurricula
Baker, B., Kanitscheider, I., Markov, T., Wu, Y., Powell, G., McGrew, B., and Mordatch, I · 2019
Cited alongside, same era.
Structured agents for physical construction
Bapst, V., Sanchez-Gonzalez, A., Doersch, C., Stachenfeld, K., Kohli, P., Battaglia, P., and Hamrick, J · 2019
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Berner, C., Brockman, G., Chan, B., Cheung, V., Debiak, P., Dennison, C., Farhi, D., Fischer, Q., Hashme, S., Hesse, C., et al · 2019
Cited alongside, same era.
One policy to control them all: Shared modular policies for agent-agnostic control
Huang, W., Mordatch, I., and Pathak, D · 2020
Later among the works it cites.
My body is a cage: the role of morphology in graph-based incompatible control
Kurin, V., Igl, M., Rocktäschel, T., Boehmer, W., and Whiteson, S · 2020
Later among the works it cites.
Towards practical multi-object manipulation using relational reinforcement learning
Li, R., Jabri, A., Darrell, T., and Agrawal, P · 2020
Later among the works it cites.
Building lego using deep generative models of graphs
Thompson, R., Ghalebi, E., DeVries, T., and Taylor, G. W · 2020
Later among the works it cites.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Yu, T., Quillen, D., He, Z., Julian, R., Hausman, K., Finn, C., and Levine, S · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cabi, S., Colmenarejo, S. G., Novikov, A., Konyushkova, K., Reed, S., Jeong, R., Zolna, K., Aytar, Y., Budden, D., Vecerik, M., et al · 2019
Cited alongside, same era.
Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning
Gupta, A., Kumar, V., Lynch, C., Levine, S., and Hausman, K · 2019
Cited alongside, same era.
Shallow-depth insertion: Peg in shallow hole through robotic in-hand manipulation
Kim, C. H. and Seo, J · 2019
Cited alongside, same era.
Ikea furniture assembly environment for long-horizon complex manipulation tasks
Lee, Y., Hu, E. S., Yang, Z., Yin, A., and Lim, J. J · 2019
Cited alongside, same era.
Learning to control self-assembling morphologies: a study of generalization via modularity
Pathak, D., Lu, C., Darrell, T., Isola, P., and Efros, A. A · 2019
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Cited alongside, same era.
Wang, R., Lehman, J., Clune, J., and Stanley, K. O · 2019
Cited alongside, same era.
Transporter networks: Rearranging the visual world for robotic manipulation
Zeng, A., Florence, P., Tompson, J., Welker, S., Chien, J., Attarian, M., Armstrong, T., Krasin, I., Duong, D., Sindhwani, V., et al · 2020
Later among the works it cites.
A system for general in-hand object re-orientation
Chen, T., Xu, J., and Agrawal, P · 2021
Later among the works it cites.
Brick-by-brick: Combinatorial construction with deep reinforcement learning
Chung, H., Kim, J., Knyazev, B., Lee, J., Taylor, G. W., Park, J., and Cho, M · 2021
Later among the works it cites.
Brax - a differentiable physics engine for large scale rigid body simulation, 2021
Freeman, C. D., Frey, E., Raichuk, A., Girgin, S., Mordatch, I., and Bachem, O · 2021
Later among the works it cites.
Long-horizon multi-robot rearrangement planning for construction assembly
Hartmann, V. N., Orthey, A., Driess, D., Oguz, O. S., and Toussaint, M · 2021
Later among the works it cites.
Generalization in dexterous manipulation via geometry-aware multi-task learning
Huang, W., Mordatch, I., Abbeel, P., and Pathak, D · 2021
Later among the works it cites.
Adversarial skill chaining for long-horizon robot manipulation via terminal state regularization
Lee, Y., Lim, J. J., Anandkumar, A., and Zhu, Y · 2021
Later among the works it cites.
Isaac gym: High performance gpu-based physics simulation for robot learning
Makoviychuk, V., Wawrzyniak, L., Guo, Y., Lu, M., Storey, K., Macklin, M., Hoeller, D., Rudin, N., Allshire, A., Handa, A., et al · 2021
Later among the works it cites.
Tool as embodiment for recursive manipulation
Noguchi, Y., Matsushima, T., Matsuo, Y., and Gu, S. S · 2021
Later among the works it cites.
Asymmetric self-play for automatic goal discovery in robotic manipulation
OpenAI, O., Plappert, M., Sampedro, R., Xu, T., Akkaya, I., Kosaraju, V., Welinder, P., D’Sa, R., Petron, A., Pinto, H. P. d. O., et al · 2021
Later among the works it cites.
Learn2assemble with structured representations and search for robotic architectural construction
Funk, N., Chalvatzaki, G., Belousov, B., and Peters, J · 2022
Closest in time.
Cliport: What and where pathways for robotic manipulation
Shridhar, M., Manuelli, L., and Fox, D · 2022
Closest in time.