Fetching the paper…
Reading the bibliography…
Learning diverse policies for non-prehensile manipulation is essential for improving skill transfer and generalization to out-of-distribution scenarios.
J. Peters and S. Schaal, “Reinforcement learning by reward-weighted regression for operational space control,” in Proceedings of the 24th international conference on Machine learning , 2007, pp. 745–750
2007
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ international conference on intelligent robots and systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
K.-T. Yu, M. Bauza, N. Fazeli, and A. Rodriguez, “More than a million ways to be pushed. a high-fidelity experimental dataset of planar pushing,” in 2016 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 2016, pp. 30–37
2016
Earlier work this paper cites.
J. Merel, L. Hasenclever, A. Galashov, A. Ahuja, V. Pham, G. Wayne, Y. W. Teh, and N. Heess, “Neural probabilistic motor primitives for humanoid control,” in International Conference on Learning Representations , 2018
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning . PMLR, 2018, pp. 1861–1870
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Earlier work this paper cites.
S. Fujimoto, H. Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in International conference on machine learning . PMLR, 2018, pp. 1587–1596
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Y. Hou and M. T. Mason, “Robust execution of contact-rich motion plans by hybrid force-velocity control,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019
2019
Earlier work this paper cites.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” Advances in neural information processing systems , vol. 32, 2019
2019
Earlier work this paper cites.
S. Kumar, A. Kumar, S. Levine, and C. Finn, “One solution is not all you need: Few-shot extrapolation via structured maxent rl,” Advances in Neural Information Processing Systems , 2020
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in neural information processing systems , 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
K. Mo, L. J. Guibas, M. Mukadam, A. Gupta, and S. Tulsiani, “Where2act: From pixels to actions for articulated 3d objects,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 6813–6823
2021
Earlier work this paper cites.
S. Fujimoto and S. S. Gu, “A minimalist approach to offline reinforcement learning,” Advances in neural information processing systems , vol. 34, pp. 20 132–20 145, 2021
2021
Earlier work this paper cites.
R. Agarwal, M. Schwarzer, P. S. Castro, A. C. Courville, and M. Bellemare, “Deep reinforcement learning at the edge of the statistical precipice,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Cited alongside, same era.
Z. Xu, Z. He, and S. Song, “Universal manipulation policy network for articulated objects,” IEEE robotics and automation letters , vol. 7, no. 2, pp. 2447–2454, 2022
2022
Cited alongside, same era.
X. Cheng, E. Huang, Y. Hou, and M. T. Mason, “Contact mode guided motion planning for quasidynamic dexterous manipulation in 3d,” in 2022 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 2730–2736
2022
Cited alongside, same era.
Z. Feldman, H. Ziesche, N. A. Vien, and D. D. Castro, “A hybrid approach for learning to shift and grasp with elaborate motion primitives,” in 2022 International Conference on Robotics and Automation, ICRA . IEEE, 2022, pp. 6365–6371
2022
Cited alongside, same era.
2023
Later among the works it cites.
W. Zhou and D. Held, “Learning to grasp the ungraspable with emergent extrinsic dexterity,” in Conference on Robot Learning . PMLR, 2023, pp. 150–160
2023
Later among the works it cites.
2023
Later among the works it cites.
G. Li, Z. Jin, M. Volpp, F. Otto, R. Lioutikov, and G. Neumann, “Prodmp: A unified perspective on dynamic and probabilistic movement primitives,” IEEE Robotics and Automation Letters , 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Chen, C. Lu, C. Ying, H. Su, and J. Zhu, “Offline reinforcement learning via high-fidelity generative behavior modeling,” in The Eleventh International Conference on Learning Representations , 2022
2022
Cited alongside, same era.
W. Zhou, B. Jiang, F. Yang, C. Paxton, and D. Held, “HACMan: Learning hybrid actor-critic maps for 6d non-prehensile manipulation,” in Conference on Robot Learning (CoRL) , vol. 229. PMLR, 2023
2023
Cited alongside, same era.
Z. Ding and C. Jin, “Consistency models as a rich and efficient policy class for reinforcement learning,” in The Twelfth International Conference on Learning Representations , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Z. Wang, J. J. Hunt, and M. Zhou, “Diffusion policies as an expressive policy class for offline reinforcement learning,” in The Eleventh International Conference on Learning Representations , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
C. Lu, H. Chen, J. Chen, H. Su, C. Li, and J. Zhu, “Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning,” in International Conference on Machine Learning . PMLR, 2023, pp. 22 825–22 855
2023
Cited alongside, same era.
M. Reuss, M. Li, X. Jia, and R. Lioutikov, “Goal conditioned imitation learning using score-based diffusion policies,” in Robotics: Science and Systems , 2023
2023
Cited alongside, same era.
Y. Song, P. Dhariwal, M. Chen, and I. Sutskever, “Consistency models,” in International Conference on Machine Learning . PMLR, 2023, pp. 32 211–32 252
2023
Later among the works it cites.
X. Jia, D. Blessing, X. Jiang, M. Reuss, A. Donat, R. Lioutikov, and G. Neumann, “Towards diverse behaviors: A benchmark for imitation learning with human demonstrations,” in The Twelfth International Conference on Learning Representations , 2024
2024
Closest in time.
S. Venkatraman, S. Khaitan, R. T. Akella, J. Dolan, J. Schneider, and G. Berseth, “Reasoning with latent diffusion in offline reinforcement learning,” in The Twelfth International Conference on Learning Representations , 2024
2024
Closest in time.
Z. Li, R. Krohn, T. Chen, A. Ajay, P. Agrawal, and G. Chalvatzaki, “Learning multimodal behaviors from scratch with diffusion policy gradient,” in The Thirty-eighth Annual Conference on Neural Information Processing Systems , 2024
2024
Closest in time.
K. Black, M. Janner, Y. Du, I. Kostrikov, and S. Levine, “Training diffusion models with reinforcement learning,” in The Twelfth International Conference on Learning Representations , 2024
2024
Closest in time.
M. Uehara, Y. Zhao, K. Black, E. Hajiramezanali, G. Scalia, N. L. Diamant, A. M. Tseng, S. Levine, and T. Biancalani, “Feedback efficient online fine-tuning of diffusion models,” in International Conference on Machine Learning (ICML) , 2024
2024
Closest in time.
2024
Closest in time.
B. Jiang, Y. Wu, W. Zhou, C. Paxton, and D. Held, “HACMan++: Spatially-Grounded Motion Primitives for Manipulation,” in Proceedings of Robotics: Science and Systems , Delft, Netherlands, July 2024
2024
Closest in time.
B. Kang, X. Ma, C. Du, T. Pang, and S. Yan, “Efficient diffusion policies for offline reinforcement learning,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Cho, J. Han, Y. Cho, and B. Kim, “CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects,” in International Conference on Learning Representations (ICLR) , 2024
2024
Closest in time.