Fetching the paper…
Reading the bibliography…
Inferring object motion representations from observations enhances the performance of robotic manipulation tasks.
“ALVINN: An Autonomous Land Vehicle in a Neural Network”
Dean Pomerleau · 1988
Earlier work this paper cites.
“Structure and function of visual area MT”
Richard Born and David Bradley · 2005
Earlier work this paper cites.
“Mujoco: A physics engine for model-based control”
Emanuel Todorov, Tom Erez and Yuval Tassa · 2012
Earlier work this paper cites.
“Learning complex dexterous manipulation with deep reinforcement learning and demonstrations”
Aravind Rajeswaran et al · 2017
Earlier work this paper cites.
“Proximal policy optimization algorithms”
John Schulman et al · 2017
Earlier work this paper cites.
“Deep object-centric representations for generalizable robot learning”
Coline Devin, Pieter Abbeel, Trevor Darrell and Sergey Levine · 2018
Earlier work this paper cites.
“4d spatio-temporal convnets: Minkowski convolutional neural networks”
Christopher Choy, JunYoung Gwak and Silvio Savarese · 2019
Earlier work this paper cites.
“Reasoning About Physical Interactions with Object-Oriented Prediction and Planning”
Michael Janner et al · 2019
Earlier work this paper cites.
“On the continuity of rotation representations in neural networks”
Yi Zhou et al · 2019
Earlier work this paper cites.
“Denoising diffusion probabilistic models”
J. Ho, A. Jain and P. Abbeel · 2020
Earlier work this paper cites.
“Object-centric task and motion planning in dynamic environments”
Toki Migimatsu and Jeannette Bohg · 2020
Earlier work this paper cites.
“Raft: Recurrent all-pairs field transforms for optical flow”
Zachary Teed and Jia Deng · 2020
Earlier work this paper cites.
“Sapien: A simulated part-based interactive environment”
Fanbo Xiang et al · 2020
Earlier work this paper cites.
“Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning”
Tianhe Yu et al · 2020
Earlier work this paper cites.
“BC-Z: Zero-Shot Task Generalization with Robotic Imitation Learning”
E. Jang and et al · 2021
Earlier work this paper cites.
“What Matters in Learning from Offline Human Demonstrations for Robot Manipulation”
A. Mandlekar and et al · 2021
Earlier work this paper cites.
“Denoising diffusion implicit models”
J. Song, C. Meng and S. Ermon · 2021
Earlier work this paper cites.
“An End-to-End Differentiable Framework for Contact-Aware Robot Design”
Jie Xu et al · 2021
Earlier work this paper cites.
“Affordance Learning from Play for Sample-Efficient Policy Learning”
J. Borja-Diaz et al · 2022
Earlier work this paper cites.
“Flowbot3d: Learning 3d articulation flow to manipulate articulated objects”
B. Eisner, H. Zhang and D. Held · 2022
Earlier work this paper cites.
“Ifor: Iterative flow minimization for robotic object rearrangement”
Ankit Goyal et al · 2022
Earlier work this paper cites.
“Behavior Transformers: Cloning k k Modes with One Stone”
Nur Shafiullah, Zichen Cui, Ariuntuya Altanzaya and Lerrel Pinto · 2022
Cited alongside, same era.
“6-DoF pose estimation of household objects for robotic manipulation: An accessible dataset and benchmark”
Stephen Tyree et al · 2022
Cited alongside, same era.
“Vrl3: A data-driven framework for visual deep reinforcement learning”
Che Wang, Xufang Luo, Keith Ross and Dongsheng Li · 2022
Cited alongside, same era.
“Viola: Object-centric imitation learning for vision-based robot manipulation”
Yifeng Zhu, Abhishek Joshi, Peter Stone and Yuke Zhu · 2022
Cited alongside, same era.
“Affordances from human videos as a versatile representation for robotics”
Shikhar Bahl et al · 2023
Cited alongside, same era.
“Dexart: Benchmarking generalizable dexterous manipulation with articulated objects”
“SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation”
Cheng-Chun Hsu et al · 2024
Closest in time.
“EgoMimic: Scaling Imitation Learning via Egocentric Video”
Simar Kareer et al · 2024
Closest in time.
“OpenVLA: An Open-Source Vision-Language-Action Model”
Moo Kim et al · 2024
Closest in time.
“Behavior Generation with Latent Actions”
Seungjae Lee et al · 2024
Closest in time.
“FoAM: Foresight-Augmented Multi-Task Imitation Policy for Robotic Manipulation”
Litao Liu et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chen Bao, Helin Xu, Yuzhe Qin and Xiaolong Wang · 2023
Cited alongside, same era.
“RT-1: Robotics Transformer for Real-World Control at Scale”
Anthony Brohan, Noah Brown, Justice Carbajal and … · 2023
Cited alongside, same era.
“Diffusion Policy: Visuomotor Policy Learning via Action Diffusion”
C. Chi and et al · 2023
Cited alongside, same era.
“From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data”
Zichen Cui, Yibin Wang, Nur Shafiullah and Lerrel Pinto · 2023
Cited alongside, same era.
“Toolflownet: Robotic manipulation with tools via predicting tool flow from point clouds”
D. Seita et al · 2023
Cited alongside, same era.
“Any-point trajectory modeling for policy learning”
C. Wen et al · 2023
Cited alongside, same era.
“Flowbot++: Learning generalized articulated objects manipulation via articulation projection”
H. Zhang, B. Eisner and D. Held · 2023
Cited alongside, same era.
Soroush Nasiriany et al · 2024
Closest in time.
“Octo: An Open-Source Generalist Robot Policy”
Octo Model Team et al · 2024
Closest in time.
“HRP: Human Affordances for Robotic Pre-Training”
Mohan Srirama, Sudeep Dasari, Shikhar Bahl and Abhinav Gupta · 2024
Closest in time.
“Robotap: Tracking arbitrary points for few-shot visual imitation”
M. Vecerik et al · 2024
Closest in time.
“RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective”
Chenxi Wang, Hongjie Fang, Hao-Shu Fang and Cewu Lu · 2024
Closest in time.
“Articulated Object Manipulation using Online Axis Estimation with SAM2-Based Tracking”
X. Wang and et al · 2024
Closest in time.
“Foundationpose: Unified 6d pose estimation and tracking of novel objects”
Bowen Wen, Wei Yang, Jan Kautz and Stan Birchfield · 2024
Closest in time.
“CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation”
Shangning Xia, Hongjie Fang, Hao-Shu Fang and Cewu Lu · 2024
Closest in time.
“Flow as the Cross-domain Manipulation Interface”
M. Xu et al · 2024
Closest in time.
“General flow as foundation affordance for scalable robot learning”
C. Yuan, C. Wen, T. Zhang and Y. Gao · 2024
Closest in time.
“RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics”
Wentao Yuan et al · 2024
Closest in time.
“3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations”
Y. Ze and et al · 2024
Closest in time.
“Leveraging Locality to Boost Sample Efficiency in Robotic Manipulation”
Tong Zhang, Yingdong Hu, Jiacheng You and Yang Gao · 2024
Closest in time.
“Vision-based Manipulation from Single Human Video with Open-World Object Graphs”
Yifeng Zhu, Arisrei Lim, Peter Stone and Yuke Zhu · 2024
Closest in time.
“Dense Policy: Bidirectional Autoregressive Learning of Actions”
Yue Su et al · 2025
Closest in time.