Fetching the paper…
Reading the bibliography…
We introduce SPOT, an object-centric imitation learning framework.
“Receding horizon control of nonlinear systems”
David Mayne and Hannah Michalska · 1988
Earlier work this paper cites.
“V-REP: A versatile and scalable robot simulation framework”
Eric Rohmer, Surya Singh and Marc Freese · 2013
Earlier work this paper cites.
JL Ba · 2016
Earlier work this paper cites.
“Deep object-centric representations for generalizable robot learning”
Coline Devin, Pieter Abbeel, Trevor Darrell and Sergey Levine · 2018
Earlier work this paper cites.
“Learning plannable representations with causal InfoGAN”
Thanard Kurutach et al · 2018
Earlier work this paper cites.
“KPAM: Keypoint affordances for category-level robotic manipulation”
Lucas Manuelli, Wei Gao, Peter Florence and Russ Tedrake · 2019
Earlier work this paper cites.
“Third-person visual imitation learning via decoupled hierarchical controller”
Pratyusha Sharma, Deepak Pathak and Abhinav Gupta · 2019
Earlier work this paper cites.
“Deep object-centric policies for autonomous driving”
Dequan Wang et al · 2019
Earlier work this paper cites.
“Denoising diffusion probabilistic models”
Jonathan Ho, Ajay Jain and Pieter Abbeel · 2020
Earlier work this paper cites.
“RLBench: The robot learning benchmark & learning environment”
Stephen James, Zicong Ma, David Arrojo and Andrew Davison · 2020
Earlier work this paper cites.
“Object-centric task and motion planning in dynamic environments”
Toki Migimatsu and Jeannette Bohg · 2020
Earlier work this paper cites.
“Graph-structured visual imitation”
Maximilian Sieb et al · 2020
Earlier work this paper cites.
“KPAM 2.0: Feedback control for category-level robotic manipulation”
Wei Gao and Russ Tedrake · 2021
Earlier work this paper cites.
“Vision-driven compliant manipulation for reliable, high-precision assembly tasks”
Andrew Morgan et al · 2021
Earlier work this paper cites.
“Learning transferable visual models from natural language supervision”
Alec Radford et al · 2021
Earlier work this paper cites.
“Denoising diffusion implicit models”
Jiaming Song, Chenlin Meng and Stefano Ermon · 2021
Earlier work this paper cites.
“Do as I can, not as I say: Grounding language in robotic affordances”
Michael Ahn et al · 2022
Earlier work this paper cites.
“Human-to-robot imitation in the wild”
Shikhar Bahl, Abhinav Gupta and Deepak Pathak · 2022
Earlier work this paper cites.
“Video pretraining (VPT): Learning to act by watching unlabeled online videos”
Bowen Baker et al · 2022
Earlier work this paper cites.
“BC-Z: Zero-shot task generalization with robotic imitation learning”
Eric Jang et al · 2022
Earlier work this paper cites.
“Adversarial imitation learning from video using a state observer”
Haresh Karnan, Faraz Torabi, Garrett Warnell and Peter Stone · 2022
Earlier work this paper cites.
“Complex in-hand manipulation via compliance-enabled finger gaiting and multi-modal planning”
Andrew Morgan et al · 2022
Earlier work this paper cites.
“R3M: A universal visual representation for robot manipulation”
Suraj Nair et al · 2022
Cited alongside, same era.
“The surprising effectiveness of representation learning for visual imitation”
Jyothish Pari, Nur Shafiullah, Sridhar Arunachalam and Lerrel Pinto · 2022
Cited alongside, same era.
“A generalist agent”
Scott Reed et al · 2022
Cited alongside, same era.
“Neural descriptor fields: SE(3)-equivariant object representations for manipulation”
Anthony Simeonov et al · 2022
Cited alongside, same era.
“Robotic telekinesis: Learning a robotic hand imitator by watching humans on YouTube”
Aravind Sivakumar, Kenneth Shaw and Deepak Pathak · 2022
Cited alongside, same era.
“Demonstrate once, imitate immediately (DOME): Learning visual servoing for one-shot imitation learning”
Eugene Valassakis, Georgios Papagiannis, Norman Di and Edward Johns · 2022
“YOLO-World: Real-time open-vocabulary object detection”
Tianheng Cheng et al · 2024
Closest in time.
“DINOBot: Robot manipulation via retrieval and alignment with vision foundation models”
Norman Di and Edward Johns · 2024
Closest in time.
“Keypoint Action tokens enable in-context imitation learning in robotics”
Norman Di and Edward Johns · 2024
Closest in time.
“Learning universal policies via text-guided video generation”
Yilun Du et al · 2024
Closest in time.
“Mobile ALOHA: Learning bimanual mobile manipulation with low-cost whole-body teleoperation”
Zipeng Fu, Tony Zhao and Chelsea Finn · 2024
Closest in time.
“RVT-2: Learning Precise Manipulation from Few Demonstrations”
Ankit Goyal et al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“You only demonstrate once: Category-level manipulation from single visual demonstration”
Bowen Wen, Wenzhao Lian, Kostas Bekris and Stefan Schaal · 2022
Cited alongside, same era.
“Diffusion policy: Visuomotor policy learning via action diffusion”
Cheng Chi et al · 2023
Cited alongside, same era.
“RVT: Robotic view transformer for 3d object manipulation”
Ankit Goyal et al · 2023
Cited alongside, same era.
“Equivariant descriptor fields: Se (3)-equivariant energy-based models for end-to-end visual robotic manipulation learning”
Hyunwoo Ryu, Hong-in Lee, Jeong-Hoon Lee and Jongeun Choi · 2023
Cited alongside, same era.
“ToolFlowNet: Robotic manipulation with tools via predicting tool flow from point clouds”
Daniel Seita et al · 2023
Cited alongside, same era.
“Perceiver-actor: A multi-task transformer for robotic manipulation”
Mohit Shridhar, Lucas Manuelli and Dieter Fox · 2023
Cited alongside, same era.
“ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation”
Wenlong Huang et al · 2024
Closest in time.
“Learning any-view 6DoF robotic grasping in cluttered scenes via neural surface rendering”
Snehal Jauhri, Ishikaa Lunawat and Georgia Chalvatzaki · 2024
Closest in time.
“CoTracker: It is better to track together”
Nikita Karaev et al · 2024
Closest in time.
“3D diffuser actor: Policy diffusion with 3D scene representations”
Tsung-Wei Ke, Nikolaos Gkanatsios and Katerina Fragkiadaki · 2024
Closest in time.
“Learning to act from actionless videos through dense correspondences”
Po-Chen Ko et al · 2024
Closest in time.
“FlowRetrieval: Flow-Guided Data Retrieval for Few-Shot Imitation Learning”
Li-Heng Lin et al · 2024
Closest in time.
“SAM-6D: Segment anything model meets zero-shot 6D object pose estimation”
Jiehong Lin, Lihua Liu, Dekun Lu and Kui Jia · 2024
Closest in time.
“RoboTAP: Tracking arbitrary points for few-shot visual imitation”
Mel Vecerik et al · 2024
Closest in time.
“FoundationPose: Unified 6D pose estimation and tracking of novel objects”
Bowen Wen, Wei Yang, Jan Kautz and Stan Birchfield · 2024
Closest in time.
“Any-point trajectory modeling for policy learning”
Chuan Wen et al · 2024
Closest in time.
“Flow as the Cross-Domain Manipulation Interface”
Mengda Xu et al · 2024
Closest in time.
“DNAct: Diffusion guided multi-task 3d policy learning”
Ge Yan, Yueh-Hua Wu and Xiaolong Wang · 2024
Closest in time.
“General flow as foundation affordance for scalable robot learning”
Chengbo Yuan, Chuan Wen, Tong Zhang and Yang Gao · 2024
Closest in time.
“3D diffusion policy”
Yanjie Ze et al · 2024
Closest in time.
“Vision-based Manipulation from Single Human Video with Open-World Object Graphs”
Yifeng Zhu, Arisrei Lim, Peter Stone and Yuke Zhu · 2024
Closest in time.