Fetching the paper…
Reading the bibliography…
This work presents an object-centric approach to learning vision-based manipulation skills from human videos.
Fast point feature histograms (fpfh) for 3d registration
R. B. Rusu, N. Blodow, and M. Beetz · 2009
Earlier work this paper cites.
Optimal detection of changepoints with a linear computational cost
R. Killick, P. Fearnhead, and I. A. Eckley · 2012
Earlier work this paper cites.
Robust reconstruction of indoor scenes
S. Choi, Q.-Y. Zhou, and V. Koltun · 2015
Earlier work this paper cites.
One-shot imitation learning
Y. Duan, M. Andrychowicz, B. Stadie, O. Jonathan Ho, J. Schneider, I. Sutskever, P. Abbeel, and W. Zaremba · 2017
Earlier work this paper cites.
Imitation from observation: Learning to imitate behaviors from raw video via context translation
Y. Liu, A. Gupta, P. Abbeel, and S. Levine · 2018
Earlier work this paper cites.
Behavioral cloning from observation
F. Torabi, G. Warnell, and P. Stone · 2018
Earlier work this paper cites.
Deep object pose estimation for semantic robotic grasping of household objects
J. Tremblay, T. To, B. Sundaralingam, Y. Xiang, D. Fox, and S. Birchfield · 2018
Earlier work this paper cites.
Deep object-centric representations for generalizable robot learning
C. Devin, P. Abbeel, T. Darrell, and S. Levine · 2018
Earlier work this paper cites.
Open3d: A modern library for 3d data processing
Q.-Y. Zhou, J. Park, and V. Koltun · 2018
Earlier work this paper cites.
Third-person visual imitation learning via decoupled hierarchical controller
P. Sharma, D. Pathak, and A. Gupta · 2019
Earlier work this paper cites.
Avid: Learning multi-stage tasks via pixel-level translation of human videos
L. Smith, N. Dhawan, M. Zhang, P. Abbeel, and S. Levine · 2019
Earlier work this paper cites.
Generative adversarial imitation from observation
F. Torabi, G. Warnell, and P. Stone · 2019
Earlier work this paper cites.
Deep object-centric policies for autonomous driving
D. Wang, C. Devin, Q.-Z. Cai, F. Yu, and T. Darrell · 2019
Earlier work this paper cites.
Structurenet: Hierarchical graph networks for 3d shape generation
K. Mo, P. Guerrero, L. Yi, H. Su, P. Wonka, N. Mitra, and L. J. Guibas · 2019
Earlier work this paper cites.
Learning to generalize across long-horizon tasks from human demonstrations
A. Mandlekar, D. Xu, R. Martín-Martín, S. Savarese, and L. Fei-Fei · 2020
Earlier work this paper cites.
Ridm: Reinforced inverse dynamics modeling for learning from a single observed demonstration
B. S. Pavse, F. Torabi, J. Hanna, G. Warnell, and P. Stone · 2020
Earlier work this paper cites.
Object-centric task and motion planning in dynamic environments
T. Migimatsu and J. Bohg · 2020
Earlier work this paper cites.
Selective review of offline change point detection methods
C. Truong, L. Oudre, and N. Vayatis · 2020
Earlier work this paper cites.
Accelerating robotic reinforcement learning via parameterized action primitives
M. Dalal, D. Pathak, and R. R. Salakhutdinov · 2021
Earlier work this paper cites.
Learning generalizable robotic reward functions from” in-the-wild” human videos
A. S. Chen, S. Nair, and C. Finn · 2021
Earlier work this paper cites.
Learning by watching: Physical imitation of manipulation skills from human videos
H. Xiong, Q. Li, Y.-C. Chen, H. Bharadhwaj, S. Sinha, and A. Garg · 2021
Earlier work this paper cites.
Imitation Learning from Observation
F. Torabi · 2021
Cited alongside, same era.
Towards open world object detection
K. Joseph, S. Khan, F. S. Khan, and V. N. Balasubramanian · 2021
Cited alongside, same era.
Coarse-to-fine imitation learning: Robot manipulation from a single demonstration
E. Johns · 2021
Cited alongside, same era.
Nerp: Neural rearrangement planning for unknown objects
A. H. Qureshi, A. Mousavian, C. Paxton, M. C. Yip, and D. Fox · 2021
Cited alongside, same era.
Hierarchical planning for long-horizon manipulation with geometric and symbolic scene graphs
Y. Zhu, J. Tremblay, S. Birchfield, and Y. Zhu · 2021
Cited alongside, same era.
Tangent space backpropagation for 3d transformation groups
Z. Teed and J. Deng · 2021
Learning to act from actionless videos through dense correspondences
P.-C. Ko, J. Mao, Y. Du, S.-H. Sun, and J. B. Tenenbaum · 2023
Later among the works it cites.
Any-point trajectory modeling for policy learning
C. Wen, X. Lin, J. So, K. Chen, Q. Dou, Y. Gao, and P. Abbeel · 2023
Later among the works it cites.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
S. Liu, Z. Zeng, T. Ren, F. Li, H. Zhang, J. Yang, C. Li, J. Yang, H. Su, J. Zhu, et al · 2023
Later among the works it cites.
Putting the object back into video object segmentation
H. K. Cheng, S. W. Oh, B. Price, J.-Y. Lee, and A. Schwing · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Augmenting reinforcement learning with behavior primitives for diverse manipulation tasks
S. Nasiriany, H. Liu, and Y. Zhu · 2022
Cited alongside, same era.
Viola: Imitation learning for vision-based manipulation with object proposal priors
Y. Zhu, A. Joshi, P. Stone, and Y. Zhu · 2022
Cited alongside, same era.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2022
Cited alongside, same era.
Vip: Towards universal visual reward and representation via value-implicit pre-training
Y. J. Ma, S. Sodhani, D. Jayaraman, O. Bastani, V. Kumar, and A. Zhang · 2022
Cited alongside, same era.
You only demonstrate once: Category-level manipulation from single visual demonstration
B. Wen, W. Lian, K. Bekris, and S. Schaal · 2022
Cited alongside, same era.
Human-to-robot imitation in the wild
S. Bahl, A. Gupta, and D. Pathak · 2022
Cited alongside, same era.
N. Karaev, I. Rocco, B. Graham, N. Neverova, A. Vedaldi, and C. Rupprecht · 2023
Later among the works it cites.
Robotap: Tracking arbitrary points for few-shot visual imitation
M. Vecerik, C. Doersch, Y. Yang, T. Davchev, Y. Aytar, G. Zhou, R. Hadsell, L. Agapito, and J. Scholz · 2023
Later among the works it cites.
Reconstructing hands in 3d with transformers
G. Pavlakos, D. Shan, I. Radosavovic, A. Kanazawa, D. Fouhey, and J. Malik · 2023
Later among the works it cites.
Zero-shot robot manipulation from passive human videos
H. Bharadhwaj, A. Gupta, S. Tulsiani, and V. Kumar · 2023
Later among the works it cites.
Graph inverse reinforcement learning from diverse videos
S. Kumar, J. Zamora, N. Hansen, R. Jangir, and X. Wang · 2023
Later among the works it cites.
Xskill: Cross embodiment skill discovery
M. Xu, Z. Xu, C. Chi, M. Veloso, and S. Song · 2023
Later among the works it cites.
Towards generalizable zero-shot manipulation via translating human interaction plans
H. Bharadhwaj, A. Gupta, V. Kumar, and S. Tulsiani · 2023
Later among the works it cites.
Open-world object manipulation using pre-trained vision-language models
A. Stone, T. Xiao, Y. Lu, K. Gopalakrishnan, K.-H. Lee, Q. Vuong, P. Wohlhart, B. Zitkovich, F. Xia, C. Finn, et al · 2023
Later among the works it cites.
Planning for multi-object manipulation with graph neural network relational classifiers
Y. Huang, A. Conkey, and T. Hermans · 2023
Later among the works it cites.
Google Labs blog, Dec. 2024
State-of-the-art video and image generation with veo 2 and imagen 3 · 2024
Closest in time.
Foundationpose: Unified 6d pose estimation and tracking of novel objects
B. Wen, W. Yang, J. Kautz, and S. Birchfield · 2024
Closest in time.
J. Xu, W. Cheng, Y. Gao, X. Wang, S. Gao, and Y. Shan · 2024
Closest in time.
Ditto: Demonstration imitation by trajectory transformation
N. Heppert, M. Argus, T. Welschehold, T. Brox, and A. Valada · 2024
Closest in time.
Okami: Teaching humanoid robots manipulation skills through single video imitation, 2024
J. Li, Y. Zhu, Y. Xie, Z. Jiang, M. Seo, G. Pavlakos, and Y. Zhu · 2024
Closest in time.
Robotic manipulation by imitating generated videos without physical demonstrations
S. Patel, S. Mohan, H. Mai, U. Jain, S. Lazebnik, and Y. Li · 2025
Closest in time.
Demodiffusion: One-shot human imitation using pre-trained diffusion policy
S. Park, H. Bharadhwaj, and S. Tulsiani · 2025
Closest in time.