Fetching the paper…
Reading the bibliography…
We explore how intermediate policy representations can facilitate generalization by providing guidance on how to perform manipulation tasks.
“Affordances in Robotic Tasks – A Survey”, 2020
Paola Ardon et al · 2004
Earlier work this paper cites.
“One-shot visual imitation learning via meta-learning”
Chelsea Finn et al · 2017
Earlier work this paper cites.
“Dex-Net 2.0: Deep Learning to Plan Robust Grasps with Synthetic Point Clouds and Analytic Grasp Metrics”
Jeffrey Mahler et al · 2017
Earlier work this paper cites.
“Scaling Egocentric Vision: The EPIC-KITCHENS Dataset”
Dima Damen et al · 2018
Earlier work this paper cites.
“Learning an embedding space for transferable robot skills”
Karol Hausman et al · 2018
Earlier work this paper cites.
“Task-embedded control networks for few-shot imitation learning”
Stephen James, Michael Bloesch and Andrew Davison · 2018
Earlier work this paper cites.
“6-DOF GraspNet: Variational Grasp Generation for Object Manipulation”
Arsalan Mousavian, Clemens Eppner and Dieter Fox · 2019
Earlier work this paper cites.
“Planning with goal-conditioned policies”
Soroush Nasiriany, Vitchyr Pong, Steven Lin and Sergey Levine · 2019
Earlier work this paper cites.
“GraspNet-1Billion: A Large-Scale Benchmark for General Object Grasping”
Hao-Shu Fang, Chenxi Wang, Minghao Gou and Cewu Lu · 2020
Earlier work this paper cites.
“Language-conditioned imitation learning for robot manipulation tasks”
Simon Stepputtis et al · 2020
Earlier work this paper cites.
“Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills”
Yevgen Chebotar et al · 2021
Earlier work this paper cites.
“BC-Z: Zero-Shot Task Generalization with Robotic Imitation Learning”
Eric Jang et al · 2021
Earlier work this paper cites.
“Accelerating reinforcement learning with learned skill priors”
Karl Pertsch, Youngwoon Lee and Joseph Lim · 2021
Earlier work this paper cites.
“Contact-GraspNet: Efficient 6-DoF Grasp Generation in Cluttered Scenes”
Martin Sundermeyer, Arsalan Mousavian, Rudolph Triebel and Dieter Fox · 2021
Earlier work this paper cites.
“RT-1: Robotics Transformer for Real-World Control at Scale”
Anthony Brohan et al · 2022
Earlier work this paper cites.
“Can Foundation Models Perform Zero-Shot Task Specification For Robot Manipulation?”
Yuchen Cui et al · 2022
Earlier work this paper cites.
“Ego4D: Around the World in 3,000 Hours of Egocentric Video”
Kristen Grauman et al · 2022
Earlier work this paper cites.
“Learning language-conditioned robot behavior from offline data and crowd-sourced annotation”
Suraj Nair et al · 2022
Cited alongside, same era.
“PaLM 2 Technical Report”, 2023
Rohan Anil et al · 2023
Cited alongside, same era.
“Affordances from Human Videos as a Versatile Representation for Robotics”
Shikhar Bahl et al · 2023
Cited alongside, same era.
“Robocat: A self-improving generalist agent for robotic manipulation”
Konstantinos Bousmalis et al · 2023
Cited alongside, same era.
“RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control”
Anthony Brohan et al · 2023
Cited alongside, same era.
“PaLM-E: An Embodied Multimodal Language Model”
Danny Driess et al · 2023
“BridgeData V2: A Dataset for Robot Learning at Scale”
Homer Walke et al · 2023
Later among the works it cites.
“RT-H: Action Hierarchies Using Language”
Suneel Belkhale et al · 2024
Closest in time.
Homanga Bharadhwaj, Roozbeh Mottaghi, Abhinav Gupta and Shubham Tulsiani · 2024
Closest in time.
“Manipulate-Anything: Automating Real-World Robots using Vision-Language Models”
Jiafei Duan et al · 2024
Closest in time.
“MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting”
Kuan Fang, Fangchen Liu, Pieter Abbeel and Sergey Levine · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“AnyGrasp: Robust and Efficient Grasp Perception in Spatial and Temporal Domains”
Hao-Shu Fang et al · 2023
Cited alongside, same era.
“RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches”, 2023
Jiayuan Gu et al · 2023
Cited alongside, same era.
“VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models”
Wenlong Huang et al · 2023
Cited alongside, same era.
“VIMA: General Robot Manipulation with Multimodal Prompts”
Yunfan Jiang et al · 2023
Cited alongside, same era.
“Mt-opt: Continuous multi-task robotic reinforcement learning at scale”
Dmitry Kalashnikov et al · 2023
Cited alongside, same era.
“Robot Learning on the Job: Human-in-the-Loop Autonomy and Learning During Deployment”
Huihan Liu et al · 2023
Cited alongside, same era.
Wenlong Huang et al · 2024
Closest in time.
“Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers”, 2024
Vidhi Jain et al · 2024
Closest in time.
“DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset”, 2024
Alexander Khazatsky et al · 2024
Closest in time.
“OpenVLA: An Open-Source Vision-Language-Action Model”
Moo Jin Kim et al · 2024
Closest in time.
“HRP: Human Affordances for Robotic Pre-Training”
Mohan Srirama, Sudeep Dasari, Shikhar Bahl and Abhinav Gupta · 2024
Closest in time.
“RT-Sketch: Goal-Conditioned Imitation Learning from Hand-Drawn Sketches”, 2024
Priya Sundaresan et al · 2024
Closest in time.
“Gemini: A Family of Highly Capable Multimodal Models”, 2024
Gemini Team · 2024
Closest in time.
“Any-point Trajectory Modeling for Policy Learning”
Chuan Wen et al · 2024
Closest in time.
“RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics”, 2024
Wentao Yuan et al · 2024
Closest in time.
“Sprint: Scalable policy pre-training via language instruction relabeling”
Jesse Zhang, Karl Pertsch, Jiahui Zhang and Joseph Lim · 2024
Closest in time.
“Transformers for one-shot visual imitation”
Sudeep Dasari and Abhinav Gupta · 2084
Closest in time.