Fetching the paper…
Reading the bibliography…
There have recently been large advances both in pre-training visual representations for robotic control and segmenting unknown category objects in general images.
“Language models are few-shot learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“The Hungarian method for the assignment problem”
Harold. Kuhn · 1955
Earlier work this paper cites.
“Objective Criteria for the Evaluation of Clustering Methods”
W.. Rand · 1971
Earlier work this paper cites.
“Comparing partitions”
Lawrence. Hubert and Phipps Arabie · 1985
Earlier work this paper cites.
“Separate visual pathways for perception and action”
Melvyn Goodale and A Milner · 1992
Earlier work this paper cites.
“‘What’and ‘where’in the human brain”
Leslie Ungerleider and James Haxby · 1994
Earlier work this paper cites.
“Distinctive image features from scale-invariant keypoints”
David Lowe · 2004
Earlier work this paper cites.
“Histograms of oriented gradients for human detection”
Navneet Dalal and Bill Triggs · 2005
Earlier work this paper cites.
“Reconstruction Bottlenecks in Object-Centric Generative Models”, 2020
Martin Engelcke, Oiwi Jones and Ingmar Posner · 2007
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng et al · 2009
Earlier work this paper cites.
“On the usefulness of ‘what’and ‘where’pathways in vision”
Edward de Haan and Alan Cowey · 2011
Earlier work this paper cites.
“On the Binding Problem in Artificial Neural Networks”, 2020
Klaus Greff, Sjoerd van Steenkiste and J“”urgen Schmidhuber · 2012
Earlier work this paper cites.
“Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks”
Shaoqing Ren, Kaiming He, Ross. Girshick and Jian Sun · 2015
Earlier work this paper cites.
“Fast r-cnn”
Ross Girshick · 2015
Earlier work this paper cites.
“Building Machines That Learn and Think Like People”
Brenden Lake, Tomer Ullman, Joshua Tenenbaum and Samuel Gershman · 2016
Earlier work this paper cites.
“Semi-supervised classification with graph convolutional networks”
Thomas Kipf and Max Welling · 2016
Earlier work this paper cites.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Earlier work this paper cites.
“On implementing 2D rectangular assignment algorithms”
David. Crouse · 2016
Earlier work this paper cites.
“Deep Object-Centric Representations for Generalizable Robot Learning”
Coline Devin, P. Abbeel, Trevor Darrell and Sergey Levine · 2017
Earlier work this paper cites.
“Deep sets”
Manzil Zaheer et al · 2017
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
Petar Velickovi“’c et al · 2017
Earlier work this paper cites.
“Mask R-CNN”
Kaiming He, Georgia Gkioxari, Piotr Doll“’ar and Ross. Girshick · 2017
Earlier work this paper cites.
“A-fast-rcnn: Hard positive generation via adversary for object detection”
Xiaolong Wang, Abhinav Shrivastava and Abhinav Gupta · 2017
Earlier work this paper cites.
“Bert: Pre-training of deep bidirectional transformers for language understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Earlier work this paper cites.
“Deep Object-Centric Policies for Autonomous Driving”
Dequan Wang et al · 2018
Earlier work this paper cites.
“Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects”
Jonathan Tremblay et al · 2018
Earlier work this paper cites.
“MONet: Unsupervised Scene Decomposition and Representation”
Christopher. Burgess et al · 2019
Cited alongside, same era.
“Contrastive Learning of Structured World Models”
Thomas Kipf, Elise van Pol and Max Welling · 2019
Cited alongside, same era.
“Graph-Structured Visual Imitation”
Maximilian Sieb et al · 2019
Cited alongside, same era.
“Unsupervised Learning of Object Keypoints for Perception and Control”
Tejas. Kulkarni et al · 2019
Cited alongside, same era.
“Unsupervised Learning of Object Structure and Dynamics from Videos”
Matthias Minderer et al · 2019
Cited alongside, same era.
“VIOLA: Imitation Learning for Vision-Based Manipulation with Object Proposal Priors”
Yifeng Zhu, Abhishek Joshi, Peter Stone and Yuke Zhu · 2022
Later among the works it cites.
“R3m: A universal visual representation for robot manipulation”
Suraj Nair et al · 2022
Later among the works it cites.
“Visuomotor control in multi-object scenes using object-aware representations”
Negin Heravi et al · 2022
Later among the works it cites.
“Generalization and Robustness Implications in Object-Centric Learning”
Andrea Dittadi et al · 2022
Later among the works it cites.
“Robotic Skill Acquisition via Instruction Augmentation with Vision-Language Models”
Ted Xiao et al · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“A perspective on objects and systematic generalization in model-based RL”
Sjoerd van Steenkiste, Klaus Greff and J“”urgen Schmidhuber · 2019
Cited alongside, same era.
“Object-centric Forward Modeling for Model Predictive Control”
Yufei Ye, Dhiraj Gandhi, Abhinav Gupta and Shubham Tulsiani · 2019
Cited alongside, same era.
“Object-Centric Task and Motion Planning in Dynamic Environments”
Toki Migimatsu and Jeannette Bohg · 2019
Cited alongside, same era.
“Object-Centric Learning with Slot Attention”
Francesco Locatello et al · 2020
Cited alongside, same era.
“SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and Decomposition”
Zhixuan Lin et al · 2020
Cited alongside, same era.
“Rlbench: The robot learning benchmark & learning environment”
Stephen James, Zicong Ma, David Arrojo and Andrew Davison · 2020
Cited alongside, same era.
“Learning transferable visual models from natural language supervision”
Alec Radford et al · 2021
Cited alongside, same era.
Later among the works it cites.
“Inductive biases for object-centric representations in the presence of complex textures”
Samuele Papa, Ole Winther and Andrea Dittadi · 2022
Later among the works it cites.
“Promising or Elusive? Unsupervised Object Segmentation from Real-world Single Images”
Yafei Yang and Bo Yang · 2022
Later among the works it cites.
“Policy architectures for compositional generalization in control”
Allan Zhou, Vikash Kumar, Chelsea Finn and Aravind Rajeswaran · 2022
Later among the works it cites.
“Vima: General robot manipulation with multimodal prompts”
Yunfan Jiang et al · 2022
Later among the works it cites.
“Learning neuro-symbolic skills for bilevel planning”
Tom Silver et al · 2022
Later among the works it cites.
“Q-attention: Enabling efficient learning for vision-based robotic manipulation”
Stephen James and Andrew Davison · 2022
Later among the works it cites.
“Human-to-Robot Imitation in the Wild”
Shikhar Bahl, Abhinav Gupta and Deepak Pathak · 2022
Later among the works it cites.
“Imagebind: One embedding space to bind them all”
Rohit Girdhar et al · 2023
Later among the works it cites.
“LIV: Language-Image Representations and Rewards for Robotic Control”
Yecheng Ma et al · 2023
Later among the works it cites.
Alexander Kirillov et al · 2023
Later among the works it cites.
“Segment Everything Everywhere All at Once”
Xueyan Zou et al · 2023
Later among the works it cites.
“Real-world robot learning with masked visual pre-training”
Ilija Radosavovic et al · 2023
Later among the works it cites.
“Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?”
Arjun Majumdar et al · 2023
Later among the works it cites.
“An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning”
Jaesik Yoon, Yi-Fu Wu, Heechul Bae and Sungjin Ahn · 2023
Later among the works it cites.
“Teach a Robot to FISH: Versatile Imitation from One Minute of Demonstrations”
Siddhant Haldar, Jyothish Pari, Ananta Rai and Lerrel Pinto · 2023
Later among the works it cites.
“PDSketch: Integrated Planning Domain Programming and Learning”
Jiayuan Mao, Tom“’as Lozano-P“’erez, Joshua Tenenbaum and Leslie Kaelbling · 2023
Later among the works it cites.
“FOCUS: Object-Centric World Models for Robotics Manipulation”
Stefano Ferraro, Pietro Mazzaglia, Tim Verbelen and Bart Dhoedt · 2023
Later among the works it cites.
“Pave the Way to Grasp Anything: Transferring Foundation Models for Universal Pick-Place Robots”
Jiange Yang et al · 2023
Later among the works it cites.
“Learning Generalizable Manipulation Policies with Object-Centric 3D Representations”
Yifeng Zhu, Zhenyu Jiang, Peter Stone and Yuke Zhu · 2023
Later among the works it cites.
“RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation”
Sourav Garg et al · 2023
Later among the works it cites.
Xu Zhao et al · 2023
Later among the works it cites.
“Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies”
Jianing Qian, Anastasios Panagopoulos and Dinesh Jayaraman · 2024
Closest in time.