Fetching the paper…
Reading the bibliography…
In robot learning, the observation space is crucial due to the distinct characteristics of different modalities, which can potentially become a bottleneck alongside policy design.
Nearest neighbor pattern classification
T. Cover and P. Hart · 1967
Earlier work this paper cites.
Pattern classification and scene analysis , volume 3
R. O. Duda, P. E. Hart, et al · 1973
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. A. Pomerleau · 1988
Earlier work this paper cites.
The farthest point strategy for progressive image sampling
Y. Eldar, M. Lindenbaum, M. Porat, and Y. Y. Zeevi · 1997
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 1999
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Torchvision the machine-vision package of torch
S. Marcel and Y. Rodriguez · 2010
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
K. Sohn, H. Lee, and X. Yan · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Earlier work this paper cites.
Deep visual foresight for planning robot motion
C. Finn and S. Levine · 2017
Earlier work this paper cites.
The" something something" video database for learning and evaluating visual common sense
R. Goyal, S. Ebrahimi Kahou, V. Michalski, J. Materzynska, S. Westphal, H. Kim, V. Haenel, I. Fruend, P. Yianilos, M. Mueller-Freitag, et al · 2017
Earlier work this paper cites.
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
S. Gu, E. Holly, T. Lillicrap, and S. Levine · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
I. Loshchilov and F. Hutter · 2017
Earlier work this paper cites.
Pointnet: Deep learning on point sets for 3d classification and segmentation
C. R. Qi, H. Su, K. Mo, and L. J. Guibas · 2017
Earlier work this paper cites.
Scaling egocentric vision: The epic-kitchens dataset
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, et al · 2018
Earlier work this paper cites.
Spnets: Differentiable fluid dynamics for deep neural networks
C. Schenck and D. Fox · 2018
Earlier work this paper cites.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Earlier work this paper cites.
Stereo magnification: Learning view synthesis using multiplane images
T. Zhou, R. Tucker, J. Flynn, G. Fyffe, and N. Snavely · 2018
Earlier work this paper cites.
4d spatio-temporal convnets: Minkowski convolutional neural networks
C. Choy, J. Gwak, and S. Savarese · 2019
Earlier work this paper cites.
PyTorch Lightning, Mar. 2019
W. Falcon and The PyTorch Lightning team · 2019
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, et al · 2019
Earlier work this paper cites.
Super-convergence: Very fast training of neural networks using large learning rates
L. N. Smith and N. Topin · 2019
Earlier work this paper cites.
Hydra - a framework for elegantly configuring complex applications
O. Yadan · 2019
Earlier work this paper cites.
On the continuity of rotation representations in neural networks
Y. Zhou, C. Barnes, J. Lu, J. Yang, and H. Li · 2019
Earlier work this paper cites.
A simple framework for contrastive learning of visual representations
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, et al · 2020
Earlier work this paper cites.
Momentum contrast for unsupervised visual representation learning
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick · 2020
Earlier work this paper cites.
Rlbench: The robot learning benchmark & learning environment
S. James, Z. Ma, D. R. Arrojo, and A. J. Davison · 2020
Cited alongside, same era.
Understanding human hands in contact at internet scale
D. Shan, J. Geng, M. Shu, and D. F. Fouhey · 2020
Cited alongside, same era.
Pointcontrast: Unsupervised pre-training for 3d point cloud understanding
S. Xie, J. Gu, D. Guo, C. R. Qi, L. Guibas, and O. Litany · 2020
Cited alongside, same era.
Learning to see before learning to act: Visual pre-training for manipulation
L. Yen-Chen, A. Zeng, S. Song, P. Isola, and T.-Y. Lin · 2020
Cited alongside, same era.
Beit: Bert pre-training of image transformers
H. Bao, L. Dong, S. Piao, and F. Wei · 2021
Cited alongside, same era.
Emerging properties in self-supervised vision transformers
Rt-2: Vision-language-action models transfer web knowledge to robotic control
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, X. Chen, K. Choromanski, T. Ding, D. Driess, A. Dubey, C. Finn, et al · 2023
Later among the works it cites.
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Later among the works it cites.
Learning human-to-robot handovers from point clouds
S. Christen, W. Yang, C. Pérez-D’Arpino, O. Hilliges, D. Fox, and Y.-W. Chao · 2023
Later among the works it cites.
Pointcept: A codebase for point cloud perception research
P. Contributors · 2023
Later among the works it cites.
Act3d: 3d feature field transformers for multi-task robotic manipulation
T. Gervet, Z. Xian, N. Gkanatsios, and K. Fragkiadaki · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Caron, H. Touvron, I. Misra, H. Jégou, J. Mairal, P. Bojanowski, and A. Joulin · 2021
Cited alongside, same era.
An empirical study of training self-supervised vision transformers
X. Chen, S. Xie, and K. He · 2021
Cited alongside, same era.
Rgb matters: Learning 7-dof grasp poses on monocular rgbd images
M. Gou, H.-S. Fang, Z. Zhu, S. Xu, C. Wang, and C. Lu · 2021
Cited alongside, same era.
Exploring data-efficient 3d scene understanding with contrastive scene contexts
J. Hou, B. Graham, M. Nießner, and S. Xie · 2021
Cited alongside, same era.
Perceiver io: A general architecture for structured inputs & outputs
A. Jaegle, S. Borgeaud, J.-B. Alayrac, C. Doersch, C. Ionescu, D. Ding, S. Koppula, D. Zoran, A. Brock, E. Shelhamer, et al · 2021
Cited alongside, same era.
A review of robot learning for manipulation: Challenges, representations, and algorithms
O. Kroemer, S. Niekum, and G. Konidaris · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Cited alongside, same era.
Maniskill2: A unified benchmark for generalizable manipulation skills
J. Gu, F. Xiang, X. Li, Z. Ling, X. Liu, T. Mu, Y. Tang, S. Tao, X. Wei, Y. Yao, et al · 2023
Later among the works it cites.
Ponder: Point cloud pre-training via neural rendering
D. Huang, S. Peng, T. He, H. Yang, X. Zhou, and W. Ouyang · 2023
Later among the works it cites.
Chain-of-thought predictive control
Z. Jia, F. Liu, V. Thumuluri, L. Chen, Z. Huang, and H. Su · 2023
Later among the works it cites.
Vima: General robot manipulation with multimodal prompts
Y. Jiang, A. Gupta, Z. Zhang, G. Wang, Y. Dou, Y. Chen, L. Fei-Fei, A. Anandkumar, Y. Zhu, and L. Fan · 2023
Later among the works it cites.
Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision
L. Ling, Y. Sheng, Z. Tu, W. Zhao, C. Xin, K. Wan, L. Yu, Q. Guo, Z. Yu, Y. Lu, et al · 2023
Later among the works it cites.
Where are we in the search for an artificial visual cortex for embodied intelligence?
A. Majumdar, K. Yadav, S. Arnaud, Y. J. Ma, C. Chen, S. Silwal, A. Jain, V.-P. Berges, P. Abbeel, J. Malik, et al · 2023
Later among the works it cites.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2023
Later among the works it cites.
Dinov2: Learning robust visual features without supervision
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, et al · 2023
Later among the works it cites.
Open x-embodiment: Robotic learning datasets and rt-x models
A. Padalkar, A. Pooley, A. Jain, A. Bewley, A. Herzog, A. Irpan, A. Khazatsky, A. Rai, A. Singh, A. Brohan, et al · 2023
Later among the works it cites.
Real-world robot learning with masked visual pre-training
I. Radosavovic, T. Xiao, S. James, P. Abbeel, J. Malik, and T. Darrell · 2023
Later among the works it cites.
Perceiver-actor: A multi-task transformer for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2023
Later among the works it cites.
Octo: An open-source generalist robot policy, 2023
O. M. Team, D. Ghosh, H. Walke, K. Pertsch, K. Black, O. Mees, S. Dasari, J. Hejna, C. Xu, J. Luo, et al · 2023
Later among the works it cites.
Masked scene contrast: A scalable framework for unsupervised 3d representation learning
X. Wu, X. Wen, X. Liu, and H. Zhao · 2023
Later among the works it cites.
Gnfactor: Multi-task real robot learning with generalizable neural feature fields
Y. Ze, G. Yan, Y.-H. Wu, A. Macaluso, Y. Ge, J. Ye, N. Hansen, L. E. Li, and X. Wang · 2023
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
T. Z. Zhao, V. Kumar, S. Levine, and C. Finn · 2023
Later among the works it cites.
Ponderv2: Pave the way for 3d foundataion model with a universal pre-training paradigm
H. Zhu, H. Yang, X. Wu, D. Huang, S. Zhang, X. He, T. He, H. Zhao, C. Shen, Y. Qiao, et al · 2023
Later among the works it cites.
Realrobot: A project for open-sourced robot learning research. https://github.com/HaoyiZhu/RealRobot
R. Contributors · 2024
Closest in time.
Low-cost robot arm
A. Koch · 2024
Closest in time.
Evaluating real-world robot manipulation policies in simulation
X. Li, K. Hsu, J. Gu, K. Pertsch, O. Mees, H. R. Walke, C. Fu, I. Lunawat, I. Sieh, S. Kirmani, et al · 2024
Closest in time.
Consistency policy: Accelerated visuomotor policies via consistency distillation
A. Prasad, K. Lin, J. Wu, L. Zhou, and J. Bohg · 2024
Closest in time.
Dexcap: Scalable and portable mocap data collection system for dexterous manipulation
C. Wang, H. Shi, W. Wang, R. Zhang, L. Fei-Fei, and C. K. Liu · 2024
Closest in time.
Spatiotemporal predictive pre-training for robotic motor control
J. Yang, B. Liu, J. Fu, B. Pan, G. Wu, and L. Wang · 2024
Closest in time.
Y. Ze, G. Zhang, K. Zhang, C. Hu, M. Wang, and H. Xu · 2024
Closest in time.