Fetching the paper…
Reading the bibliography…
Enabling robotic manipulation that generalizes to out-of-distribution scenes is a crucial step toward open-world embodied intelligence.
J. J. Gibson, “The ecological approach to the visual perception of pictures,”
1978
Earlier work this paper cites.
S. H. Creem-Regehr and J. N. Lee, “Neural representations of graspable objects: are tools special?”
2005
Earlier work this paper cites.
Y. Zhu, A. Fathi, and L. Fei-Fei, “Reasoning about object affordances in a knowledge base representation,” in
2014
Earlier work this paper cites.
F. Saxen and A. Al-Hamadi, “Color-based skin segmentation: An evaluation of the state of the art,” in
2014
Earlier work this paper cites.
A. Myers, C. L. Teo, C. Fermüller, and Y. Aloimonos, “Affordance detection of tool parts from geometric features,” in
2015
Earlier work this paper cites.
M. Hassan and A. Dharmaratne, “Attribute based affordance detection from human-object interaction images,” in
2016
Earlier work this paper cites.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price
2018
Earlier work this paper cites.
K. Fang, T.-L. Wu, D. Yang, S. Savarese, and J. J. Lim, “Demo2vec: Reasoning object affordances from online videos,” in
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” in
2018
Earlier work this paper cites.
T. Nagarajan, C. Feichtenhofer, and K. Grauman, “Grounded human-object interaction hotspots from video,” in
2019
Earlier work this paper cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in
2020
Earlier work this paper cites.
D. Shan, J. Geng, M. Shu, and D. F. Fouhey, “Understanding human hands in contact at internet scale,” in
2020
Earlier work this paper cites.
——, “The epic-kitchens dataset: Collection, challenges and baselines,”
2020
Earlier work this paper cites.
S. Amir, Y. Gandelsman, S. Bagon, and T. Dekel, “Deep vit features as dense visual descriptors,”
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
K. Mo, L. J. Guibas, M. Mukadam, A. Gupta, and S. Tulsiani, “Where2act: From pixels to actions for articulated 3d objects,” in
2021
Earlier work this paper cites.
Z. Hou, B. Yu, Y. Qiao, X. Peng, and D. Tao, “Affordance transfer learning for human-object interaction detection,” in
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” in
2021
Earlier work this paper cites.
P. Mandikal and K. Grauman, “Learning dexterous grasping with object-centric visual affordances,” in
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
M. Goyal, S. Modi, R. Goyal, and S. Gupta, “Human hands as probes for interactive object understanding,” in
2022
Cited alongside, same era.
Y. Wang, R. Wu, K. Mo, J. Ke, Q. Fan, L. J. Guibas, and H. Dong, “Adaafford: Learning to adapt manipulation affordance for 3d articulated objects via few-shot interactions,” in
2022
Cited alongside, same era.
K. Mo, Y. Qin, F. Xiang, H. Su, and L. Guibas, “O2o-afford: Annotation-free large-scale object-object affordance learning,” in
2022
Cited alongside, same era.
S. Liu, S. Tripathi, S. Majumdar, and X. Wang, “Joint hand motion and interaction hotspots prediction from egocentric videos,” in
2022
Cited alongside, same era.
H. Luo, W. Zhai, J. Zhang, Y. Cao, and D. Tao, “Learning affordance grounding from exocentric images,” in
2022
Cited alongside, same era.
H. Geng, Z. Li, Y. Geng, J. Chen, H. Dong, and H. Wang, “Partmanip: Learning cross-category generalizable part manipulation policy from point cloud observations,” in
2023
Later among the works it cites.
Y. Geng, B. An, H. Geng, Y. Chen, Y. Yang, and H. Dong, “Rlafford: End-to-end affordance learning for robotic manipulation,” in
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Y.-H. Wu, J. Wang, and X. Wang, “Learning generalizable dexterous manipulation from human grasp affordance,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Damen, H. Doughty, G. M. Farinella, A. Furnari, E. Kazakos, J. Ma, D. Moltisanti, J. Munro, T. Perrett, W. Price
2022
Cited alongside, same era.
K. Grauman, A. Westbury, E. Byrne, Z. Chavis, A. Furnari, R. Girdhar, J. Hamburger, H. Jiang, M. Liu, X. Liu
2022
Cited alongside, same era.
S. Bahl, A. Gupta, and D. Pathak, “Human-to-robot imitation in the wild,”
2022
Cited alongside, same era.
——, “Dexvip: Learning dexterous grasping with human hand pose priors from video,” in
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Y.-L. Li, H. Fan, Z. Qiu, Y. Dou, L. Xu, H.-S. Fang, P. Guo, H. Su, D. Wang, W. Wu
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Ye, X. Li, A. Gupta, S. De Mello, S. Birchfield, J. Song, S. Tulsiani, and S. Liu, “Affordance diffusion: Synthesizing hand-object interactions,” in
2023
Later among the works it cites.
R. Mendonca, S. Bahl, and D. Pathak, “Structured world models from human videos,”
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
L. Medeiros, “lang-segment-anything,”
2023
Later among the works it cites.
Y. Xu, W. Wan, J. Zhang, H. Liu, Z. Shan, H. Shen, R. Wang, H. Geng, Y. Weng, J. Chen
2023
Later among the works it cites.
2023
Later among the works it cites.
A. Rashid, S. Sharma, C. M. Kim, J. Kerr, L. Y. Chen, A. Kanazawa, and K. Goldberg, “Language embedded radiance fields for zero-shot task-oriented grasping,” in
2023
Later among the works it cites.
P. Li, T. Liu, Y. Li, Y. Geng, Y. Zhu, Y. Yang, and S. Huang, “Gendexgrasp: Generalizable dexterous grasping,” in
2023
Later among the works it cites.
G. Luo, L. Dunlap, D. H. Park, A. Holynski, and T. Darrell, “Diffusion hyperfeatures: Searching through time and space for semantic correspondence,” in
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Li, S. Dikhale, S. Iba, and N. Jamali, “Vihope: Visuotactile in-hand object 6d pose estimation with shape completion,”
2023
Later among the works it cites.
2023
Later among the works it cites.
N. Di Palo and E. Johns, “On the effectiveness of retrieval, alignment, and replay in manipulation,”
2024
Closest in time.