Fetching the paper…
Reading the bibliography…
Visual affordance learning is a key component for robots to understand how to interact with objects.
M. A. Fischler and R. C. Bolles, “Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography,” Communications of the ACM , vol. 24, no. 6, pp. 381–395, 1981
1981
Earlier work this paper cites.
M. J. Swain and D. H. Ballard, “Color indexing,” International journal of computer vision , vol. 7, no. 1, 1991
1991
Earlier work this paper cites.
M. Müller, “Dynamic time warping,” Information Retrieval for Music and Motion , 2007
2007
Earlier work this paper cites.
H. Bay, A. Ess, T. Tuytelaars, and L. Van Gool, “Speeded-up robust features (surf),” Computer vision and image understanding , vol. 110, no. 3, 2008
2008
Earlier work this paper cites.
A. Gupta, A. Kembhavi, and L. S. Davis, “Observing human-object interactions: Using spatial and functional compatibility for recognition,” IEEE transactions on pattern analysis and machine intelligence , vol. 31, no. 10, pp. 1775–1789, 2009
2009
Earlier work this paper cites.
T. Judd, K. Ehinger, F. Durand, and A. Torralba, “Learning to predict where humans look,” in 2009 IEEE 12th International Conference on Computer Vision , 2009, pp. 2106–2113
2009
Earlier work this paper cites.
J. J. Gibson, The ecological approach to visual perception: classic edition . Psychology press, 2014
2014
Earlier work this paper cites.
V. G. Kim, S. Chaudhuri, L. Guibas, and T. Funkhouser, “Shape2pose: Human-centric shape analysis,” ACM Transactions on Graphics (TOG) , vol. 33, no. 4, pp. 1–12, 2014
2014
Earlier work this paper cites.
S. Kazemzadeh, V. Ordonez, M. Matten, and T. Berg, “ReferItGame: Referring to objects in photographs of natural scenes,” in Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2014, pp. 787–798
2014
Earlier work this paper cites.
A. Myers, C. L. Teo, C. Fermüller, and Y. Aloimonos, “Affordance detection of tool parts from geometric features,” in ICRA , 2015
2015
Earlier work this paper cites.
S. Gupta, P. Arbeláez, R. Girshick, and J. Malik, “Indoor scene understanding with rgb-d images: Bottom-up segmentation, object detection and semantic segmentation,” International Journal of Computer Vision , vol. 112, pp. 133–149, 2015
2015
Earlier work this paper cites.
S. Gupta and J. Malik, “Visual semantic role labeling,” arXiv preprint arXiv:1505.04474 , 2015
2015
Earlier work this paper cites.
B. A. Plummer, L. Wang, C. M. Cervantes, J. C. Caicedo, J. Hockenmaier, and S. Lazebnik, “Flickr30k Entities: Collecting Region-to-Phrase Correspondences for Richer Image-to-Sentence Models,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
L. Yu, P. Poirson, S. Yang, A. C. Berg, and T. L. Berg, “Modeling Context in Referring Expressions,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2016
2016
Earlier work this paper cites.
J. Mao, J. Huang, A. Toshev, O. Camburu, A. Yuille, and K. Murphy, “Generation and Comprehension of Unambiguous Object Descriptions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Earlier work this paper cites.
A. Nguyen, D. Kanoulas, D. G. Caldwell, and N. G. Tsagarakis, “Object-based affordances detection with convolutional neural networks and dense conditional random fields,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2017
2017
Earlier work this paper cites.
J. Sawatzky, A. Srikantha, and J. Gall, “Weakly supervised affordance detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017
2017
Earlier work this paper cites.
M. Honnibal and I. Montani, “spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing,” 2017
2017
Earlier work this paper cites.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, and M. Wray, “Scaling egocentric vision: The epic-kitchens dataset,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 720–736
2018
Earlier work this paper cites.
C.-Y. Chuang, J. Li, A. Torralba, and S. Fidler, “Learning to act properly: Predicting and explaining affordances from images,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Cited alongside, same era.
V. Dumoulin, E. Perez, N. Schucher, F. Strub, H. d. Vries, A. Courville, and Y. Bengio, “Feature-wise transformations,” Distill , vol. 3, no. 7, p. e11, 2018
2018
Cited alongside, same era.
Z. Bylinskii, T. Judd, A. Oliva, A. Torralba, and F. Durand, “What do different evaluation metrics tell us about saliency models?” IEEE transactions on pattern analysis and machine intelligence , vol. 41, no. 3, pp. 740–757, 2018
2018
Cited alongside, same era.
T. Nagarajan, C. Feichtenhofer, and K. Grauman, “Grounded human-object interaction hotspots from video,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2019, pp. 8688–8697
2019
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
S. Liu, S. Tripathi, S. Majumdar, and X. Wang, “Joint hand motion and interaction hotspots prediction from egocentric videos,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 3282–3292
2022
Later among the works it cites.
L. Mur-Labadia, J. J. Guerrero, and R. Martinez-Cantin, “Multi-label affordance mapping from egocentric vision,” in ICCV , 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Liu, C. Liu, Y. Bai, and A. L. Yuille, “CLEVR-Ref+: Diagnosing Visual Reasoning With Referring Expressions,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 4185–4195
2019
Cited alongside, same era.
T. Nagarajan, Y. Li, C. Feichtenhofer, and K. Grauman, “Ego-topo: Environment affordances from egocentric video,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Cited alongside, same era.
C. Wu, Z. Lin, S. Cohen, T. Bui, and S. Maji, “Phrasecut: Language-based image segmentation in the wild,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2020, pp. 10 216–10 225
2020
Cited alongside, same era.
D. Shan, J. Geng, M. Shu, and D. F. Fouhey, “Understanding human hands in contact at internet scale,” in CVPR , 2020
2020
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in European conference on computer vision , 2020, pp. 213–229
2020
Cited alongside, same era.
Y. Zha, S. Bhambri, and L. Guan, “Contrastively learning visual attention as affordance cues from demonstrations for robotic grasping,” in 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2021, pp. 7835–7842
2021
Cited alongside, same era.
A. Kamath, M. Singh, Y. LeCun, G. Synnaeve, I. Misra, and N. Carion, “Mdetr - modulated detection for end-to-end multi-modal understanding,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2021, pp. 1780–1790
2021
Cited alongside, same era.
S. Deng, X. Xu, C. Wu, K. Chen, and K. Jia, “3d affordancenet: A benchmark for visual object affordance understanding,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 1778–1787
2021
Cited alongside, same era.
T. Nguyen, M. N. Vu, A. Vuong, D. Nguyen, T. Vo, N. Le, and A. Nguyen, “Open-vocabulary affordance detection in 3d point clouds,” in IROS , 2023
2023
Later among the works it cites.
G. Li, V. Jampani, D. Sun, and L. Sevilla-Lara, “Locate: Localize and transfer object parts for weakly supervised affordance grounding,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 10 922–10 931
2023
Later among the works it cites.
J. Chen, D. Gao, K. Q. Lin, and M. Z. Shou, “Affordance grounding from demonstration video to target image,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 6799–6808
2023
Later among the works it cites.
S. Bahl, R. Mendonca, L. Chen, U. Jain, and D. Pathak, “Affordances from human videos as a versatile representation for robotics,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 13 778–13 790
2023
Later among the works it cites.
Z. Khalifa and S. A. A. Shah, “A large scale multi-view rgbd visual affordance learning dataset,” in 2023 IEEE International Conference on Image Processing (ICIP) , 2023, pp. 1325–1329
2023
Later among the works it cites.
Z. Yu, Y. Huang, R. Furuta, T. Yagi, Y. Goutsu, and Y. Sato, “Fine-grained affordance annotation for egocentric hand-object interaction videos,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , 2023, pp. 2155–2163
2023
Later among the works it cites.
R. Fan, T. Wang, M. Hirano, and Y. Yamakawa, “One-shot affordance learning (osal): Learning to manipulate articulated objects by observing once,” in 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2023, pp. 2955–2962
2023
Later among the works it cites.
J. Tang, G. Zheng, J. Yu, and S. Yang, “Cotdet: Affordance knowledge prompting for task driven object detection,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 3068–3078
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Wang, M. K. Singh, and L. Torresani, “Ego-only: Egocentric action detection without exocentric transferring,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 5250–5261
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.