Fetching the paper…
Reading the bibliography…
Affordance grounding aims to localize the interaction regions for the manipulated objects in the scene image according to given instructions.
Swain, M.J., Ballard, D.H.: Color indexing. International journal of computer vision 7
1991
Earlier work this paper cites.
Peters, R.J., Iyer, A., Itti, L., Koch, C.: Components of bottom-up gaze allocation in natural images. Vision research 45
2005
Earlier work this paper cites.
Nguyen, A., Kanoulas, D., Caldwell, D., Tsagarakis, N.: Object-based affordances detection with convolutional neural networks and dense conditional random fields (09 2017). https://doi.org/10.1109/IROS.2017.8206484
2017
Earlier work this paper cites.
Nguyen, A., Kanoulas, D., Caldwell, D.G., Tsagarakis, N.G.: Object-based affordances detection with convolutional neural networks and dense conditional random fields. In: 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). pp. 5908–5915. IEEE (2017)
2017
Earlier work this paper cites.
Bylinskii, Z., Judd, T., Oliva, A., Torralba, A., Durand, F.: What do different evaluation metrics tell us about saliency models? IEEE transactions on pattern analysis and machine intelligence 41
2018
Earlier work this paper cites.
Chuang, C.Y., Li, J., Torralba, A., Fidler, S.: Learning to act properly: Predicting and explaining affordances from images. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. pp. 975–983 (2018)
2018
Earlier work this paper cites.
Do, T.T., Nguyen, A., Reid, I.: Affordancenet: An end-to-end deep learning approach for object affordance detection. In: 2018 IEEE international conference on robotics and automation (ICRA). pp. 5882–5889. IEEE (2018)
2018
Earlier work this paper cites.
Fang, K., Wu, T.L., Yang, D., Savarese, S., Lim, J.J.: Demo2vec: Reasoning object affordances from online videos. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. pp. 2139–2147 (2018)
2018
Earlier work this paper cites.
Mi, J., Tang, S., Deng, Z., Goerner, M., Zhang, J.: Object affordance based multimodal fusion for natural human-robot interaction. Cognitive Systems Research 54
2019
Earlier work this paper cites.
Nagarajan, T., Feichtenhofer, C., Grauman, K.: Grounded human-object interaction hotspots from video. In: Proceedings of the IEEE/CVF International Conference on Computer Vision. pp. 8688–8697 (2019)
2019
Earlier work this paper cites.
Mi, J., Liang, H., Katsakis, N., Tang, S., Li, Q., Zhang, C., Zhang, J.: Intention-related natural language grounding via object affordance detection and intention semantic extraction. Frontiers in Neurorobotics 14
2020
Earlier work this paper cites.
Zhao, X., Cao, Y., Kang, Y.: Object affordance detection with relationship-aware network. Neural Computing and Applications 32
2020
Earlier work this paper cites.
Caron, M., Touvron, H., Misra, I., Jégou, H., Mairal, J., Bojanowski, P., Joulin, A.: Emerging properties in self-supervised vision transformers. In: Proceedings of the IEEE/CVF international conference on computer vision. pp. 9650–9660 (2021)
2021
Earlier work this paper cites.
Deng, S., Xu, X., Wu, C., Chen, K., Jia, K.: 3d affordancenet: A benchmark for visual object affordance understanding. In: proceedings of the IEEE/CVF conference on computer vision and pattern recognition. pp. 1778–1787 (2021)
2021
Earlier work this paper cites.
Hassanin, M., Khan, S., Tahtali, M.: Visual affordance and function understanding: A survey. ACM Computing Surveys (CSUR) 54
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.: Learning transferable visual models from natural language supervision. In: International conference on machine learning. pp. 8748–8763. PMLR (2021)
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
Grauman, K., Westbury, A., Byrne, E., Chavis, Z., Furnari, A., Girdhar, R., Hamburger, J., Jiang, H., Liu, M., Liu, X., et al.: Ego4d: Around the world in 3,000 hours of egocentric video. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 18995–19012 (2022)
2022
Cited alongside, same era.
Huang, W., Abbeel, P., Pathak, D., Mordatch, I.: Language models as zero-shot planners: Extracting actionable knowledge for embodied agents. In: International Conference on Machine Learning. pp. 9118–9147. PMLR (2022)
2022
Cited alongside, same era.
Li, J., Li, D., Xiong, C., Hoi, S.: Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation. In: International Conference on Machine Learning. pp. 12888–12900. PMLR (2022)
2023
Later among the works it cites.
Li, G., Jampani, V., Sun, D., Sevilla-Lara, L.: Locate: Localize and transfer object parts for weakly supervised affordance grounding. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 10922–10931 (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Li, Y.L., Xu, Y., Xu, X., Mao, X., Yao, Y., Liu, S., Lu, C.: Beyond object recognition: A new benchmark towards object concept learning. In: Proceedings of the IEEE/CVF International Conference on Computer Vision. pp. 20029–20040 (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
Lu, L., Zhai, W., Luo, H., Kang, Y., Cao, Y.: Phrase-based affordance detection via cyclic bilateral interaction. IEEE Transactions on Artificial Intelligence (2022)
2022
Cited alongside, same era.
Lüddecke, T., Ecker, A.: Image segmentation using text and image prompts. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition. pp. 7086–7096 (2022)
2022
Cited alongside, same era.
Luo, H., Zhai, W., Zhang, J., Cao, Y., Tao, D.: Learning affordance grounding from exocentric images. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 2252–2261 (2022)
2022
Cited alongside, same era.
Wang, S., Zhou, Z., Kan, Z.: When transformer meets robotic grasping: Exploits context for efficient grasp detection. IEEE robotics and automation letters 7
2022
Cited alongside, same era.
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Xia, F., Chi, E., Le, Q.V., Zhou, D., et al.: Chain-of-thought prompting elicits reasoning in large language models. Advances in Neural Information Processing Systems 35
2022
Cited alongside, same era.
Zhai, W., Luo, H., Zhang, J., Cao, Y., Tao, D.: One-shot object affordance detection in the wild. International Journal of Computer Vision 130
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
Liu, H., Li, C., Wu, Q., Lee, Y.J.: Visual instruction tuning. In: NeurIPS (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Luo, H., Zhai, W., Zhang, J., Cao, Y., Tao, D.: Grounded affordance from exocentric view. International Journal of Computer Vision pp. 1–25 (2023)
2023
Later among the works it cites.
Luo, H., Zhai, W., Zhang, J., Cao, Y., Tao, D.: Learning visual affordance grounding from demonstration videos. IEEE Transactions on Neural Networks and Learning Systems (2023)
2023
Later among the works it cites.
Nguyen, T., Vu, M.N., Vuong, A., Nguyen, D., Vo, T., Le, N., Nguyen, A.: Open-vocabulary affordance detection in 3d point clouds. In: 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). pp. 5692–5698. IEEE (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhou, Z., Wang, S., Chen, Z., Cai, M., Kan, Z.: A novel framework for improved grasping of thin and stacked objects. IEEE Transactions on Artificial Intelligence (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
Feng, G., Zhang, B., Gu, Y., Ye, H., He, D., Wang, L.: Towards revealing the mystery behind chain of thought: a theoretical perspective. Advances in Neural Information Processing Systems 36
2024
Closest in time.
2024
Closest in time.
Li, Z., Peng, B., He, P., Galley, M., Gao, J., Yan, X.: Guiding large language models via directional stimulus prompting. Advances in Neural Information Processing Systems 36
2024
Closest in time.
2024
Closest in time.
Wang, S., Zhou, Z., Li, B., Li, Z., Kan, Z.: Multi-modal interaction with transformers: bridging robots and human with natural language. Robotica 42
2024
Closest in time.