Fetching the paper…
Reading the bibliography…
The household rearrangement task involves spotting misplaced objects in a scene and accommodate them with proper places.
From 3d scene geometry to human workspace. In CVPR 2011 . IEEE, 1961–1968
Abhinav Gupta, Scott Satkin, Alexei A Efros, and Martial Hebert. 2011 · 1968
Earlier work this paper cites.
The theory of affordances
James J Gibson. 1977 · 1977
Earlier work this paper cites.
Cumulated gain-based evaluation of IR techniques
Kalervo Järvelin and Jaana Kekäläinen. 2002 · 2002
Earlier work this paper cites.
Noise robust spectral clustering. In 2007 IEEE 11th International Conference on Computer Vision . IEEE, 1–8
Zhenguo Li, Jianzhuang Liu, Shifeng Chen, and Xiaoou Tang. 2007 · 2007
Earlier work this paper cites.
A tutorial on spectral clustering
Ulrike Von Luxburg. 2007 · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
What makes a chair a chair?. In CVPR 2011 . IEEE, 1529–1536
Helmut Grabner, Juergen Gall, and Luc Van Gool. 2011 · 2011
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images. In Computer Vision–ECCV 2012: 12th European Conference on Computer Vision, Florence, Italy, October 7-13, 2012, Proceedings, Part V 12 . Springer, 746–760
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus. 2012 · 2012
Earlier work this paper cites.
Physically grounded spatio-temporal object affordances. In Computer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part III 13 . Springer, 831–847
Hema S Koppula and Ashutosh Saxena. 2014 · 2014
Earlier work this paper cites.
SceneGrok: Inferring action maps in 3D environments
Manolis Savva, Angel X Chang, Pat Hanrahan, Matthew Fisher, and Matthias Nießner. 2014 · 2014
Earlier work this paper cites.
Activity-centric scene synthesis for functional 3D scene modeling
Matthew Fisher, Manolis Savva, Yangyan Li, Pat Hanrahan, and Matthias Nießner. 2015 · 2015
Earlier work this paper cites.
Pigraphs: learning interaction snapshots from observations
Manolis Savva, Angel X Chang, Pat Hanrahan, Matthew Fisher, and Matthias Nießner. 2016 · 2016
Earlier work this paper cites.
Inferring forces and learning human utilities from videos. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 3823–3833
Yixin Zhu, Chenfanfu Jiang, Yibiao Zhao, Demetri Terzopoulos, and Song-Chun Zhu. 2016 · 2016
Earlier work this paper cites.
Synthesizing training data for object detection in indoor scenes
Georgios Georgakis, Arsalan Mousavian, Alexander C Berg, and Jana Kosecka. 2017 · 2017
Earlier work this paper cites.
Demo2vec: Reasoning object affordances from online videos. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 2139–2147
Kuan Fang, Te-Lin Wu, Daniel Yang, Silvio Savarese, and Joseph J Lim. 2018 · 2018
Earlier work this paper cites.
Functionality representations and applications for shape analysis. In Computer Graphics Forum , Vol. 37. Wiley Online Library, 603–624
Ruizhen Hu, Manolis Savva, and Oliver van Kaick. 2018 · 2018
Earlier work this paper cites.
Putting humans in a scene: Learning affordance in 3d indoor environments. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 12368–12376
Xueting Li, Sifei Liu, Kihwan Kim, Xiaolong Wang, Ming-Hsuan Yang, and Jan Kautz. 2019 · 2019
Earlier work this paper cites.
Grounded human-object interaction hotspots from video. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 8688–8697
Tushar Nagarajan, Christoph Feichtenhofer, and Kristen Grauman. 2019 · 2019
Earlier work this paper cites.
Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part I 16 . Springer, 422–440
Panos Achlioptas, Ahmed Abdelreheem, Fei Xia, Mohamed Elhoseiny, and Leonidas Guibas. 2020 · 2020
Earlier work this paper cites.
Rearrangement: A challenge for embodied ai
Dhruv Batra, Angel X Chang, Sonia Chernova, Andrew J Davison, Jia Deng, Vladlen Koltun, Sergey Levine, Jitendra Malik, Igor Mordatch, Roozbeh Mottaghi, et al · 2020
Earlier work this paper cites.
Ego-topo: Environment affordances from egocentric video. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 163–172
Tushar Nagarajan, Yanghao Li, Christoph Feichtenhofer, and Kristen Grauman. 2020 · 2020
Earlier work this paper cites.
A survey on performance metrics for object-detection algorithms. In 2020 international conference on systems, signals and image processing (IWSSIP) . IEEE, 237–242
Rafael Padilla, Sergio L Netto, and Eduardo AB Da Silva. 2020 · 2020
Cited alongside, same era.
Fusion-aware point convolution for online semantic 3d scene segmentation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 4534–4543
Jiazhao Zhang, Chenyang Zhu, Lintao Zheng, and Kai Xu. 2020 · 2020
Cited alongside, same era.
3d affordancenet: A benchmark for visual object affordance understanding. In proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 1778–1787
Shengheng Deng, Xun Xu, Chaozheng Wu, Ke Chen, and Kui Jia. 2021 · 2021
Cited alongside, same era.
Synthesizing scene-aware virtual reality teleport graphs
Changyang Li, Haikun Huang, Jyh-Ming Lien, and Lap-Fai Yu. 2021 · 2021
Cited alongside, same era.
Ocrtoc: A cloud-based competition and benchmark for robotic grasping and manipulation
Efficient Object Rearrangement via Multi-view Fusion
Dehao Huang, Chao Tang, and Hong Zhang. 2023 · 2023
Later among the works it cites.
ConceptFusion: Open-set Multimodal 3D Mapping
Krishna Murthy Jatavallabhula, Alihusein Kuwajerwala, Qiao Gu, Mohd Omama, Tao Chen, Shuang Li, Ganesh Iyer, Soroush Saryazdi, Nikhil Keetha, Ayush Tewari, Joshua B. Tenenbaum, Celso Miguel de Melo, Madhava Krishna, Liam Paull, Florian Shkurti, and Antonio Torralba. 2023 · 2023
Later among the works it cites.
Mukul Khanna, Yongsen Mao, Hanxiao Jiang, Sanjay Haresh, Brennan Shacklett, Dhruv Batra, Alexander Clegg, Eric Undersander, Angel X. Chang, and Manolis Savva. 2023 · 2023
Later among the works it cites.
Putting people in their place: Affordance-aware human insertion into scenes. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 17089–17099
Sumith Kulal, Tim Brooks, Alex Aiken, Jiajun Wu, Jimei Yang, Jingwan Lu, Alexei A Efros, and Krishna Kumar Singh. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ziyuan Liu, Wei Liu, Yuzhe Qin, Fanbo Xiang, Minghao Gou, Songyan Xin, Maximo A Roa, Berk Calli, Hao Su, Yu Sun, et al · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision. In International conference on machine learning . PMLR, 8748–8763
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Relationship oriented affordance learning through manipulation graph construction
Chao Tang, Jingwen Yu, Weinan Chen, and Hong Zhang. 2021 · 2021
Cited alongside, same era.
Visual room rearrangement. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 5922–5931
Luca Weihs, Matt Deitke, Aniruddha Kembhavi, and Roozbeh Mottaghi. 2021 · 2021
Cited alongside, same era.
Scanqa: 3d question answering for spatial scene understanding. In proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 19129–19139
Daichi Azuma, Taiki Miyanishi, Shuhei Kurita, and Motoaki Kawanabe. 2022 · 2022
Cited alongside, same era.
Disarm: displacement aware relation module for 3d detection. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 16980–16989
Yao Duan, Chenyang Zhu, Yuqing Lan, Renjiao Yi, Xinwang Liu, and Kai Xu. 2022 · 2022
Cited alongside, same era.
Housekeep: Tidying virtual households using commonsense reasoning. In European Conference on Computer Vision . Springer, 355–373
Yash Kant, Arun Ramachandran, Sriram Yenamandra, Igor Gilitschenski, Dhruv Batra, Andrew Szot, and Harsh Agrawal. 2022 · 2022
Cited alongside, same era.
IFR-Explore: Learning Inter-object Functional Relationships in 3D Indoor Scenes. In International Conference on Learning Representations
QI LI, Kaichun Mo, Yanchao Yang, Hang Zhao, and Leonidas Guibas. 2022 · 2022
Cited alongside, same era.
One-Shot Open Affordance Learning with Foundation Models
Gen Li, Deqing Sun, Laura Sevilla-Lara, and Varun Jampani. 2023 · 2023
Later among the works it cites.
StructDiffusion: Language-Guided Creation of Physically-Valid Structures using Unseen Objects. In RSS 2023
Weiyu Liu, Yilun Du, Tucker Hermans, Sonia Chernova, and Chris Paxton. 2023 · 2023
Later among the works it cites.
Ovir-3d: Open-vocabulary 3d instance retrieval without training on 3d data. In Conference on Robot Learning . PMLR, 1610–1620
Shiyang Lu, Haonan Chang, Eric Pu Jing, Abdeslam Boularias, and Kostas Bekris. 2023 · 2023
Later among the works it cites.
Open-vocabulary affordance detection in 3d point clouds. In 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 5692–5698
Toan Nguyen, Minh Nhat Vu, An Vuong, Dzung Nguyen, Thieu Vo, Ngan Le, and Anh Nguyen. 2023 · 2023
Later among the works it cites.
Grid: Scene-graph-based instruction-driven robotic task planning
Zhe Ni, Xiao-Xin Deng, Cong Tai, Xin-Yue Zhu, Xiang Wu, Yong-Jin Liu, and Long Zeng. 2023 · 2023
Later among the works it cites.
Habitat 3.0: A co-habitat for humans, avatars and robots
Xavier Puig, Eric Undersander, Andrew Szot, Mikael Dallaire Cote, Tsung-Yen Yang, Ruslan Partsey, Ruta Desai, Alexander William Clegg, Michal Hlavac, So Yeon Min, et al · 2023
Later among the works it cites.
Saynav: Grounding large language models for dynamic planning to navigation in new environments
Abhinav Rajvanshi, Karan Sikka, Xiao Lin, Bhoram Lee, Han-Pang Chiu, and Alvaro Velasquez. 2023 · 2023
Later among the works it cites.
Sayplan: Grounding large language models using 3d scene graphs for scalable robot task planning. In 7th Annual Conference on Robot Learning
Krishan Rana, Jesse Haviland, Sourav Garg, Jad Abou-Chakra, Ian Reid, and Niko Suenderhauf. 2023 · 2023
Later among the works it cites.
Open-vocabulary affordance detection using knowledge distillation and text-point correlation
Tuan Van Vo, Minh Nhat Vu, Baoru Huang, Toan Nguyen, Ngan Le, Thieu Vo, and Anh Nguyen. 2023 · 2023
Later among the works it cites.
Chat-3d: Data-efficiently tuning large language model for universal dialogue of 3d scenes
Zehan Wang, Haifeng Huang, Yang Zhao, Ziang Zhang, and Zhou Zhao. 2023 · 2023
Later among the works it cites.
TidyBot: Personalized Robot Assistance with Large Language Models
Jimmy Wu, Rika Antonova, Adam Kan, Marion Lepert, Andy Zeng, Shuran Song, Jeannette Bohg, Szymon Rusinkiewicz, and Thomas Funkhouser. 2023 · 2023
Later among the works it cites.
HomeRobot: Open Vocab Mobile Manipulation
Sriram Yenamandra, Arun Ramachandran, Karmesh Yadav, Austin Wang, Mukul Khanna, Theophile Gervet, Tsung-Yen Yang, Vidhi Jain, Alex William Clegg, John Turner, Zsolt Kira, Manolis Savva, Angel Chang, Devendra Singh Chaplot, Dhruv Batra, Roozbeh Mottaghi, Yonatan Bisk, and Chris Paxton. 2023 · 2023
Later among the works it cites.
3d-aware object goal navigation via simultaneous exploration and identification. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6672–6682
Jiazhao Zhang, Liu Dai, Fanpeng Meng, Qingnan Fan, Xuelin Chen, Kai Xu, and He Wang. 2023 · 2023
Later among the works it cites.
Dongge Han, Trevor McInroe, Adam Jelley, Stefano V Albrecht, Peter Bell, and Amos Storkey. 2024 · 2024
Closest in time.
Advances in Data-Driven Analysis and Synthesis of 3D Indoor Scenes. In Computer Graphics Forum , Vol. 43. Wiley Online Library, e14927
Akshay Gadi Patil, Supriya Gadi Patil, Manyi Li, Matthew Fisher, Manolis Savva, and Hao Zhang. 2024 · 2024
Closest in time.
Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance
Zan Wang, Yixin Chen, Baoxiong Jia, Puhao Li, Jinlu Zhang, Jingze Zhang, Tengyu Liu, Yixin Zhu, Wei Liang, and Siyuan Huang. 2024 · 2024
Closest in time.
RAIL: Robot Affordance Imagination with Large Language Models
Ceng Zhang, Xin Meng, Dongchen Qi, and Gregory S Chirikjian. 2024 · 2024
Closest in time.