Fetching the paper…
Reading the bibliography…
Embodied Question Answering (EQA) is a challenging task in embodied intelligence that requires agents to dynamically explore 3D environments, actively gather visual information, and perform multi-step reasoning to answer questions.
A frontier-based approach for autonomous exploration
Brian Yamauchi · 1997
Earlier work this paper cites.
Frontier-based exploration using multiple robots
Brian Yamauchi · 1998
Earlier work this paper cites.
Frontier-based probabilistic strategies for sensor-based exploration
Luigi Freda and Giuseppe Oriolo · 2005
Earlier work this paper cites.
Evaluating the efficiency of frontier-based exploration strategies
Dirk Holz, Nicola Basilico, Francesco Amigoni, and Sven Behnke · 2010
Earlier work this paper cites.
The question answering systems: A survey
Ali Mohamed Nabil Allam and Mohamed Hassan Haggag · 2012
Earlier work this paper cites.
Interactive environment exploration in clutter
Megha Gupta, Thomas Rühr, Michael Beetz, and Gaurav S Sukhatme · 2013
Earlier work this paper cites.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Earlier work this paper cites.
Embodied question answering
Abhishek Das, Samyak Datta, Georgia Gkioxari, Stefan Lee, Devi Parikh, and Dhruv Batra · 2018
Earlier work this paper cites.
Iqa: Visual question answering in interactive environments
Daniel Gordon, Aniruddha Kembhavi, Mohammad Rastegari, Joseph Redmon, Dieter Fox, and Ali Farhadi · 2018
Earlier work this paper cites.
Videonavqa: Bridging the gap between visual and embodied question answering
Cătălina Cangea, Eugene Belilovsky, Pietro Liò, and Aaron Courville · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et al · 2019
Earlier work this paper cites.
Deep reinforcement learning robot for search and rescue applications: Exploration in unknown cluttered environments
Farzad Niroui, Kaicheng Zhang, Zendai Kashino, and Goldie Nejat · 2019
Earlier work this paper cites.
Embodied question answering in photorealistic environments with point cloud perception
Erik Wijmans, Samyak Datta, Oleksandr Maksymets, Abhishek Das, Georgia Gkioxari, Stefan Lee, Irfan Essa, Devi Parikh, and Dhruv Batra · 2019
Earlier work this paper cites.
Multi-target embodied question answering
Licheng Yu, Xinlei Chen, Georgia Gkioxari, Mohit Bansal, Tamara L Berg, and Dhruv Batra · 2019
Earlier work this paper cites.
Depth and video segmentation based visual attention for embodied question answering
Haonan Luo, Guosheng Lin, Yazhou Yao, Fayao Liu, Zichuan Liu, and Zhenmin Tang · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Cross-modal causal relational reasoning for event-level visual question answering
Yang Liu, Guanbin Li, and Liang Lin · 2023
Cited alongside, same era.
Knowledge-based embodied question answering
Sinan Tan, Mengmeng Ge, Di Guo, Huaping Liu, and Fuchun Sun · 2023
Cited alongside, same era.
Tidybot: Personalized robot assistance with large language models
Jimmy Wu, Rika Antonova, Adam Kan, Marion Lepert, Andy Zeng, Shuran Song, Jeannette Bohg, Szymon Rusinkiewicz, and Thomas Funkhouser · 2023
Cited alongside, same era.
Efficienteqa: An efficient approach for open vocabulary embodied question answering
Kai Cheng, Zhengyuan Li, Xingpeng Sun, Byung-Cheol Min, Amrit Singh Bedi, and Aniket Bera · 2024
Openeqa: Embodied question answering in the era of foundation models
Arjun Majumdar, Anurag Ajay, Xiaohan Zhang, Pranav Putta, Sriram Yenamandra, Mikael Henaff, Sneha Silwal, Paul Mcvay, Oleksandr Maksymets, Sergio Arnaud, et al · 2024
Later among the works it cites.
Explore until confident: Efficient exploration for embodied question answering
Allen Z Ren, Jaden Clark, Anushri Dixit, Masha Itkina, Anirudha Majumdar, and Dorsa Sadigh · 2024
Later among the works it cites.
Map-based modular approach for zero-shot embodied question answering
Koya Sakamoto, Daichi Azuma, Taiki Miyanishi, Shuhei Kurita, and Motoaki Kawanabe · 2024
Later among the works it cites.
Grapheqa: Using 3d semantic scene graphs for real-time embodied question answering
Saumya Saxena, Blake Buchanan, Chris Paxton, Bingqing Chen, Narunas Vaskevicius, Luigi Palmieri, Jonathan Francis, and Oliver Kroemer · 2024
Later among the works it cites.
An improved frontier-based robot exploration strategy combined with deep reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
S-eqa: Tackling situational queries in embodied question answering
Vishnu Sashank Dorbala, Prasoon Goyal, Robinson Piramuthu, Michael Johnston, Reza Ghanadhan, and Dinesh Manocha · 2024
Cited alongside, same era.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al · 2024
Cited alongside, same era.
Aaron Hurst, Adam Lerer, Adam P Goucher, Adam Perelman, Aditya Ramesh, Aidan Clark, AJ Ostrow, Akila Welihinda, Alan Hayes, Alec Radford, et al · 2024
Cited alongside, same era.
From image to language: A critical analysis of visual question answering (vqa) approaches, challenges, and opportunities
Md Farhan Ishmam, Md Sakib Hossain Shovon, Muhammad Firoz Mridha, and Nilanjan Dey · 2024
Cited alongside, same era.
Hal-eval: A universal and fine-grained hallucination evaluation framework for large vision language models
Chaoya Jiang, Hongrui Jia, Mengfan Dong, Wei Ye, Haiyang Xu, Ming Yan, Ji Zhang, and Shikun Zhang · 2024
Cited alongside, same era.
Dream2real: Zero-shot 3d object rearrangement with vision-language models
Ivan Kapelyukh, Yifei Ren, Ignacio Alzugaray, and Edward Johns · 2024
Cited alongside, same era.
Prismatic vlms: Investigating the design space of visually-conditioned language models
Siddharth Karamcheti, Suraj Nair, Ashwin Balakrishna, Percy Liang, Thomas Kollar, and Dorsa Sadigh · 2024
Cited alongside, same era.
Rui Wang, Jie Zhang, Ming Lyu, Cheng Yan, and Yaowei Chen · 2024
Later among the works it cites.
Noisyeqa: Benchmarking embodied question answering against noisy queries
Tao Wu, Chuhao Zhou, Yen Heng Wong, Lin Gu, and Jianfei Yang · 2024
Later among the works it cites.
Rt-grasp: Reasoning tuning robotic grasping via multi-modal large language model
Jinxuan Xu, Shiyu Jin, Yutian Lei, Yuqian Zhang, and Liangjun Zhang · 2024
Later among the works it cites.
Tree-of-reasoning question decomposition for complex question answering with large language models
Kun Zhang, Jiali Zeng, Fandong Meng, Yuanzhuo Wang, Shiqi Sun, Long Bai, Huawei Shen, and Jie Zhou · 2024
Later among the works it cites.
Towards learning a generalist model for embodied navigation
Duo Zheng, Shijia Huang, Lin Zhao, Yiwu Zhong, and Liwei Wang · 2024
Later among the works it cites.
Cross-modal causal relation alignment for video question grounding
Weixing Chen, Yang Liu, Binglin Chen, Jiandong Su, Yongsen Zheng, and Liang Lin · 2025
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al · 2025
Closest in time.
Dspnet: Dual-vision scene perception for robust 3d question answering
Jingzhou Luo, Yang Liu, Weixing Chen, Zhen Li, Yaowei Wang, Guanbin Li, and Liang Lin · 2025
Closest in time.
Towards long-horizon vision-language navigation: Platform, benchmark and method
Xinshuai Song, Weixing Chen, Yang Liu, Weikai Chen, Guanbin Li, and Liang Lin · 2025
Closest in time.
Cityeqa: A hierarchical llm agent on embodied question answering benchmark in city space
Yong Zhao, Kai Xu, Zhengqiu Zhu, Yue Hu, Zhiheng Zheng, Yingfeng Chen, Yatai Ji, Chen Gao, Yong Li, and Jincai Huang · 2025
Closest in time.