Fetching the paper…
Reading the bibliography…
We present and tackle the problem of Embodied Question Answering (EQA) with Situational Queries (S-EQA) in a household environment.
D. Müllner, “Modern hierarchical, agglomerative clustering algorithms,”
2011
Earlier work this paper cites.
X. Puig, K. Ra, M. Boben, J. Li, T. Wang, S. Fidler, and A. Torralba, “Virtualhome: Simulating household activities via programs,” in
2018
Earlier work this paper cites.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied question answering,” in
2018
Earlier work this paper cites.
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi, “Iqa: Visual question answering in interactive environments,” in
2018
Earlier work this paper cites.
L. Yu, X. Chen, G. Gkioxari, M. Bansal, T. L. Berg, and D. Batra, “Multi-target embodied question answering,” in
2019
Earlier work this paper cites.
Y. Li, A. X. Feng, J. Li, S. Mumick, A. Halevy, V. Li, and W.-C. Tan, “Subjective databases,”
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
OpenAI, “Language models are unsupervised multitask learners,”
2020
Earlier work this paper cites.
D. Gao, R. Wang, Z. Bai, and X. Chen, “Env-qa: A video question answering benchmark for comprehensive understanding of dynamic environments,” in
2021
Earlier work this paper cites.
W. Kim, B. Son, and I. Kim, “Vilt: Vision-and-language transformer without convolution or region supervision,” in
2021
Earlier work this paper cites.
J. Duan, S. Yu, H. L. Tan, H. Zhu, and C. Tan, “A survey of embodied ai: From simulators to research tasks,”
2022
Cited alongside, same era.
2022
Cited alongside, same era.
J. Huang and K. C.-C. Chang, “Towards reasoning in large language models: A survey,”
2022
Cited alongside, same era.
2022
Cited alongside, same era.
P. Cascante-Bonilla, H. Wu, L. Wang, R. S. Feris, and V. Ordonez, “Simvqa: Exploring simulated environments for visual question answering,” in
Q. Sima, S. Tan, H. Liu, F. Sun, W. Xu, and L. Fu, “Embodied referring expression for manipulation question answering in interactive environment,” in
2023
Later among the works it cites.
——, “Gpt-4 technical report,” 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
J. Li, D. Li, C. Xiong, and S. Hoi, “Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,” in
2022
Cited alongside, same era.
2022
Cited alongside, same era.
H. Luo, G. Lin, F. Shen, X. Huang, Y. Yao, and H. Shen, “Robust-eqa: robust learning for embodied question answering with noisy labels,”
2023
Cited alongside, same era.
S. Tan, M. Ge, D. Guo, H. Liu, and F. Sun, “Knowledge-based embodied question answering,”
2023
Cited alongside, same era.
“Amazon Mechanical Turk — mturk.com,”
Cited in the paper.
2023
Later among the works it cites.
D. Shah, B. Osiński, S. Levine
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Mu, Q. Zhang, M. Hu, W. Wang, M. Ding, J. Jin, B. Wang, J. Dai, Y. Qiao, and P. Luo, “Embodiedgpt: Vision-language pre-training via embodied chain of thought,”
2023
Later among the works it cites.