Fetching the paper…
Reading the bibliography…
In this paper, we propose a novel Knowledge-based Embodied Question Answering (K-EQA) task, in which the agent intelligently explores the environment to answer various questions with the knowledge.
B. PostgreSQL, “Postgresql,”
1996
Earlier work this paper cites.
R. Coulom, “Efficient selectivity and backup operators in monte-carlo tree search,” in
2006
Earlier work this paper cites.
2014
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. Lawrence Zitnick, and D. Parikh, “Vqa: Visual question answering,” in
2015
Earlier work this paper cites.
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein, “Neural module networks,” in
2016
Earlier work this paper cites.
C. Xiong, S. Merity, and R. Socher, “Dynamic memory networks for visual and textual question answering,” in
2016
Earlier work this paper cites.
R. Speer, J. Chin, and C. Havasi, “Conceptnet 5.5: An open multilingual graph of general knowledge,”
2016
Earlier work this paper cites.
L. Dong and M. Lapata, “Language to logical form with neural attention,” in
2016
Earlier work this paper cites.
Q. Wu, C. Shen, P. Wang, A. Dick, and A. Van Den Hengel, “Image captioning and visual question answering based on attributes and external knowledge,”
2017
Earlier work this paper cites.
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra, “Visual dialog,” in
2017
Earlier work this paper cites.
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. Lawrence Zitnick, and R. Girshick, “Clevr: A diagnostic dataset for compositional language and elementary visual reasoning,” in
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask r-cnn,” in
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi, “Iqa: Visual question answering in interactive environments,” in
2018
Cited alongside, same era.
P. Wang, Q. Wu, C. Shen, A. Dick, and A. van den Hengel, “Fvqa: Fact-based visual question answering,”
2018
Cited alongside, same era.
2018
Cited alongside, same era.
F. Liu, T. Xiang, T. M. Hospedales, W. Yang, and C. Sun, “Inverse visual question answering: A new benchmark and vqa diagnosis tool,”
2018
C. Chen, U. Jain, C. Schissler, S. V. A. Gari, Z. Al-Halah, V. K. Ithapu, P. Robinson, and K. Grauman, “Audio-visual embodied navigation,”
2019
Later among the works it cites.
H. Luo, G. Lin, Z. Liu, F. Liu, Z. Tang, and Y. Yao, “Segeqa: Video segmentation based visual attention for embodied question answering,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
U. Jain, L. Weihs, E. Kolve, M. Rastegari, S. Lazebnik, A. Farhadi, A. G. Schwing, and A. Kembhavi, “Two body problem: Collaborative visual task completion,” in
2019
Later among the works it cites.
A. Das, T. Gervet, J. Romoff, D. Batra, D. Parikh, M. Rabbat, and J. Pineau, “Tarmac: Targeted multi-agent communication,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
R. Hu, J. Andreas, T. Darrell, and K. Saenko, “Explainable neural computation via stack neural module networks,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Cited alongside, same era.
L. Yu, X. Chen, G. Gkioxari, M. Bansal, T. L. Berg, and D. Batra, “Multi-target embodied question answering,” in
2019
Cited alongside, same era.
K. Marino, M. Rastegari, A. Farhadi, and R. Mottaghi, “Ok-vqa: A visual question answering benchmark requiring external knowledge,” in
2019
Cited alongside, same era.
2019
Later among the works it cites.
J. Liang, L. Jiang, L. Cao, Y. Kalantidis, L.-J. Li, and A. G. Hauptmann, “Focal visual-text attention for memex question answering,”
2019
Later among the works it cites.
2019
Later among the works it cites.
D. A. Hudson and C. D. Manning, “Gqa: A new dataset for real-world visual reasoning and compositional question answering,” in
2019
Later among the works it cites.
R. Vedantam, K. Desai, S. Lee, M. Rohrbach, D. Batra, and D. Parikh, “Probabilistic neural symbolic models for interpretable visual question answering,” in
2019
Later among the works it cites.
Q. Cao, X. Liang, B. Li, and L. Lin, “Interpretable visual question answering by reasoning on dependency trees,”
2019
Later among the works it cites.
I. Armeni, Z.-Y. He, J. Gwak, A. R. Zamir, M. Fischer, J. Malik, and S. Savarese, “3d scene graph: A structure for unified semantics, 3d space, and camera,” in
2019
Later among the works it cites.
F. Sukkar, G. Best, C. Yoo, and R. Fitch, “Multi-robot region-of-interest reconstruction with dec-mcts,” in
2019
Later among the works it cites.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied question answering,” in
2063
Closest in time.