Fetching the paper…
Reading the bibliography…
A wide variety of agentic AI applications - ranging from cognitive assistants for dementia patients to robotics - demand a robust memory system grounded in reality.
Critique of pure reason. 1781,
I. Kant, · 1908
Earlier work this paper cites.
The anatomy of a large-scale hypertextual web search engine,
S. Bin, K. L. Page, · 1998
Earlier work this paper cites.
3D scene graph: A structure for unified semantics, 3D space, and camera,
I. Armeni, Z.-Y. He, J. Gwak, A. R. Zamir, M. Fischer, J. Malik, S. Savarese, · 2019
Earlier work this paper cites.
3D scene graph: A sparse and semantic representation of physical environments for intelligent agents,
U.-H. Kim, J.-M. Park, T.-J. Song, J.-H. Kim, · 2019
Earlier work this paper cites.
Action representation for intelligent agents using Memory Nets,
J. Eggert, J. Deigmöller, L. Fischer, A. Richter, · 2020
Earlier work this paper cites.
3D dynamic scene graphs: Actionable spatial perception with places, objects, and humans,
A. Rosinol, A. Gupta, M. Abate, J. Shi, L. Carlone, · 2020
Earlier work this paper cites.
Knowledge graphs,
A. Hogan, E. Blomqvist, M. Cochez, C. d’Amato, G. D. Melo, C. Gutierrez, S. Kirrane, J. E. L. Gayo, R. Navigli, S. Neumaier, et al., · 2021
Earlier work this paper cites.
Taskography: Evaluating robot task planning over large 3D scene graphs,
C. Agia, K. M. Jatavallabhula, M. Khodeir, O. Miksik, V. Vineet, M. Mukadam, L. Paull, F. Shkurti, · 2022
Earlier work this paper cites.
LifelongMemory: Leveraging LLMs for answering queries in long-form egocentric videos,
Y. Wang, Y. Yang, M. Ren, · 2023
Cited alongside, same era.
S. Kashmira, J. L. Dantanarayana, J. Brodsky, A. Mahendra, Y. Kang, K. Flautner, L. Tang, J. Mars, · 2024
Cited alongside, same era.
A review of visual slam for robotics: Evolution, properties, and future applications,
B. Al-Tawil, T. Hempel, A. Abdelrahman, A. Al-Hamadi, · 2024
Cited alongside, same era.
Embodied-RAG: General non-parametric embodied memory for retrieval and generation,
Q. Xie, S. Y. Min, T. Zhang, K. Xu, A. Bajaj, R. Salakhutdinov, M. Johnson-Roberson, Y. Bisk, · 2024
Cited alongside, same era.
VideoAgent: A memory-augmented multimodal agent for video understanding,
Text2cypher: Bridging natural language and graph databases,
M. G. Ozsoy, L. Messallem, J. Besga, G. Minneci, · 2024
Later among the works it cites.
From local to global: A graph RAG approach to query-focused summarization,
D. Edge, H. Trinh, N. Cheng, J. Bradley, A. Chao, A. Mody, S. Truitt, J. Larson, · 2024
Later among the works it cites.
J. Achiam, et al., · 2024
Later among the works it cites.
Tulip agent–enabling LLM-based agents to solve tasks using large tool libraries,
F. Ocker, D. Tanneberg, J. Eggert, M. Gienger, · 2024
Later among the works it cites.
Visual large language models for generalized and specialized applications,
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Fan, X. Ma, R. Wu, Y. Du, J. Li, Z. Gao, Q. Li, · 2024
Cited alongside, same era.
Amego: Active memory from long egocentric videos,
G. Goletto, T. Nagarajan, G. Averta, D. Damen, · 2024
Cited alongside, same era.
Omniquery: Contextually augmenting captured multimodal memory to enable personal question answering,
J. N. Li, Z. J. Zhang, J. Ma, · 2024
Cited alongside, same era.
HippoRAG: Neurobiologically inspired long-term memory for large language models,
B. J. Gutiérrez, Y. Shu, Y. Gu, M. Yasunaga, Y. Su, · 2024
Cited alongside, same era.
Y. Li, Z. Lai, W. Bao, Z. Tan, A. Dao, K. Sui, J. Shen, D. Liu, H. Liu, Y. Kong, · 2025
Closest in time.
J. Eggert, F. Ocker, Graph based memory extension for large language models, 2025. US Patent App. 18/898,607
2025
Closest in time.
MemPal: Leveraging multimodal AI and LLMs for voice-activated object retrieval in homes of older adults,
N. Maniar, S. W. Chan, W. Zulfikar, S. Ren, C. Xu, P. Maes, · 2025
Closest in time.