Fetching the paper…
Reading the bibliography…
Recent advances in Large Language Models (LLMs) have shown impressive capabilities in various applications, yet LLMs face challenges such as limited context windows and difficulties in generalization.
“The magical number seven plus or minus two: some limits on our capacity for processing information”, 1956
George. Miller · 1956
Earlier work this paper cites.
“Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks”, 2021
Patrick Lewis et al · 2005
Earlier work this paper cites.
“Conjunctive representation of position, direction, and velocity in entorhinal cortex”
Francesca Sargolini et al · 2006
Earlier work this paper cites.
“Spatial memory: how egocentric and allocentric combine” Epub 2006 Oct 30
Neil Burgess · 2006
Earlier work this paper cites.
“Remembering the past and imagining the future: A neural model of spatial memory and imagery”
Peter Byrne, Suzanna Becker and Neil Burgess · 2007
Earlier work this paper cites.
“Thinking, fast and slow”
Daniel Kahneman · 2011
Earlier work this paper cites.
“Attention Is All You Need”, 2017
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“Vector-based navigation using grid-like representations in artificial agents”
Andrea Banino et al · 2018
Earlier work this paper cites.
“Flexible coding of object motion in multiple reference frames by parietal cortex neurons”
R. Sasaki, A. Anzai and D.. Angelaki · 2020
Earlier work this paper cites.
“Assured Learning-enabled Autonomy: A Metacognitive Reinforcement Learning Framework”, 2021
Aquib Mustafa, Majid Mazouchi, Subramanya Nageshrao and Hamidreza Modares · 2021
Earlier work this paper cites.
“Grid Cell Path Integration For Movement-Based Visual Object Recognition”, 2021
Niels Leadholm, Marcus Lewis and Subutai Ahmad · 2021
Earlier work this paper cites.
“Computational Metacognition”, 2022
Michael Cox et al · 2022
Cited alongside, same era.
“Enhancing metacognitive reinforcement learning using reward structures and feedback” These authors contributed equally
Paul. Krueger, Falk Lieder and Thomas. Griffiths · 2022
Cited alongside, same era.
“Grand unified theory of mind and brain-part i: Space-time approach to dynamic connectomes of c. elegans and human brains by mepmos”, 2022
Katsushi Arisaka · 2022
Cited alongside, same era.
“Grid cells and their potential application in AI”, 2022
Jason Toy · 2022
Cited alongside, same era.
“Textbooks Are All You Need II: phi-1.5 technical report”, 2023
Yuanzhi Li et al · 2023
Later among the works it cites.
“Phi-2: The surprising power of small language models” https://www.microsoft.com/en-us/research/blog/phi-2-the-surprising-power-of-small-language-models/
Marah Abdin et al · 2023
Later among the works it cites.
“Llama 2: Open Foundation and Fine-Tuned Chat Models”, 2023
Hugo Touvron et al · 2023
Later among the works it cites.
“GPT-4 Technical Report”, 2023
OpenAI et al · 2023
Later among the works it cites.
“Improved Baselines with Visual Instruction Tuning”
Haotian Liu, Chunyuan Li, Yuheng Li and Yong Lee · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
James.. Whittington, Joseph Warren and Tim.J. Behrens · 2022
Cited alongside, same era.
“Communicative Agents for Software Development”, 2023
Chen Qian et al · 2023
Cited alongside, same era.
“Generative Agents: Interactive Simulacra of Human Behavior”, 2023
Joon Park et al · 2023
Cited alongside, same era.
“Breaking Bad: Unraveling Influences and Risks of User Inputs to ChatGPT for Game Story Generation”
Pittawat Taveekitworachai et al · 2023
Cited alongside, same era.
Albert. Jiang et al · 2023
Cited alongside, same era.
“Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena”, 2023
Lianmin Zheng et al · 2023
Cited alongside, same era.
Later among the works it cites.
“SILO Language Models: Isolating Legal Risk In a Nonparametric Datastore”, 2023
Sewon Min et al · 2023
Later among the works it cites.
“Efficient Memory Management for Large Language Model Serving with PagedAttention”
Woosuk Kwon et al · 2023
Later among the works it cites.
“Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing” under review
Anonymous · 2023
Later among the works it cites.
Albert. Jiang et al · 2024
Closest in time.
“TinyLlama: An Open-Source Small Language Model”, 2024
Peiyuan Zhang, Guangtao Zeng, Tianduo Wang and Wei Lu · 2024
Closest in time.