Fetching the paper…
Reading the bibliography…
Large language models (LLMs) suffer from temporal misalignment issues especially across long span of time.
On the mathematical foundations of theoretical statistics
Ronald A Fisher. 1922 · 1922
Earlier work this paper cites.
Back to the future: Towards explainable temporal reasoning with large language models
Chenhan Yuan, Qianqian Xie, Jimin Huang, and Sophia Ananiadou. 2024 · 1974
Earlier work this paper cites.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Sutime: A library for recognizing and normalizing time expressions
Angel X Chang and Christopher D Manning. 2012 · 2012
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
A Waswani, N Shazeer, N Parmar, J Uszkoreit, L Jones, A Gomez, L Kaiser, and I Polosukhin. 2017 · 2017
Earlier work this paper cites.
Diachronic degradation of language models: Insights from social media
Kokil Jaidka, Niyati Chhaya, and Lyle Ungar. 2018 · 2018
Earlier work this paper cites.
Knowledge neurons in pretrained transformers
Damai Dai, Li Dong, Yaru Hao, Zhifang Sui, Baobao Chang, and Furu Wei. 2021 · 2021
Earlier work this paper cites.
Editing factual knowledge in language models
Nicola De Cao, Wilker Aziz, and Ivan Titov. 2021 · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Earlier work this paper cites.
Mind the gap: Assessing temporal generalization in neural language models
Angeliki Lazaridou, Adhi Kuncoro, Elena Gribovskaya, Devang Agrawal, Adam Liska, Tayfun Terzi, Mai Gimenez, Cyprien de Masson d’Autume, Tomas Kocisky, Sebastian Ruder, et al. 2021 · 2021
Earlier work this paper cites.
Eric Mitchell, Charles Lin, Antoine Bosselut, Chelsea Finn, and Christopher D Manning. 2021 · 2021
Earlier work this paper cites.
Temporal effects on pre-trained models for language processing tasks
Oshin Agarwal and Ani Nenkova. 2022 · 2022
Cited alongside, same era.
Time-aware language models as temporal knowledge bases
Bhuwan Dhingra, Jeremy R Cole, Julian Martin Eisenschlos, Daniel Gillick, Jacob Eisenstein, and William W Cohen. 2022 · 2022
Cited alongside, same era.
Temporalwiki: A lifelong benchmark for training and evaluating ever-evolving language models
Joel Jang, Seonghyeon Ye, Changho Lee, Sohee Yang, Joongbo Shin, Janghoon Han, Gyeonghun Kim, and Minjoon Seo. 2022 · 2022
Cited alongside, same era.
Timelms: Diachronic language models from twitter
Daniel Loureiro, Francesco Barbieri, Leonardo Neves, Luis Espinosa Anke, and Jose Camacho-Collados. 2022 · 2022
Cited alongside, same era.
Time waits for no one! analysis and challenges of temporal misalignment
Kelvin Luu, Daniel Khashabi, Suchin Gururangan, Karishma Mandyam, and Noah A Smith. 2022 · 2022
Time is encoded in the weights of finetuned language models
Kai Nylund, Suchin Gururangan, and Noah A Smith. 2023 · 2023
Later among the works it cites.
Towards benchmarking and improving the temporal reasoning capability of large language models
Qingyu Tan, Hwee Tou Ng, and Lidong Bing. 2023 · 2023
Later among the works it cites.
Tram: Benchmarking temporal reasoning for large language models
Yuqing Wang and Yun Zhao. 2023 · 2023
Later among the works it cites.
Once upon a time in graph: Relative-time pretraining for complex temporal reasoning
Sen Yang, Xin Li, Lidong Bing, and Wai Lam. 2023 · 2023
Later among the works it cites.
Mitigating temporal misalignment by discarding outdated facts
Michael JQ Zhang and Eunsol Choi. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mass-editing memory in a transformer
Kevin Meng, Arnab Sen Sharma, Alex Andonian, Yonatan Belinkov, and David Bau. 2022 · 2022
Cited alongside, same era.
Impact of pretraining term frequencies on few-shot numerical reasoning
Razeghi Yasaman, Robert Logan IV, Gardner Matt, and Singh Sameer. 2022 · 2022
Cited alongside, same era.
Modifying memories in transformer models
Chen Zhu, Ankit Singh Rawat, Manzil Zaheer, Srinadh Bhojanapalli, Daliang Li, Felix Yu, and Sanjiv Kumar. 2020 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Cited alongside, same era.
Language models represent space and time
Wes Gurnee and Max Tegmark. 2023 · 2023
Cited alongside, same era.
Dynamic benchmarking of masked language models on temporal concept drift with multiple views
Katerina Margatina, Shuai Wang, Yogarshi Vyas, Neha Anna John, Yassine Benajiba, and Miguel Ballesteros. 2023 · 2023
Cited alongside, same era.
R Thomas McCoy, Shunyu Yao, Dan Friedman, Matthew Hardy, and Thomas L Griffiths. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
Remember this event that year? assessing temporal information and reasoning in large language models
Himanshu Beniwal, Dishant Patel, Hritik Ladia, Ankit Yadav, Mayank Singh, et al. 2024 · 2024
Later among the works it cites.
Chew: A dataset of changing events in wikipedia
Hsuvas Borkakoty and Luis Espinosa-Anke. 2024 · 2024
Later among the works it cites.
Felix Drinkall, Eghbal Rahimikia, Janet B Pierrehumbert, and Stefan Zohren. 2024 · 2024
Later among the works it cites.
A pretrainer’s guide to training data: Measuring the effects of data age, domain coverage, quality, & toxicity
Shayne Longpre, Gregory Yauney, Emily Reif, Katherine Lee, Adam Roberts, Barret Zoph, Denny Zhou, Jason Wei, Kevin Robinson, David Mimno, et al. 2024 · 2024
Later among the works it cites.
An Yang, Baosong Yang, Beichen Zhang, Binyuan Hui, Bo Zheng, Bowen Yu, Chengyuan Li, Dayiheng Liu, Fei Huang, Haoran Wei, et al. 2024 · 2024
Later among the works it cites.
Set the clock: Temporal alignment of pretrained language models
Bowen Zhao, Zander Brumbaugh, Yizhong Wang, Hannaneh Hajishirzi, and Noah A Smith. 2024 · 2024
Later among the works it cites.
A diachronic language model for long-time span classical chinese
Yuting Wei, Meiling Li, Yangfu Zhu, Yuanxing Xu, Yuqing Li, and Bin Wu. 2025 · 2025
Closest in time.