Fetching the paper…
Reading the bibliography…
Although Large Language Models (LLMs) are effective in performing various NLP tasks, they still struggle to handle tasks that require extensive, real-world knowledge, especially when dealing with long-tail facts (facts related to long-tail entities).
From TreeBank to PropBank,
P. Kingsbury, M. Palmer, · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation,
K. Papineni, S. Roukos, T. Ward, W.-J. Zhu, · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries,
C.-Y. Lin, · 2004
Earlier work this paper cites.
TAGME: on-the-fly annotation of short text fragments (by wikipedia entities),
P. Ferragina, U. Scaiella, · 2010
Earlier work this paper cites.
Abstract Meaning Representation for sembanking,
L. Banarescu, C. Bonial, S. Cai, M. Georgescu, K. Griffitt, U. Hermjakob, K. Knight, P. Koehn, M. Palmer, N. Schneider, · 2013
Earlier work this paper cites.
Wikidata: a free collaborative knowledgebase,
D. Vrandecic, M. Krötzsch, · 2014
Earlier work this paper cites.
MS MARCO: A human generated machine reading comprehension dataset,
T. Nguyen, M. Rosenberg, X. Song, J. Gao, S. Tiwary, R. Majumder, L. Deng, · 2016
Earlier work this paper cites.
TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension,
M. Joshi, E. Choi, D. Weld, L. Zettlemoyer, · 2017
Earlier work this paper cites.
P. Velickovic, G. Cucurull, A. Casanova, A. Romero, P. Liò, Y. Bengio, · 2017
Earlier work this paper cites.
T-REx: A large scale alignment of natural language with knowledge base triples,
H. Elsahar, P. Vougiouklis, A. Remaci, C. Gravier, J. Hare, F. Laforest, E. Simperl, · 2018
Earlier work this paper cites.
Language models as knowledge bases?,
F. Petroni, T. Rocktäschel, S. Riedel, P. Lewis, A. Bakhtin, Y. Wu, A. Miller, · 2019
Earlier work this paper cites.
Natural questions: A benchmark for question answering research,
T. Kwiatkowski, J. Palomaki, O. Redfield, M. Collins, A. Parikh, C. Alberti, D. Epstein, I. Polosukhin, J. Devlin, K. Lee, K. Toutanova, L. Jones, M. Kelcey, M.-W. Chang, A. M. Dai, J. Uszkoreit, Q. Le, S. Petrov, · 2019
Earlier work this paper cites.
Wizard of wikipedia: Knowledge-powered conversational agents,
E. Dinan, S. Roller, K. Shuster, A. Fan, M. Auli, J. Weston, · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding,
J. Devlin, M.-W. Chang, K. Lee, K. Toutanova, · 2019
Earlier work this paper cites.
Scalable zero-shot entity linking with dense entity retrieval,
L. Wu, F. Petroni, M. Josifoski, S. Riedel, L. Zettlemoyer, · 2020
Earlier work this paper cites.
KILT: a benchmark for knowledge intensive language tasks,
F. Petroni, A. Piktus, A. Fan, P. Lewis, M. Yazdani, N. De Cao, J. Thorne, Y. Jernite, V. Karpukhin, J. Maillard, V. Plachouras, T. Rocktäschel, S. Riedel, · 2021
Cited alongside, same era.
A survey on complex knowledge base question answering: Methods, challenges and solutions,
Y. Lan, G. He, J. Jiang, J. Jiang, W. X. Zhao, J. Wen, · 2021
Cited alongside, same era.
Leveraging Abstract Meaning Representation for knowledge base question answering,
P. Kapanipathi, I. Abdelaziz, S. Ravishankar, S. Roukos, A. Gray, R. Fernandez Astudillo, M. Chang, C. Cornelio, S. Dana, A. Fokoue, D. Garg, A. Gliozzo, S. Gurajada, H. Karanam, N. Khan, D. Khandelwal, Y.-S. Lee, Y. Li, F. Luus, N. Makondo, N. Mihindukulasooriya, T. Naseem, S. Neelam, L. Popa, R. Gangi Reddy, R. Riegel, G. Rossiello, U. Sharma, G. P. S. Bhargav, M. Yu, · 2021
Cited alongside, same era.
Retrieval, re-ranking and multi-task learning for knowledge-base question answering,
Z. Wang, P. Ng, R. Nallapati, B. Xiang, · 2021
Cited alongside, same era.
A semantics-aware transformer model of relation linking for knowledge base question answering,
T. Naseem, S. Ravishankar, N. Mihindukulasooriya, I. Abdelaziz, Y.-S. Lee, P. Kapanipathi, S. Roukos, A. Gliozzo, A. Gray, · 2021
MENLI: robust evaluation metrics from natural language inference,
Y. Chen, S. Eger, · 2022
Later among the works it cites.
OpenAI, · 2023
Later among the works it cites.
R. Anil, A. M. Dai, O. Firat, M. Johnson, D. Lepikhin, A. Passos, S. Shakeri, E. Taropa, P. Bailey, Z. Chen, E. Chu, J. H. Clark, L. E. Shafey, Y. Huang, K. Meier-Hellstern, G. Mishra, E. Moreira, M. Omernick, K. Robinson, S. Ruder, Y. Tay, K. Xiao, Y. Xu, Y. Zhang, G. H. Ábrego, J. Ahn, J. Austin, P. Barham, J. Botha, J. Bradbury, S. Brahma, K. Brooks, M. Catasta, Y. Cheng, C. Cherry, C. A. Choquette-Choo, A. Chowdhery, C. Crepy, S. Dave, M. Dehghani, S. Dev, J. Devlin, M. Díaz, N. Du, E. Dyer, V. Feinberg, F. Feng, V. Fienber, M. Freitag, X. Garcia, S. Gehrmann, L. Gonzalez, et al., · 2023
Later among the works it cites.
When not to trust language models: Investigating effectiveness of parametric and non-parametric memories,
A. Mallen, A. Asai, V. Zhong, R. Das, D. Khashabi, H. Hajishirzi, · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
One SPRING to rule them both: Symmetric AMR semantic parsing and generation without a complex pipeline,
M. Bevilacqua, R. Blloshmi, R. Navigli, · 2021
Cited alongside, same era.
Can NLI models verify QA systems’ predictions?,
J. Chen, E. Choi, G. Durrett, · 2021
Cited alongside, same era.
P. He, J. Gao, W. Chen, · 2021
Cited alongside, same era.
KQA pro: A dataset with explicit compositional programs for complex question answering over knowledge base,
S. Cao, J. Shi, L. Pan, L. Nie, Y. Xiang, L. Hou, J. Li, B. He, H. Zhang, · 2022
Cited alongside, same era.
Conversational question answering: a survey.,
M. Zaib, W. E. Zhang, Q. Z. Sheng, A. Mahmood, Y. Zhang, · 2022
Cited alongside, same era.
A survey on complex factual question answering,
L. Zhang, J. Zhang, X. Ke, H. Li, X. Huang, Z. Shao, S. Cao, X. Lv, · 2022
Cited alongside, same era.
Re2G: Retrieve, rerank, generate,
M. Glass, G. Rossiello, M. F. M. Chowdhury, A. Naik, P. Cai, A. Gliozzo, · 2022
Cited alongside, same era.
B. Peng, M. Galley, P. He, H. Cheng, Y. Xie, Y. Hu, Q. Huang, L. Liden, Z. Yu, W. Chen, J. Gao, · 2023
Later among the works it cites.
Z. Hu, V. Gutiérrez-Basulto, Z. Xiang, R. Li, J. Z. Pan, · 2023
Later among the works it cites.
Semantic parsing for conversational question answering over knowledge graphs,
L. Perez-Beltrachini, P. Jain, E. Monti, M. Lapata, · 2023
Later among the works it cites.
FastRAT: Fast and Efficient Cross-lingual Text-to-SQL Semantic Parsing,
P. Vougiouklis, N. Papasarantopoulos, D. Zheng, D. Tuckey, C. Diao, Z. Shen, J. Z. Pan, · 2023
Later among the works it cites.
Large Language Models and Knowledge Graphs: Opportunities and Challenges,
J. Z. Pan, S. Razniewski, J.-C. Kalo, S. Singhania, J. Chen, S. Dietze, H. Jabeen, J. Omeliyanenko, W. Zhang, M. Lissandrini, R. Biswas, G. de Melo, A. Bonifati, E. Vakaj, M. Dragoni, , D. Graux, · 2023
Later among the works it cites.
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, T. B. Hashimoto, Stanford alpaca: An instruction-following llama model, https://github.com/tatsu-lab/stanford_alpaca , 2023
2023
Later among the works it cites.
Instruction tuning with GPT-4,
B. Peng, C. Li, P. He, M. Galley, J. Gao, · 2023
Later among the works it cites.
Self-rag: Learning to retrieve, generate, and critique through self-reflection,
A. Asai, Z. Wu, Y. Wang, A. Sil, H. Hajishirzi, · 2023
Later among the works it cites.
Efficient memory management for large language model serving with pagedattention,
W. Kwon, Z. Li, S. Zhuang, Y. Sheng, L. Zheng, C. H. Yu, J. Gonzalez, H. Zhang, I. Stoica, · 2023
Later among the works it cites.
Knowledge-augmented language model verification,
J. Baek, S. Jeong, M. Kang, J. Park, S. Hwang, · 2023
Later among the works it cites.
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning,
D. Zheng, M. Lapata, J. Z. Pan, · 2024
Closest in time.