Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have transformed how people interact with artificial intelligence (AI) systems, achieving state-of-the-art results in various tasks, including scientific discovery and hypothesis generation.
Language models are few-shot learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Zhang, T.; Kishore, V.; Wu, F.; Weinberger, K. Q.; and Artzi, Y. 2019 · 1904
Earlier work this paper cites.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Zhao, W.; Peyrard, M.; Liu, F.; Gao, Y.; Meyer, C. M.; and Eger, S. 2019 · 1909
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
Automatic machine translation evaluation in many languages via zero-shot paraphrasing
Thompson, B.; and Post, M. 2020 · 2004
Earlier work this paper cites.
Bartscore: Evaluating generated text as text generation
Yuan, W.; Neubig, G.; and Liu, P. 2021 · 2021
Earlier work this paper cites.
Large language models are zero-shot reasoners
Kojima, T.; Gu, S. S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y. 2022 · 2022
Earlier work this paper cites.
The impact of large language models on scientific discovery: a preliminary study using gpt-4
AI4Science, M. R.; and Quantum, M. A. 2023 · 2023
Earlier work this paper cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
Bubeck, S.; Chandrasekaran, V.; Eldan, R.; Gehrke, J.; Horvitz, E.; Kamar, E.; Lee, P.; Lee, Y. T.; Li, Y.; Lundberg, S.; et al. 2023 · 2023
Cited alongside, same era.
Can large language models be an alternative to human evaluations?
Chiang, C.-H.; and Lee, H.-y. 2023 · 2023
Cited alongside, same era.
The semantic scholar open data platform
Kinney, R.; Anastasiades, C.; Authur, R.; Beltagy, I.; Bragg, J.; Buraczynski, A.; Cachola, I.; Candra, S.; Chandrasekhar, Y.; Cohan, A.; et al. 2023 · 2023
Cited alongside, same era.
G-eval: Nlg evaluation using gpt-4 with better human alignment
Liu, Y.; Iter, D.; Xu, Y.; Wang, S.; Xu, R.; and Zhu, C. 2023 · 2023
Cited alongside, same era.
Baek, J.; Jauhar, S. K.; Cucerzan, S.; and Hwang, S. J. 2024 · 2024
Closest in time.
Benchmarking foundation models with language-model-as-an-examiner
Bai, Y.; Ying, J.; Cao, Y.; Lv, X.; He, Y.; Wang, X.; Yu, J.; Zeng, K.; Xiao, Y.; Lyu, H.; et al. 2024 · 2024
Closest in time.
GPTScore: Evaluate as You Desire
Fu, J.; Ng, S.-K.; Jiang, Z.; and Liu, P. 2024 · 2024
Closest in time.
Google Scholar Top Publications
Google Scholar. 2024 · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reid, M.; Savinov, N.; Teplyashin, D.; Lepikhin, D.; Lillicrap, T.; Alayrac, J.-b.; Soricut, R.; Lazaridou, A.; Firat, O.; Schrittwieser, J.; et al. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Moore, S.; Tong, R.; Singh, A.; Liu, Z.; Hu, X.; Lu, Y.; Liang, J.; Cao, C.; Khosravi, H.; Denny, P.; et al. 2023 · 2023
Cited alongside, same era.
OpenAI. 2023 · 2023
Cited alongside, same era.
Qiu, L.; Jiang, L.; Lu, X.; Sclar, M.; Pyatkin, V.; Bhagavatula, C.; Wang, B.; Kim, Y.; Choi, Y.; Dziri, N.; et al. 2023 · 2023
Cited alongside, same era.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H.; Martin, L.; Stone, K.; Albert, P.; Almahairi, A.; Babaei, Y.; Bashlykov, N.; Batra, S.; Bhargava, P.; Bhosale, S.; et al. 2023 · 2023
Cited alongside, same era.
Learning to generate novel scientific directions with contextualized literature-based discovery
Wang, Q.; Downey, D.; Ji, H.; and Hope, T. 2023a
Cited in the paper.
Scimon: Scientific inspiration machines optimized for novelty
Wang, Q.; Downey, D.; Ji, H.; and Hope, T. 2023b
Cited in the paper.
Large language models in health care: Development, applications, and challenges
Yang, R.; Tan, T. F.; Lu, W.; Thirunavukarasu, A. J.; Ting, D. S. W.; and Liu, N. 2023a
Cited in the paper.
Large language models for automated open-domain scientific hypotheses discovery
Yang, Z.; Du, X.; Li, J.; Zheng, J.; Poria, S.; and Cambria, E. 2023b
Cited in the paper.
Vu, T.; Krishna, K.; Alzubi, S.; Tar, C.; Faruqui, M.; and Sung, Y.-H. 2024 · 2024
Closest in time.
An LLM-based Knowledge Synthesis and Scientific Reasoning Framework for Biomedical Discovery
Wysocki, O.; Wysocka, M.; Carvalho, D.; Bogatu, A. T.; Gusicuma, D. M.; Delmas, M.; Unsworth, H.; and Freitas, A. 2024 · 2024
Closest in time.
Hypothesis Generation with Large Language Models
Zhou, Y.; Liu, H.; Srivastava, T.; Mei, H.; and Tan, C. 2024 · 2024
Closest in time.