Fetching the paper…
Reading the bibliography…
Recent proprietary large language models (LLMs), such as GPT-4, have achieved a milestone in tackling diverse challenges in the biomedical domain, ranging from multiple-choice questions to long-form generations.
Huggingface’s transformers: State-of-the-art natural language processing
Wolf, T. et al · 1910
Earlier work this paper cites.
Knn model-based approach in classification
Guo, G. et al · 2003
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. (2004) · 2004
Earlier work this paper cites.
Measuring massive multitask language understanding
Hendrycks, D. et al · 2009
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Bajaj, P. et al · 2016
Earlier work this paper cites.
Overview of the medical question answering task at trec 2017 liveqa
Abacha, A. B. et al · 2017
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Christiano, P. F. et al · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J. et al · 2017
Earlier work this paper cites.
Constituency parsing with a self-attentive encoder
Kitaev, N. and Klein, D. (2018) · 2018
Earlier work this paper cites.
Bridging the gap between consumers’ medication questions and trusted answers
Abacha, A. B. et al · 2019
Earlier work this paper cites.
Pubmedqa: A dataset for biomedical research question answering
Jin, Q. et al · 2019
Earlier work this paper cites.
Multilingual constituency parsing with self-attention and pre-training
Kitaev, N. et al · 2019
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A. et al · 2019
Earlier work this paper cites.
Multi-passage bert: A globally normalized bert model for open-domain question answering
Wang, Z. et al · 2019
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Zhang, T. et al · 2019
Earlier work this paper cites.
Retrieval augmented language model pre-training
Guu, K. et al · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Karpukhin, V. et al · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Lewis, P. et al · 2020
Cited alongside, same era.
Zero: Memory optimizations toward training trillion parameter models
Rajbhandari, S. et al · 2020
Cited alongside, same era.
What disease does this patient have? a large-scale open domain question answering dataset from medical exams
Jin, D. et al · 2021
Cited alongside, same era.
Generation-augmented retrieval for open-domain question answering
Mao, Y. et al · 2021
Cited alongside, same era.
Hallucinated but factual! inspecting the factuality of hallucinations in abstractive summarization
Cao, M. et al · 2022
Cited alongside, same era.
Meditron-70b: Scaling medical pretraining for large language models
Chen, Z. et al · 2023
Later among the works it cites.
Mol-instructions: A large-scale biomolecular instruction dataset for large language models
Fang, Y. et al · 2023
Later among the works it cites.
Medalpaca–an open-source collection of medical conversational ai models and training data
Han, T. et al · 2023
Later among the works it cites.
Survey of hallucination in natural language generation
Ji, Z. et al · 2023
Later among the works it cites.
Active retrieval augmented generation
Jiang, Z. et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scaling instruction-finetuned language models
Chung, H. W. et al · 2022
Cited alongside, same era.
Flashattention: Fast and memory-efficient exact attention with io-awareness
Dao, T. et al · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Ouyang, L. et al · 2022
Cited alongside, same era.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Pal, A. et al · 2022
Cited alongside, same era.
Large language models encode clinical knowledge
Singhal, K. et al · 2022
Cited alongside, same era.
Galactica: A large language model for science
Taylor, R. et al · 2022
Cited alongside, same era.
Medcpt: Contrastive pre-trained transformers with large-scale pubmed search logs for zero-shot biomedical information retrieval
Jin, Q. et al · 2023
Later among the works it cites.
Knowledge-augmented reasoning distillation for small language models in knowledge-intensive tasks
Kang, M. et al · 2023
Later among the works it cites.
Efficient memory management for large language model serving with pagedattention
Kwon, W. et al · 2023
Later among the works it cites.
Meddm: Llm-executable clinical guidance tree for clinical decision-making
Li, B. et al · 2023
Later among the works it cites.
Capabilities of gpt-4 on medical challenge problems
Nori, H. et al · 2023
Later among the works it cites.
Enhancing retrieval-augmented large language models with iterative retrieval-generation synergy
Shao, Z. et al · 2023
Later among the works it cites.
Alpaca: A strong, replicable instruction-following model
Taori, R. et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H. et al · 2023
Later among the works it cites.
Augmenting black-box llms with medical textbooks for clinical question answering
Wang, Y. et al · 2023
Later among the works it cites.
Alpacare: Instruction-tuned large language models for medical application
Zhang, X. et al · 2023
Later among the works it cites.