Fetching the paper…
Reading the bibliography…
Excellence in a wide variety of medical applications poses considerable challenges for AI, requiring advanced reasoning, access to up-to-date medical knowledge and understanding of complex multimodal data.
MIMIC-CXR-JPG, a large publicly available database of labeled chest radiographs
A. E. Johnson, T. J. Pollard, N. R. Greenbaum, M. P. Lungren, C.-y. Deng, Y. Peng, Z. Lu, R. G. Mark, S. J. Berkowitz, and S. Horng · 1901
Earlier work this paper cites.
Diagnostic strategies in the hypothesis-directed pathfinder system
E. Horvitz, D. Heckerman, B. N. Nathwani, and L. M. Fagan · 1984
Earlier work this paper cites.
When less is more: a practical approach to searching for evidence-based answers
K. K. Grandage, D. C. Slawson, and A. F. Shaughnessy · 2002
Earlier work this paper cites.
Causes and prevention of laparoscopic bile duct injuries: analysis of 252 cases from a human factors and cognitive psychology perspective
L. W. Way, L. Stewart, W. Gantert, K. Liu, C. M. Lee, K. Whang, and J. G. Hunter · 2003
Earlier work this paper cites.
The intersections of gender and class in health status and health care
A. Iyer, G. Sen, and P. Östlin · 2008
Earlier work this paper cites.
Rationale and use of the critical view of safety in laparoscopic cholecystectomy
S. M. Strasberg and M. L. Brunt · 2010
Earlier work this paper cites.
Challenges and opportunities facing medical education
P. Densen · 2011
Earlier work this paper cites.
Mining electronic health records: towards better research applications and clinical care
P. B. Jensen, L. J. Jensen, and S. Brunak · 2012
Earlier work this paper cites.
Gender disparities in health care
J. A. Kent, V. Patel, and N. A. Varela · 2012
Earlier work this paper cites.
Standards for reporting plain language summaries (pls) for cochrane diagnostic test accuracy reviews, 2014
Cochrane · 2014
Earlier work this paper cites.
A simple effective method for generation of a permanent record of the critical view of safety during laparoscopic cholecystectomy by intraoperative “doublet” photography
D. E. Sanford and S. M. Strasberg · 2014
Earlier work this paper cites.
Fto obesity variant circuitry and adipocyte browning in humans
M. Claussnitzer, S. N. Dankel, K.-H. Kim, G. Quon, W. Meuleman, C. Haugen, V. Glunk, I. S. Sousa, J. L. Beaudry, V. Puviindran, et al · 2015
Earlier work this paper cites.
Information overload in healthcare: too much of a good thing?
I. Klerings, A. S. Weinhandl, and K. J. Thaler · 2015
Earlier work this paper cites.
Racial bias in health care and health: challenges and opportunities
D. R. Williams and R. Wyatt · 2015
Earlier work this paper cites.
Extracting information from the text of electronic medical records to improve case detection: a systematic review
E. Ford, J. A. Carroll, H. E. Smith, D. Scott, and J. A. Cassell · 2016
Earlier work this paper cites.
MIMIC-III, a freely accessible critical care database
A. E. Johnson, T. J. Pollard, L. Shen, L.-w. H. Lehman, M. Feng, M. Ghassemi, B. Moody, P. Szolovits, L. Anthony Celi, and R. G. Mark · 2016
Earlier work this paper cites.
Endonet: a deep architecture for recognition tasks on laparoscopic videos
A. P. Twinanda, S. Shehata, D. Mutter, J. Marescaux, M. De Mathelin, and N. Padoy · 2016
Earlier work this paper cites.
Clinical reasoning: defining it, teaching it, assessing it, studying it
L. D. Gruppen · 2017
Earlier work this paper cites.
Health inequities, social determinants, and intersectionality
N. López and V. L. Gadsden · 2017
Earlier work this paper cites.
Making sense of conflicting science information: Exploring bias in the search engine result page
A. Novin and E. Meyers · 2017
Earlier work this paper cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
N. Shazeer, A. Mirhoseini, K. Maziarz, A. Davis, Q. Le, G. Hinton, and J. Dean · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Implementing machine learning in health care—addressing ethical challenges
D. S. Char, N. H. Shah, and D. Magnus · 2018
Earlier work this paper cites.
Endo3d: Online workflow analysis for endoscopic surgeries based on 3d cnn and lstm
W. Chen, J. Feng, J. Lu, and J. Zhou · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Earlier work this paper cites.
Radiology objects in context (roco): a multimodal image dataset
O. Pelka, S. Koitka, J. Rückert, F. Nensa, and C. M. Friedrich · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
A. Radford, K. Narasimhan, T. Salimans, I. Sutskever, et al · 2018
Earlier work this paper cites.
Fairvis: Visual analytics for discovering intersectional bias in machine learning
Á. A. Cabrera, W. Epperson, F. Hohman, M. Kahng, J. Morgenstern, and D. H. Chau · 2019
Earlier work this paper cites.
Transformer-XL: Attentive language models beyond a fixed-length context
Z. Dai, Z. Yang, Y. Yang, J. Carbonell, Q. V. Le, and R. Salakhutdinov · 2019
Earlier work this paper cites.
CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison
J. Irvin, P. Rajpurkar, M. Ko, Y. Yu, S. Ciurea-Ilcus, C. Chute, H. Marklund, B. Haghgoo, R. Ball, K. Shpanskaya, et al · 2019
Earlier work this paper cites.
Associations between age discrimination and health and wellbeing: cross-sectional and prospective analysis of the english longitudinal study of ageing
S. E. Jackson, R. A. Hackett, and A. Steptoe · 2019
Earlier work this paper cites.
Weakly supervised convolutional lstm approach for tool tracking in laparoscopic videos
C. I. Nwoye, D. Mutter, J. Marescaux, and N. Padoy · 2019
Earlier work this paper cites.
Dissecting racial bias in an algorithm used to manage the health of populations
Z. Obermeyer, B. Powers, C. Vogeli, and S. Mullainathan · 2019
Earlier work this paper cites.
After visit summary: Not an afterthought
E. Sieferd, N. Mohanty, and R. J. Holden · 2019
Earlier work this paper cites.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Earlier work this paper cites.
Sex and gender differences and biases in artificial intelligence for biomedicine and healthcare
D. Cirillo, S. Catuara-Solarz, C. Morey, E. Guney, L. Subirats, S. Mellino, A. Gigante, A. Valencia, M. J. Rementeria, A. S. Chadha, et al · 2020
Earlier work this paper cites.
Randaugment: Practical automated data augmentation with a reduced search space
E. D. Cubuk, B. Zoph, J. Shlens, and Q. V. Le · 2020
Earlier work this paper cites.
PathVQA: 30000+ questions for medical visual question answering
X. He, Z. Cai, W. Wei, Y. Zhang, L. Mou, E. Xing, and P. Xie · 2020
Earlier work this paper cites.
Coronavirus goes viral: quantifying the covid-19 misinformation epidemic on twitter
R. Kouzy, J. Abi Jaoude, A. Kraitem, M. B. El Alam, B. Karam, E. Adib, J. Zarka, C. Traboulsi, E. W. Akl, and K. Baddour · 2020
Earlier work this paper cites.
PAD-UFES-20: A skin lesion dataset composed of patient data and clinical images collected from smartphones
A. G. Pacheco, G. R. Lima, A. S. Salomao, B. Krohling, I. P. Biral, G. G. de Angelo, F. C. Alves Jr, J. G. Esgario, A. C. Simora, P. B. Castro, et al · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Earlier work this paper cites.
Information overload in emergency medicine physicians: a multisite case study exploring the causes, impact, and solutions in four north england national health service trusts
L. Sbaffi, J. Walton, J. Blenkinsopp, and G. Walton · 2020
Earlier work this paper cites.
Lower socioeconomic status and the acceleration of aging: An outcome-wide analysis
A. Steptoe and P. Zaninotto · 2020
Earlier work this paper cites.
PTB-XL, a large publicly available electrocardiography dataset
P. Wagner, N. Strodthoff, R.-D. Bousseljot, D. Kreiseler, F. I. Lunze, W. Samek, and T. Schaeffter · 2020
Earlier work this paper cites.
Fairness with overlapping groups; a probabilistic perspective
F. Yang, M. Cisse, and S. Koyejo · 2020
Earlier work this paper cites.
Paragraph-level simplification of medical texts
A. Devaraj, I. Marshall, B. Wallace, and J. J. Li · 2021
Earlier work this paper cites.
A real-time spatiotemporal ai model analyzes skill in open surgical videos
E. D. Goodman, K. K. Patel, Y. Zhang, W. Locke, C. J. Kennedy, R. Mehrotra, S. Ren, M. Y. Guan, M. Downing, H. W. Chen, et al · 2021
Earlier work this paper cites.
What disease does this patient have? a large-scale open domain question answering dataset from medical exams
D. Jin, E. Pan, N. Oufattole, W.-H. Weng, H. Fang, and P. Szolovits · 2021
Cited alongside, same era.
Linking the fto obesity rs1421085 variant circuitry to cellular, metabolic, and organismal phenotypes in vivo
S. Laber, S. Forcisi, L. Bentley, J. Petzold, F. Moritz, K. S. Smirnov, L. Al Sadat, I. Williamson, S. Strobel, T. Agnew, et al · 2021
Cited alongside, same era.
Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering
B. Liu, L.-M. Zhan, L. Xu, L. Ma, Y. Yang, and X.-M. Wu · 2021
Cited alongside, same era.
P. Mascagni, D. Alapatt, A. Garcia, N. Okamoto, A. Vardazaryan, G. Costamagna, B. Dallemagne, and N. Padoy · 2021
Cited alongside, same era.
Health inequities in lgbt people and nursing interventions to reduce them: A systematic review
Generative artificial intelligence for chest radiograph interpretation in the emergency department
J. Huang, L. Neill, M. Wittbrodt, D. Melnick, M. Klug, M. Thompson, J. Bailitz, T. Loftus, S. Malik, A. Phull, et al · 2023
Later among the works it cites.
MAIRA-1: A specialised large multimodal model for radiology report generation
S. L. Hyland, S. Bannur, K. Bouzid, D. C. Castro, M. Ranjit, A. Schwaighofer, F. Pérez-García, V. Salvatelli, S. Srivastav, A. Thieme, et al · 2023
Later among the works it cites.
Accuracy of a generative artificial intelligence model in a complex diagnostic challenge
Z. Kanjee, B. Crowe, and A. Rodman · 2023
Later among the works it cites.
A comparative study of pretrained language models for long clinical text
Y. Li, R. M. Wehbe, F. S. Ahmad, H. Wang, and Y. Luo · 2023
Later among the works it cites.
A translational perspective towards clinical ai fairness
M. Liu, Y. Ning, S. Teixayavong, M. Mertens, J. Xu, D. S. W. Ting, L. T.-E. Cheng, J. C. L. Ong, Z. L. Teo, T. F. Tan, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Medina-Martínez, C. Saus-Ortega, M. M. Sánchez-Lorente, E. M. Sosa-Palanca, P. García-Martínez, and M. I. Mármol-López · 2021
Cited alongside, same era.
WebGPT: Browser-assisted question-answering with human feedback
R. Nakano, J. Hilton, S. Balaji, J. Wu, L. Ouyang, C. Kim, C. Hesse, S. Jain, V. Kosaraju, W. Saunders, et al · 2021
Cited alongside, same era.
Mitigating ethnic disparities in covid-19 and beyond
M. S. Razai, H. K. Kankam, A. Majeed, A. Esmail, and D. R. Williams · 2021
Cited alongside, same era.
Worst of both worlds: Biases compound in pre-trained vision-and-language models
T. Srinivasan and Y. Bisk · 2021
Cited alongside, same era.
Finetuned language models are zero-shot learners
J. Wei, M. Bosma, V. Y. Zhao, K. Guu, A. W. Yu, B. Lester, N. Du, A. M. Dai, and Q. V. Le · 2021
Cited alongside, same era.
Flamingo: a visual language model for few-shot learning
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds, et al · 2022
Cited alongside, same era.
Pathways: Asynchronous distributed dataflow for ML
P. Barham, A. Chowdhery, J. Dean, S. Ghemawat, S. Hand, D. Hurt, M. Isard, H. Lim, R. Pang, S. Roy, et al · 2022
Cited alongside, same era.
PaLI: A jointly-scaled multilingual language-image model
X. Chen, X. Wang, S. Changpinyo, A. Piergiovanni, P. Padlewski, D. Salz, S. Goodman, A. Grycner, B. Mustafa, L. Beyer, et al · 2022
Cited alongside, same era.
Later among the works it cites.
A foundational multimodal vision language ai assistant for human pathology
M. Y. Lu, B. Chen, D. F. Williamson, R. J. Chen, K. Ikamura, G. Gerber, I. Liang, L. P. Le, T. Ding, A. V. Parwani, et al · 2023
Later among the works it cites.
Multimodal composite association score: Measuring gender bias in generative multimodal models
A. Mandal, S. Leavy, and S. Little · 2023
Later among the works it cites.
Towards accurate differential diagnosis with large language models
D. McDuff, M. Schaekermann, T. Tu, A. Palepu, A. Wang, J. Garrison, K. Singhal, Y. Sharma, S. Azizi, K. Kulkarni, et al · 2023
Later among the works it cites.
Can generalist foundation models outcompete special-purpose tuning? case study in medicine
H. Nori, Y. T. Lee, S. Zhang, D. Carignan, R. Edgar, N. Fusi, N. King, J. Larson, Y. Li, W. Liu, et al · 2023
Later among the works it cites.
Ecg-qa: A comprehensive question answering dataset combined with electrocardiogram
J. Oh, G. Lee, S. Bae, J.-m. Kwon, and E. Choi · 2023
Later among the works it cites.
Large language models propagate race-based medicine
J. A. Omiye, J. C. Lester, S. Spichak, V. Rotemberg, and R. Daneshjou · 2023
Later among the works it cites.
LongBoX: Evaluating transformers on long-sequence clinical tasks, 2023
M. Parmar, A. Naik, H. Gupta, D. Agrawal, and C. Baral · 2023
Later among the works it cites.
Hyena hierarchy: Towards larger convolutional language models
M. Poli, S. Massaroli, E. Nguyen, D. Y. Fu, T. Dao, S. Baccus, Y. Bengio, S. Ermon, and C. Ré · 2023
Later among the works it cites.
Reasoning with language model prompting: A survey, 2023
S. Qiao, Y. Ou, N. Zhang, X. Chen, Y. Yao, S. Deng, C. Tan, F. Huang, and H. Chen · 2023
Later among the works it cites.
ToolLLM: Facilitating large language models to master 16000+ real-world apis
Y. Qin, S. Liang, Y. Ye, K. Zhu, L. Yan, Y. Lu, Y. Lin, X. Cong, X. Tang, B. Qian, et al · 2023
Later among the works it cites.
Cholec80-cvs: An open dataset with an evaluation of strasberg’s critical view of safety for ai
M. S. Ríos, M. A. Molina-Rodriguez, D. Londoño, C. A. Guillén, S. Sierra, F. Zapata, and L. F. Giraldo · 2023
Later among the works it cites.
Evaluating AI systems under uncertain ground truth: a case study in dermatology, 2023
D. Stutz, A. T. Cemgil, A. G. Roy, T. Matejovicova, M. Barsbey, P. Strachan, M. Schaekermann, J. Freyberg, R. Rikhye, B. Freeman, J. P. Matos, U. Telang, D. R. Webster, Y. Liu, G. S. Corrado, Y. Matias, P. Kohli, Y. Liu, A. Doucet, and A. Karthikesalingam · 2023
Later among the works it cites.
A. Toma, P. R. Lawler, J. Ba, R. G. Krishnan, B. B. Rubin, and B. Wang · 2023
Later among the works it cites.
LLaMA: Open and efficient foundation language models
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, et al · 2023
Later among the works it cites.
Med-halt: Medical domain hallucination test for large language models
L. K. Umapathi, A. Pal, and M. Sankarasubbu · 2023
Later among the works it cites.
N. Varshney, W. Yao, H. Zhang, J. Chen, and D. Yu · 2023
Later among the works it cites.
Visual answer localization with cross-modal mutual knowledge transfer
Y. Weng and B. Li · 2023
Later among the works it cites.
S. Xu, L. Yang, C. Kelly, M. Sieniek, T. Kohlberger, M. Ma, W.-H. Weng, A. Kiraly, S. Kazemzadeh, Z. Melamed, et al · 2023
Later among the works it cites.
Tree of thoughts: Deliberate problem solving with large language models, 2023
S. Yao, D. Yu, J. Zhao, I. Shafran, T. L. Griffiths, Y. Cao, and K. Narasimhan · 2023
Later among the works it cites.
MMMU: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
X. Yue, Y. Ni, K. Zhang, T. Zheng, R. Liu, G. Zhang, S. Stevens, D. Jiang, W. Ren, Y. Sun, et al · 2023
Later among the works it cites.
Least-to-most prompting enables complex reasoning in large language models, 2023
D. Zhou, N. Schärli, L. Hou, J. Wei, N. Scales, X. Wang, D. Schuurmans, C. Cui, O. Bousquet, Q. Le, and E. Chi · 2023
Later among the works it cites.
Graph of thoughts: Solving elaborate problems with large language models, 2024
M. Besta, N. Blach, A. Kubicek, R. Gerstenberger, M. Podstawski, L. Gianinazzi, J. Gajda, T. Lehmann, H. Niewiadomski, P. Nyczyk, and T. Hoefler · 2024
Closest in time.
Retrieval-augmented generation for large language models: A survey, 2024
Y. Gao, Y. Xiong, X. Gao, K. Jia, J. Pan, Y. Bi, Y. Dai, J. Sun, Q. Guo, M. Wang, and H. Wang · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Gemini Team, Google · 2024
Closest in time.
Analyzing surgical technique in diverse open surgical videos with multitask machine learning
E. D. Goodman, K. K. Patel, Y. Zhang, W. Locke, C. J. Kennedy, R. Mehrotra, S. Ren, M. Guan, O. Zohar, M. Downing, et al · 2024
Closest in time.
ToolkenGPT: Augmenting frozen language models with massive tools via tool embeddings
S. Hao, T. Liu, Z. Wang, and Z. Hu · 2024
Closest in time.
GeneGPT: Augmenting large language models with domain tools for improved access to biomedical information
Q. Jin, Y. Yang, Q. Chen, and Z. Lu · 2024
Closest in time.
LLaVa-Med: Training a large language-and-vision assistant for biomedicine in one day
C. Li, C. Wong, S. Zhang, N. Usuyama, H. Liu, J. Yang, T. Naumann, H. Poon, and J. Gao · 2024
Closest in time.
Lost in the middle: How language models use long contexts
N. F. Liu, K. Lin, J. Hewitt, A. Paranjape, M. Bevilacqua, F. Petroni, and P. Liang · 2024
Closest in time.
Papers with code - medical, 2024
Meta · 2024
Closest in time.
A toolbox for surfacing health equity harms and biases in large language models
S. R. Pfohl, H. Cole-Lewis, R. Sayres, D. Neal, M. Asiedu, A. Dieng, N. Tomasev, Q. M. Rashid, S. Azizi, N. Rostamzadeh, et al · 2024
Closest in time.
Toolformer: Language models can teach themselves to use tools
T. Schick, J. Dwivedi-Yu, R. Dessì, R. Raileanu, M. Lomeli, E. Hambro, L. Zettlemoyer, N. Cancedda, and T. Scialom · 2024
Closest in time.
Consensus, dissensus and synergy between clinicians and specialist foundation models in radiology report generation
R. Tanno, D. Barrett, A. Sellergren, S. Ghaisas, S. Dathathri, A. See, J. Welbl, K. Singhal, S. Azizi, T. Tu, et al · 2024
Closest in time.
Image challenge
The New England Journal of Medicine · 2024
Closest in time.
Electrocardiogram instruction tuning for report generation, 2024
Z. Wan, C. Liu, X. Wang, C. Tao, H. Shen, Z. Peng, J. Fu, R. Arcucci, H. Yao, and M. Zhang · 2024
Closest in time.
Exploring the reasoning abilities of multimodal large language models (mllms): A comprehensive survey on emerging trends in multimodal reasoning, 2024
Y. Wang, W. Chen, X. Han, X. Lin, H. Zhao, Y. Liu, B. Zhai, J. Yuan, Q. You, and H. Yang · 2024
Closest in time.
A. Ward, J. Li, J. Wang, S. Lakshminarasimhan, A. Carrick, B. Campana, J. Hartford, T. Tiyasirichokchai, S. Virmani, R. Wong, et al · 2024
Closest in time.
How well do llms cite relevant medical references? an evaluation framework and analyses
K. Wu, E. Wu, A. Cassasola, A. Zhang, K. Wei, T. Nguyen, S. Riantawan, P. S. Riantawan, D. E. Ho, and J. Zou · 2024
Closest in time.
Almanac—retrieval-augmented language models for clinical medicine
C. Zakka, R. Shad, A. Chaurasia, A. R. Dalal, J. L. Kim, M. Moor, R. Fong, C. Phillips, K. Alexander, E. Ashley, et al · 2024
Closest in time.
J. M. Zambrano Chaves, S.-C. Huang, Y. Xu, H. Xu, N. Usuyama, S. Zhang, F. Wang, Y. Xie, M. Khademi, Z. Yang, et al · 2024
Closest in time.
Raft: Adapting language model to domain specific rag
T. Zhang, S. G. Patil, N. Jain, S. Shen, M. Zaharia, I. Stoica, and J. E. Gonzalez · 2024
Closest in time.