Fetching the paper…
Reading the bibliography…
The rapid development of artificial intelligence has constantly reshaped the field of intelligent healthcare and medicine.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
A general framework for multiresolution image fusion: from pixels to regions
G. Piella · 2003
Earlier work this paper cites.
The unified medical language system (umls): integrating biomedical terminology
O. Bodenreider · 2004
Earlier work this paper cites.
High-level information fusion: An overview
P. H. Foo and G. W. Ng · 2013
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow et al · 2014
Earlier work this paper cites.
A general framework for image fusion based on multi-scale transform and sparse representation
Y. Liu et al · 2015
Earlier work this paper cites.
Multi-modality medical image fusion using discrete wavelet transform
V. Bhavana and H. Krishnappa · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger et al · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He et al · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
T. N. Kipf and M. Welling · 2016
Earlier work this paper cites.
Preparing a collection of radiology examinations for distribution and retrieval
D. Demner-Fushman et al · 2016
Earlier work this paper cites.
Generating binary tags for fast medical image retrieval based on convolutional nets and radon transform
X. Liu et al · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He et al · 2016
Earlier work this paper cites.
3d u-net: learning dense volumetric segmentation from sparse annotation
Ö. Çiçek et al · 2016
Earlier work this paper cites.
Attention is all you need
A. Vaswani et al · 2017
Earlier work this paper cites.
Overview of imageclefcaption 2017 - image caption prediction and concept detection for biomedical images
C. Eickhoff et al · 2017
Earlier work this paper cites.
Pixel-level image fusion: A survey of the state of the art
S. Li et al · 2017
Earlier work this paper cites.
Medical image synthesis with context-aware generative adversarial networks
D. Nie et al · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman et al · 2017
Earlier work this paper cites.
Tienet: Text-image embedding network for common thorax disease classification and reporting in chest x-rays
X. Wang et al · 2018
Earlier work this paper cites.
Overview of the imageclef 2018 caption prediction tasks
A. G. S. de Herrera et al · 2018
Earlier work this paper cites.
On the automatic generation of medical imaging reports
B. Jing et al · 2018
Earlier work this paper cites.
Radiology objects in context (ROCO): A multimodal image dataset
O. Pelka et al · 2018
Earlier work this paper cites.
Overview of imageclef 2018 medical domain visual question answering task
S. A. Hasan et al · 2018
Earlier work this paper cites.
A dataset of clinically generated visual questions and answers about radiology images
J. J. Lau et al · 2018
Earlier work this paper cites.
On the automatic generation of medical imaging reports
B. Jing et al · 2018
Earlier work this paper cites.
Cross-modality image synthesis from unpaired data using cyclegan: Effects of gradient consistency loss and training data size
Y. Hiasa et al · 2018
Earlier work this paper cites.
Synthesizing missing pet from mri with cycle-consistent generative adversarial networks for alzheimer’s disease diagnosis
Y. Pan et al · 2018
Earlier work this paper cites.
Generation of structural mr images from amyloid pet: application to mr-less quantification
H. Choi and D. S. Lee · 2018
Earlier work this paper cites.
Yolov3: An incremental improvement
J. Redmon and A. Farhadi · 2018
Earlier work this paper cites.
Multimodal machine learning: A survey and taxonomy
T. Baltrusaitis et al · 2019
Earlier work this paper cites.
Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports
A. E. Johnson et al · 2019
Earlier work this paper cites.
Vqa-med: Overview of the medical visual question answering task at imageclef 2019
A. Ben Abacha et al · 2019
Earlier work this paper cites.
Show, describe and conclude: On exploiting the structure information of chest x-ray reports
B. Jing et al · 2019
Earlier work this paper cites.
Overcoming data limitation in medical visual question answering
B. D. Nguyen et al · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin et al · 2019
Earlier work this paper cites.
Cross-modality synthesis from ct to pet using fcn and gan networks for improved automated lesion detection
A. Ben-Cohen et al · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Y. Liu et al · 2019
Earlier work this paper cites.
Publicly available clinical bert embeddings
E. Alsentzer et al · 2019
Earlier work this paper cites.
Automated diagnosis of multi-class brain abnormalities using mri images: a deep convolutional neural network based method
D. R. Nayak et al · 2020
Earlier work this paper cites.
Padchest: A large chest x-ray image dataset with multi-label annotated reports
A. Bustos et al · 2020
Earlier work this paper cites.
Medicat: A dataset of medical images, captions, and textual references
S. Subramanian et al · 2020
Earlier work this paper cites.
Overview of the vqa-med task at imageclef 2020: Visual question answering and generation in the medical domain
A. B. Abacha et al · 2020
Earlier work this paper cites.
Towards visual dialog for radiology
O. Kovaleva et al · 2020
Earlier work this paper cites.
Pathvqa: 30000+ questions for medical visual question answering
X. He et al · 2020
Earlier work this paper cites.
Advances in multimodal data fusion in neuroimaging: Overview, challenges, and novel orientation
Y.-D. Zhang et al · 2020
Earlier work this paper cites.
Bertscore: Evaluating text generation with BERT
T. Zhang et al · 2020
Earlier work this paper cites.
Cross-modality medical image retrieval with deep features
A. Mbilinyi and H. Schuldt · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
J. Ho et al · 2020
Earlier work this paper cites.
Resunet-a: A deep learning framework for semantic segmentation of remotely sensed data
F. I. Diakogiannis et al · 2020
Earlier work this paper cites.
Pubmed parser: A python parser for pubmed open-access xml subset and medline xml dataset xml dataset
T. Achakulvisut et al · 2020
Earlier work this paper cites.
Learning visual-semantic embeddings for reporting abnormal findings on chest x-rays
J. Ni et al · 2020
Earlier work this paper cites.
Autoprompt: Eliciting knowledge from language models with automatically generated prompts
T. Shin et al · 2020
Earlier work this paper cites.
A review of applications in federated learning
L. Li et al · 2020
Earlier work this paper cites.
Domain adaptation for medical image analysis: a survey
H. Guan and M. Liu · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford et al · 2021
Earlier work this paper cites.
Multiple instance captioning: Learning representations from histopathology textbooks and articles
J. Gamper and N. Rajpoot · 2021
Earlier work this paper cites.
FFA-IR: towards an explainable and reliable medical report generation benchmark
M. Li et al · 2021
Earlier work this paper cites.
Overview of the vqa-med task at imageclef 2021: Visual question answering and generation in the medical domain
A. Ben Abacha et al · 2021
Earlier work this paper cites.
SLAKE: A semantically-labeled knowledge-enhanced dataset for medical visual question answering
B. Liu et al · 2021
Earlier work this paper cites.
Multi scale decomposition based medical image fusion using convolutional neural network and sparse representation
D. S. Shibu and S. S. Priyadharsini · 2021
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
A. Dosovitskiy et al · 2021
Earlier work this paper cites.
Variational topic inference for chest x-ray report generation
I. Najdenkoska et al · 2021
Earlier work this paper cites.
Radgraph: Extracting clinical entities and relations from radiology reports
S. Jain et al · 2021
Earlier work this paper cites.
Multiple meta-model quantifying for medical visual question answering
T. Do et al · 2021
Earlier work this paper cites.
Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition
S. Huang et al · 2021
Earlier work this paper cites.
Domain-specific language model pretraining for biomedical natural language processing
Y. Gu et al · 2021
Earlier work this paper cites.
High-performance large-scale image recognition without normalization
A. Brock et al · 2021
Earlier work this paper cites.
Emerging properties in self-supervised vision transformers
M. Caron et al · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
E. J. Hu et al · 2021
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
P. Esser et al · 2021
Cited alongside, same era.
Scaling up visual and vision-language representation learning with noisy text supervision
C. Jia et al · 2021
Cited alongside, same era.
Underdiagnosis bias of artificial intelligence algorithms applied to chest radiographs in under-served patient populations
L. Seyyed-Kalantari et al · 2021
Cited alongside, same era.
Producing personalized statin treatment plans to optimize clinical outcomes using big data and machine learning
K. Zhang et al · 2023
Later among the works it cites.
Xraygpt: Chest radiographs summarization using medical vision-language models
O. Thawkar et al · 2023
Later among the works it cites.
Llava-med: Training a large language-and-vision assistant for biomedicine in one day
C. Li et al · 2023
Later among the works it cites.
Med-flamingo: a multimodal medical few-shot learner
M. Moor et al · 2023
Later among the works it cites.
Towards generalist foundation model for radiology
C. Wu et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C.-L. Chi et al · 2022
Cited alongside, same era.
Multimodal biomedical ai
J. N. Acosta et al · 2022
Cited alongside, same era.
A survey on deep learning and explainability for automatic report generation from medical images
P. Messina et al · 2022
Cited alongside, same era.
Sf-net: A multi-task model for brain tumor segmentation in multimodal mri via image fusion
Y. Liu et al · 2022
Cited alongside, same era.
Y. Chen et al · 2022
Cited alongside, same era.
Improving the factual correctness of radiology report generation with semantic rewards
J. Delbrouck et al · 2022
Cited alongside, same era.
OVQA: A clinically generated visual question answering dataset
Y. Huang et al · 2022
Cited alongside, same era.
C. Pellegrini et al · 2023
Later among the works it cites.
Qilin-med-vl: Towards chinese large vision-language model for general healthcare
J. Liu et al · 2023
Later among the works it cites.
Maira-1: A specialised large multimodal model for radiology report generation
S. L. Hyland et al · 2023
Later among the works it cites.
A foundational multimodal vision language ai assistant for human pathology
M. Y. Lu et al · 2023
Later among the works it cites.
Medxchat: Bridging cxr modalities with a unified multimodal large model
L. Yang et al · 2023
Later among the works it cites.
Eva: Exploring the limits of masked visual representation learning at scale
Y. Fang et al · 2023
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, March 2023
W.-L. Chiang et al · 2023
Later among the works it cites.
A. Q. Jiang et al · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
H. Touvron et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
H. Touvron et al · 2023
Later among the works it cites.
Efficient and effective text encoding for chinese llama and alpaca
Y. Cui et al · 2023
Later among the works it cites.
Pmc-llama: Further finetuning llama on medical papers
C. Wu et al · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
L. Zhang et al · 2023
Later among the works it cites.
Lmm-assisted breast cancer treatment target segmentation with consistency embedding
K. Kim et al · 2023
Later among the works it cites.
Segvol: Universal and interactive volumetric medical image segmentation
Y. Du et al · 2023
Later among the works it cites.
Foundational models in medical imaging: A comprehensive survey and future vision
B. Azad et al · 2023
Later among the works it cites.
Contrastive graph representations for logical formulas embedding
Q. Lin et al · 2023
Later among the works it cites.
Building a knowledge graph to enable precision medicine
P. Chandak et al · 2023
Later among the works it cites.
Rethinking tokenizer and decoder in masked graph modeling for molecules
Z. Liu et al · 2023
Later among the works it cites.
Multimodal chain-of-thought reasoning in language models
Z. Zhang et al · 2023
Later among the works it cites.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao et al · 2023
Later among the works it cites.
Techs: Temporal logical graph networks for explainable extrapolation reasoning
Q. Lin et al · 2023
Later among the works it cites.
Leveraging biomolecule and natural language through multi-modal learning: A survey
Q. Pei et al · 2024
Closest in time.
Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Y. Hu et al · 2024
Closest in time.
A visual-language foundation model for computational pathology
M. Y. Lu et al · 2024
Closest in time.
Pre-trained multimodal large language model enhances dermatological diagnosis using skingpt-4
J. Zhou et al · 2024
Closest in time.
Work like a doctor: Unifying scan localizer and dynamic generator for automated computed tomography report generation
Y. Tang et al · 2024
Closest in time.
Hybrid-net: A fusion of densenet169 and advanced machine learning classifiers for enhanced brain tumor diagnosis
S. U. R. Khan et al · 2024
Closest in time.
Improving radiology report generation quality and diversity through reinforcement learning and text augmentation
D. Parres et al · 2024
Closest in time.
Dtan: Diffusion-based text attention network for medical image segmentation
Y. Zhao et al · 2024
Closest in time.
Diffusion model-based text-guided enhancement network for medical image segmentation
Z. Dong et al · 2024
Closest in time.
Pathasst: A generative foundation ai assistant towards artificial general intelligence of pathology
Y. Sun et al · 2024
Closest in time.
I. E. Hamamci et al · 2024
Closest in time.
Pairaug: What can augmented image-text pairs do for radiology?
Y. Xie et al · 2024
Closest in time.
Z. Li et al · 2024
Closest in time.
M. H. Phan et al · 2024
Closest in time.
Knowledge-enhanced visual-language pretraining for computational pathology
X. Zhou et al · 2024
Closest in time.
Devide: Faceted medical knowledge for improved medical vision-language pre-training
H. Luo et al · 2024
Closest in time.
A. Q. Jiang et al · 2024
Closest in time.
Minigpt-4: Enhancing vision-language understanding with advanced large language models
D. Zhu et al · 2024
Closest in time.
Llm-cxr: Instruction-finetuned llm for cxr image understanding and generation
S. Lee et al · 2024
Closest in time.
Towards generalist biomedical ai
T. Tu et al · 2024
Closest in time.
Chexagent: Towards a foundation model for chest x-ray interpretation
Z. Chen et al · 2024
Closest in time.
M3d: Advancing 3d medical image analysis with multi-modal large language models
F. Bai et al · 2024
Closest in time.
Dia-llama: Towards large language model-driven ct report generation
Z. Chen et al · 2024
Closest in time.
Training small multimodal models to bridge biomedical competency gap: A case study in radiology imaging
J. M. Zambrano Chaves et al · 2024
Closest in time.
Wolf: Large language model framework for cxr understanding
S. Kang et al · 2024
Closest in time.
Rad-dino: Exploring scalable medical image encoders beyond text supervision
F. Pérez-García et al · 2024
Closest in time.
Y. Kim et al · 2024
Closest in time.
Towards a general-purpose foundation model for computational pathology
R. J. Chen et al · 2024
Closest in time.
Detecting and evaluating medical hallucinations in large vision language models
J. Chen et al · 2024
Closest in time.
A. Pal and M. Sankarasubbu · 2024
Closest in time.
Data-centric foundation models in computational healthcare: A survey
Y. Zhang et al · 2024
Closest in time.
Continual self-supervised learning: Towards universal multi-modal medical data representation learning
Y. Ye et al · 2024
Closest in time.
Mini-gemini: Mining the potential of multi-modality vision language models
Y. Li et al · 2024
Closest in time.
Monkey: Image resolution and text label are important things for large multi-modal models
Z. Li et al · 2024
Closest in time.
Mitigating large language model hallucinations via autonomous knowledge graph-based retrofitting
X. Guan et al · 2024
Closest in time.
Next-gpt: Any-to-any multimodal llm
S. Wu et al · 2024
Closest in time.
Graph of thoughts: Solving elaborate problems with large language models
M. Besta et al · 2024
Closest in time.
Symbol-llm: Towards foundational symbol-centric interface for large language models
F. Xu et al · 2024
Closest in time.
Robust visual question answering: Datasets, methods, and future challenges
J. Ma et al · 2024
Closest in time.
Hallucination of multimodal large language models: A survey
Z. Bai et al · 2024
Closest in time.