Fetching the paper…
Reading the bibliography…
Data imbalance is a fundamental challenge in applying language models to biomedical applications, particularly in ICD code prediction tasks where label and demographic distributions are uneven.
E. Alsentzer, J. Murphy, W. Boag, W.-H. Weng, D. Jindi, T. Naumann, and M. McDermott, “Publicly available clinical BERT embeddings,” in Proceedings of the 2nd Clinical Natural Language Processing Workshop , A. Rumshisky, K. Roberts, S. Bethard, and T. Naumann, Eds. Minneapolis, Minnesota, USA: Association for Computational Linguistics, Jun. 2019, pp. 72–78. [Online]. Available: https://aclanthology.org/W19-1909
1909
Earlier work this paper cites.
A. L. Goldberger, L. A. Amaral, L. Glass, J. M. Hausdorff, P. C. Ivanov, R. G. Mark, J. E. Mietus, G. B. Moody, C.-K. Peng, and H. E. Stanley, “Physiobank, physiotoolkit, and physionet: Components of a new research resource for complex physiologic signals,” Circulation , vol. 101, no. 23, pp. e215–e220, 2000
2000
Earlier work this paper cites.
2008
Earlier work this paper cites.
G. Mercuro, M. Deidda, A. Piras, C. C. Dessalvi, S. Maffei, and G. M. Rosano, “Gender determinants of cardiovascular risk factors and diseases,” Journal of Cardiovascular Medicine , vol. 11, no. 3, pp. 207–220, 2010
2010
Earlier work this paper cites.
O. L. Quintero, M. J. Amador-Patarroyo, G. Montoya-Ortiz, A. Rojas-Villarraga, and J.-M. Anaya, “Autoimmune disease and gender: plausible mechanisms for the female predominance of autoimmunity,” Journal of autoimmunity , vol. 38, no. 2-3, pp. J109–J119, 2012
2012
Earlier work this paper cites.
L. Dixon, J. Li, J. Sorensen, N. Thain, and L. Vasserman, “Measuring and mitigating unintended bias in text classification,” in Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , 2018, pp. 67–73
2018
Earlier work this paper cites.
S.-C. Tsai, T.-Y. Chang, and Y.-N. Chen, “Leveraging hierarchical category knowledge for data-imbalanced multi-label diagnostic text understanding,” in Proceedings of the 10th international workshop on health text mining and information analysis (LOUHI) , 2019, pp. 39–43
2019
Earlier work this paper cites.
K. Xu, M. Lam, J. Pang, X. Gao, C. Band, P. Mathur, F. Papay, A. K. Khanna, J. B. Cywinski, K. Maheshwari, P. Xie, and E. P. Xing, “Multimodal machine learning for automated icd coding,” in Proceedings of the 4th Machine Learning for Healthcare Conference , ser. Proceedings of Machine Learning Research, vol. 106. PMLR, 09–10 Aug 2019, pp. 197–215. [Online]. Available: https://proceedings.mlr.press/v106/xu19a.html
2019
Earlier work this paper cites.
S. Wang, M. E. Elkin, and X. Zhu, “Imbalanced learning for hospital readmission prediction using national readmission database,” in 2020 IEEE International Conference on Knowledge Graph (ICKG) . IEEE, 2020, pp. 116–122
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
S. Kino, Y.-T. Hsu, K. Shiba, Y.-S. Chien, C. Mita, I. Kawachi, and A. Daoud, “A scoping review on the use of machine learning in research on social determinants of health: Trends and research prospects,” SSM - Population Health , vol. 15, p. 100836, 2021. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S2352827321001117
2021
Earlier work this paper cites.
S. Ji, M. Hölttä, and P. Marttinen, “Does the magic of bert apply to medical code assignment? a quantitative study,” Computers in biology and medicine , vol. 139, p. 104998, 2021
2021
Earlier work this paper cites.
H. Lu, L. Ehwerhemuepha, and C. Rakovski, “A comparative study on deep learning models for text classification of unstructured medical notes with various levels of class imbalance,” BMC medical research methodology , vol. 22, no. 1, p. 181, 2022
2022
Cited alongside, same era.
K. De Angeli, S. Gao, I. Danciu, E. B. Durbin, X.-C. Wu, A. Stroup, J. Doherty, S. Schwartz, C. Wiggins, M. Damesyn, L. Coyle, L. Penberthy, G. D. Tourassi, and H.-J. Yoon, “Class imbalance in out-of-distribution datasets: Improving the robustness of the textcnn for the classification of rare cancer types,” Journal of Biomedical Informatics , vol. 125, p. 103957, 2022. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S1532046421002860
2022
Cited alongside, same era.
E. Röösli, S. Bozkurt, and T. Hernandez-Boussard, “Peeking into a black box, the fairness and generalizability of a MIMIC-III benchmarking model,” Scientific Data , vol. 9, no. 1, p. 24, jan 2022. [Online]. Available: https://www.nature.com/articles/s41597-021-01110-7
2022
Cited alongside, same era.
S. Henning, W. Beluch, A. Fraser, and A. Friedrich, “A survey of methods for addressing class imbalance in deep-learning based natural language processing,” in Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics , A. Vlachos and I. Augenstein, Eds. Dubrovnik, Croatia: Association for Computational Linguistics, May 2023, pp. 523–540. [Online]. Available: https://aclanthology.org/2023.eacl-main.38
2023
Later among the works it cites.
M. Y. Yang, G. H. Kwak, T. Pollard, L. A. Celi, and M. Ghassemi, “Evaluating the impact of social determinants on health prediction in the intensive care unit,” in Proceedings of the 2023 AAAI/ACM Conference on AI, Ethics, and Society , ser. AIES ’23. New York, NY, USA: Association for Computing Machinery, 2023, p. 333–350. [Online]. Available: https://doi.org/10.1145/3600211.3604719
2023
Later among the works it cites.
A. Johnson, T. Pollard, S. Horng, L. A. Celi, and R. Mark, “Mimic-iv-note: Deidentified free-text clinical notes (version 2.2),” 2023. [Online]. Available: https://doi.org/10.13026/1n74-ne17
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
Y. Li, R. M. Wehbe, F. S. Ahmad, H. Wang, and Y. Luo, “Clinical-longformer and clinical-bigbird: Transformers for long clinical sequences,” 2022
2022
Cited alongside, same era.
W. Lyu, X. Dong, R. Wong, S. Zheng, K. Abell-Hart, F. Wang, and C. Chen, “A multimodal transformer: Fusing clinical notes with structured ehr data for interpretable in-hospital mortality prediction,” in AMIA Annual Symposium Proceedings , vol. 2022. American Medical Informatics Association, 2022, p. 719
2022
Cited alongside, same era.
Y. Shi, T. ValizadehAslani, J. Wang, P. Ren, Y. Zhang, M. Hu, L. Zhao, and H. Liang, “Improving imbalanced learning by pre-finetuning with data augmentation,” in Proceedings of the Fourth International Workshop on Learning with Imbalanced Domains: Theory and Applications , ser. Proceedings of Machine Learning Research, N. Moniz, P. Branco, L. Torgo, N. Japkowicz, M. Wozniak, and S. Wang, Eds., vol. 183. PMLR, 23 Sep 2022, pp. 68–82
2022
Cited alongside, same era.
Y. Wu, I.-C. Huang, and X. Huang, “Token imbalance adaptation for radiology report generation,” in Proceedings of the Conference on Health, Inference, and Learning , ser. Proceedings of Machine Learning Research, B. J. Mortazavi, T. Sarker, A. Beam, and J. C. Ho, Eds., vol. 209. PMLR, 22 Jun–24 Jun 2023, pp. 72–85. [Online]. Available: https://proceedings.mlr.press/v209/wu23a.html
2023
Cited alongside, same era.
A. E. W. Johnson, L. Bulgarelli, L. Shen, A. Gayles, A. Shammout, S. Horng, T. J. Pollard, S. Hao, B. Moody, B. Gow, L.-w. H. Lehman, L. A. Celi, and R. G. Mark, “MIMIC-IV, a freely accessible electronic health record dataset,” Scientific Data , vol. 10, no. 1, p. 1, 2023. [Online]. Available: https://doi.org/10.1038/s41597-022-01899-x
2023
Cited alongside, same era.
N. A. Cloutier and N. Japkowicz, “Fine-tuned generative llm oversampling can improve performance over traditional techniques on multiclass imbalanced text classification,” in 2023 IEEE International Conference on Big Data (BigData) . IEEE, 2023, pp. 5181–5186
2023
Cited alongside, same era.
M. Wornow, Y. Xu, R. Thapa, B. Patel, E. Steinberg, S. Fleming, M. A. Pfeffer, J. Fries, and N. H. Shah, “The shaky foundations of large language models and foundation models for electronic health records,” npj Digital Medicine , vol. 6, no. 1, p. 135, jul 2023. [Online]. Available: https://www.nature.com/articles/s41746-023-00879-8
2023
Cited alongside, same era.
Y. Jin, Y. Xiong, D. Shi, Y. Lin, L. He, Y. Zhang, J. M. Plasek, L. Zhou, D. W. Bates, and C. Tang, “Learning from undercoded clinical records for automated international classification of diseases (icd) coding,” Journal of the American Medical Informatics Association , vol. 30, no. 3, pp. 438–446, 2023
2023
Later among the works it cites.
T. Kuroiwa, A. Sarcon, T. Ibara, E. Yamada, A. Yamamoto, K. Tsukamoto, and K. Fujita, “The potential of chatgpt as a self-diagnostic tool in common orthopedic diseases: exploratory study,” Journal of Medical Internet Research , vol. 25, p. e47621, 2023
2023
Later among the works it cites.
Y. Li, R. M. Wehbe, F. S. Ahmad, H. Wang, and Y. Luo, “A comparative study of pretrained language models for long clinical text,” Journal of the American Medical Informatics Association , vol. 30, no. 2, pp. 340–347, 2023
2023
Later among the works it cites.
S. Gonçalves and G. Ehrensperger, “Icd-mappings,” 2023. [Online]. Available: https://github.com/snovaisg/ICD-Mappings
2023
Later among the works it cites.
S. Nerella, S. Bandyopadhyay, J. Zhang, M. Contreras, S. Siegel, A. Bumin, B. Silva, J. Sena, B. Shickel, A. Bihorac et al. , “Transformers and large language models in healthcare: A review,” Artificial Intelligence in Medicine , p. 102900, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
C.-Y. Hsu, K. Cox, J. Xu, Z. Tan, T. Zhai, M. Hu, D. Pratt, T. Chen, Z. Hu, and Y. Ding, “Thought graph: Generating thought process for biological reasoning,” in Companion Proceedings of the ACM on Web Conference 2024 , 2024, pp. 537–540
2024
Closest in time.