A neural probabilistic language model
Bengio Y, Ducharme R, Vincent P · 2000
Earlier work this paper cites.
De-identification algorithm for free-text nursing notes
M M , G D Cliffford Douglass, Andrew Reisner, W J Long, G B Moody, R G Mark · 2005
Earlier work this paper cites.
Automated de-identification of free-text medical records
Ishna Neamatullah, Margaret M Douglass, Li-Wei H Lehman, Andrew Reisner, Mauricio Villarroel, William J Long, Peter Szolovits, George B Moody, Roger G Mark, Gari D Clifford · 2008
Earlier work this paper cites.
Recurrent neural network based language model
Mikolov T, Karafiát M, Burget L, Cernockỳ J, Khudanpur S · 2010
Earlier work this paper cites.
Long short-term memory
Graves A, Graves A · 2012
Earlier work this paper cites.
The Impact of De-identification on Downstream Named Entity Recognition in Clinical Text
Hanna Berg, Aron Henriksson, Hercules Dalianis · 2012
Earlier work this paper cites.
Automatic detection of protected health information from clinic narratives
Hui Yang JMG · 2015
Earlier work this paper cites.
Annotating longitudinal clinical narratives for de-identification: The 2014 i2b2/UTHealth corpus
Stubbs A, Uzuner Ö · 2015
Earlier work this paper cites.
MIMIC-III, a freely accessible critical care database
Johnson AE, Pollard TJ, Shen L, Lehman LwH, Feng M, Ghassemi M, et al · 2016
Earlier work this paper cites.
Gate-variants of gated recurrent unit (GRU) neural networks
Dey R, Salem FM · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, et al · 2017
Earlier work this paper cites.
Sharp nearby, fuzzy far away: How neural language models use context
Khandelwal U, He H, Qi P, Jurafsky D · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin J, Chang MW, Lee K, Toutanova K · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training. 2018
Radford A, Narasimhan K, Salimans T, Sutskever I, et al · 2018
Earlier work this paper cites.
Privacy protection of patient medical images using digital watermarking technique for E-healthcare system
Asokan Sivaprakash, Samuel NE Rajan, and Sundaramoorthy Selvaperumal · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu Y, Ott M, Goyal N, Du J, Joshi M, Chen D, et al · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford A, Wu J, Child R, Luan D, Amodei D, Sutskever I, et al · 2019
Earlier work this paper cites.
Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Lewis M, Liu Y, Goyal N, Ghazvininejad M, Mohamed A, Levy O, et al · 2019
Earlier work this paper cites.
Pseudonymisation of Swedish electronic patient records using a rule-based approach
Hercules Dalianis · 2019
Earlier work this paper cites.
A study of deep learning methods for de-identification of clinical notes in cross-institute settings
Xi Yang, Tianchen Lyu, Qian Li, Chih-Yin Lee, Jiang Bian, William R Hogan, Yonghui Wu · 2019
Earlier work this paper cites.
Publicly available clinical BERT embeddings
Alsentzer E, Murphy JR, Boag W, Weng WH, Jin D, Naumann T, et al · 2019
Earlier work this paper cites.
De-identification of electronic health record using neural network
Tanbir Ahmed, Md Momin Al Aziz, Noman Mohammed · 2020
Earlier work this paper cites.
Language models are few-shot learners
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, et al · 2020
Earlier work this paper cites.
Pre-trained models for natural language processing: A survey
Qiu X, Sun T, Xu Y, Shao Y, Dai N, Huang X · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel C, Shazeer N, Roberts A, Lee K, Narang S, Matena M, et al · 2020
Earlier work this paper cites.
It’s not just size that matters: Small language models are also few-shot learners
Schick T, Schütze H · 2020
Earlier work this paper cites.
Autoprompt: Eliciting knowledge from language models with automatically generated prompts
Shin T, Razeghi Y, Logan IV RL, Wallace E, Singh S · 2020
Earlier work this paper cites.
Survey on RNN and CRF models for de-identification of medical free text
Joffrey L Leevy, Taghi M Khoshgoftaar, Flavio Villanustre · 2020
Earlier work this paper cites.
Crosslingual named entity recognition for clinical de-identification applied to a COVID-19 Italian data set
Catelli R, Gargiulo F, Casola V, De Pietro G, Fujita H, Esposito M · 2020
Earlier work this paper cites.
Accelerating training of transformer-based language models with progressive layer dropping
Zhang M, He Y · 2020
Earlier work this paper cites.
BioBERT: a pre-trained biomedical language representation model for biomedical text mining
Lee J, Yoon W, Kim S, Kim D, Kim S, So CH, et al · 2020
Earlier work this paper cites.
Challenges and opportunities beyond structured data in analysis of electronic health records
Tayefi M, Ngo P, Chomutare T, Dalianis H, Salvi E, Budrionis A, et al · 2021
Earlier work this paper cites.
Bio-signal data sharing security through watermarking: a technical survey
Nandita Sharma, Ashima Anand, Amit Kumar Singh · 2021
Earlier work this paper cites.
A mathematical framework for transformer circuits
Elhage N, Nanda N, Olsson C, Henighan T, Joseph N, Mann B, et al · 2021
Earlier work this paper cites.
Jurassic-1: Technical details and evaluation
Lieber O, Sharir O, Lenz B, Shoham Y · 2021
Earlier work this paper cites.
GPT understands, too
Liu X, Zheng Y, Du Z, Ding M, Qian Y, Yang Z, et al · 2021
Earlier work this paper cites.
Chestxraybert: A pretrained language model for chest radiology report summarization
Cai X, Liu S, Han J, Yang L, Liu Z, Liu T · 2021
Earlier work this paper cites.