Fetching the paper…
Reading the bibliography…
Natural Language Processing (NLP) has been revolutionized by the use of Pre-trained Language Models (PLMs) such as BERT.
W. Yu et al. , “Dict-bert: Enhancing language model pre-training with dictionary,” in Findings Assoc. Comput. Linguistic , 2022, pp. 1907–1918
1918
Earlier work this paper cites.
M. Glass, G. Rossiello, M. F. M. Chowdhury, and A. Gliozzo, “Robust retrieval augmented generation for zero-shot slot filling,” in Proc. 2021 Conf. Emplir. Methods Natural Lang. Process. , 2021, pp. 1939–1949
1949
Earlier work this paper cites.
X. Pan, B. Zhang, J. May, J. Nothman, K. Knight, and H. Ji, “Cross-lingual name tagging and linking for 282 languages,” in Proc. 55th Annu. Meeting Assoc. Comput. Linguistics , vol. 1, 2017, pp. 1946–1958
1958
Earlier work this paper cites.
G. A. Miller, “Wordnet: a lexical database for english,” Commun. ACM , vol. 38, no. 11, pp. 39–41, 1995
1995
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural Comput. , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
P. Singh, T. Lin, E. T. Mueller, G. Lim, T. Perkins, and W. L. Zhu, “Open mind common sense: Knowledge acquisition from the general public,” in OTM Confederated Int. Conf., “On the Move to Meaningful Internet Systems” . Berlin, Germany: Springer, 2002, pp. 1223–1237
2002
Earlier work this paper cites.
E. F. Sang and F. De Meulder, “Introduction to the conll-2003 shared task: Language-independent named entity recognition,” in Proc. Conference North Amer. Assoc. Comput. Linguistics , 2003, pp. 142–147
2003
Earlier work this paper cites.
J.-D. Kim, T. Ohta, Y. Tsuruoka, Y. Tateisi, and N. Collier, “Introduction to the bio-entity recognition task at jnlpba,” in Proc. Int. Joint Workshop Natural Lang. Process. Biomedicine and its Appl. Citeseer, 2004, pp. 70–75
2004
Earlier work this paper cites.
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. Ives, “Dbpedia: A nucleus for a web of open data,” in The semantic web . Springer, 2007, pp. 722–735
2007
Earlier work this paper cites.
Ö. Uzuner, Y. Luo, and P. Szolovits, “Evaluating the state-of-the-art in automatic de-identification,” J. Amer. Med. Inform. Assoc. , vol. 14, no. 5, pp. 550–563, 2007
2007
Earlier work this paper cites.
K. Kipper, A. Korhonen, N. Ryant, and M. Palmer, “A large-scale classification of english verbs,” Lang. Resour. Eval. , vol. 42, no. 1, pp. 21–40, 2008
2008
Earlier work this paper cites.
K. Bollacker, C. Evans, P. Paritosh, T. Sturge, and J. Taylor, “Freebase: a collaboratively created graph database for structuring human knowledge,” in Proc. ACM SIGMOD Int. Conf. Manage. Data , 2008, pp. 1247–1250
2008
Earlier work this paper cites.
L. Smith et al. , “Overview of biocreative ii gene mention recognition,” Genome Biol. , vol. 9, no. 2, pp. 1–19, 2008
2008
Earlier work this paper cites.
J. Scott and G. Marshall, A dictionary of sociology . Oxford University Press, USA, 2009
2009
Earlier work this paper cites.
A. Carlson, J. Betteridge, B. Kisiel, B. Settles, E. R. Hruschka, and T. M. Mitchell, “Toward an architecture for never-ending language learning,” in Proc. 24th AAAI Conf. Artif. Intell. , 2010
2010
Earlier work this paper cites.
Ö. Uzuner, B. R. South, S. Shen, and S. L. DuVall, “2010 i2b2/va challenge on concepts, assertions, and relations in clinical text,” J. Amer. Med. Inform. Assoc. , vol. 18, no. 5, pp. 552–556, 2011
2011
Earlier work this paper cites.
S. S. Dasgupta, S. N. Ray, and P. Talukdar, “Hyte: Hyperplane-based temporally aware knowledge graph embedding,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2018, pp. 2001–2011
2011
Earlier work this paper cites.
C. M. Meyer and I. Gurevych, “Wiktionary: A new rival for expert-built lexicons? exploring the possibilities of collaborative lexicography,” in Electronic Lexicography . Oxford University Press, 11 2012. [Online]. Available: https://doi.org/10.1093/acprof:oso/9780199654864.003.0013
2012
Earlier work this paper cites.
E. M. Van Mulligen et al. , “The eu-adr corpus: annotated drugs, diseases, targets, and their relationships,” J. Biomed. Inform. , vol. 45, no. 5, pp. 879–884, 2012
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Proc. Int. Conf. Neural Inf. Process. Syst , vol. 26, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
M. Herrero-Zazo, I. Segura-Bedmar, P. Martínez, and T. Declerck, “The ddi corpus: An annotated corpus with pharmacological substances and drug–drug interactions,” J. Biomed. Inform. , vol. 46, no. 5, pp. 914–920, 2013
2013
Earlier work this paper cites.
R. Socher et al. , “Recursive deep models for semantic compositionality over a sentiment treebank,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2013, pp. 1631–1642
2013
Earlier work this paper cites.
J. Berant, A. Chou, R. Frostig, and P. Liang, “Semantic parsing on freebase from question-answer pairs,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2013, pp. 1533–1544
2013
Earlier work this paper cites.
J. Pennington, R. Socher, and C. D. Manning, “Glove: Global vectors for word representation,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2014, pp. 1532–1543
2014
Earlier work this paper cites.
D. Vrandečić and M. Krötzsch, “Wikidata: a free collaborative knowledgebase,” Commun. ACM , vol. 57, no. 10, pp. 78–85, 2014
2014
Earlier work this paper cites.
R. I. Doğan, R. Leaman, and Z. Lu, “Ncbi disease corpus: a resource for disease name recognition and concept normalization,” J. Biomed. Inform. , vol. 47, pp. 1–10, 2014
2014
Earlier work this paper cites.
M. Pontiki, D. Galanis, J. Pavlopoulos, H. Papageorgiou, I. Androutsopoulos, and S. Manandhar, “Semeval-2014 task 4: Aspect based sentiment analysis,” in Proc. 8th Int Workshop Semantic Eval. , 2014
2014
Earlier work this paper cites.
X. Ling, S. Singh, and D. S. Weld, “Design challenges for entity linking,” Trans. Assoc. Comput. Linguistics , vol. 3, pp. 315–328, 2015
2015
Earlier work this paper cites.
A. Stubbs, C. Kotfila, and Ö. Uzuner, “Automated systems for the de-identification of longitudinal clinical narratives: Overview of 2014 i2b2/uthealth shared task track 1,” J. Biomed. Inform. , vol. 58, pp. S11–S19, 2015
2015
Earlier work this paper cites.
À. Bravo, J. Piñero, N. Queralt-Rosinach, M. Rautschka, and L. I. Furlong, “Extraction of relations between genes and diseases from text and large-scale data analysis: implications for translational research,” BMC Bioinformatics , vol. 16, no. 1, pp. 1–17, 2015
2015
Earlier work this paper cites.
X. Zhang, J. Zhao, and Y. LeCun, “Character-level convolutional networks for text classification,” in Proc. Int. Conf. Neural Inf. Process. Syst. , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang, “Squad: 100,000+ questions for machine comprehension of text,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2016, pp. 2383–2392
2016
Earlier work this paper cites.
J. Li et al. , “Biocreative v cdr task corpus: A resource for chemical disease relation extraction,” Database(Oxford) , vol. 2016, 2016
2016
Earlier work this paper cites.
G.-A. Levow, “The third international chinese language processing bakeoff: Word segmentation and named entity recognition,” in Proc. Fifth SIGHAN Workshop Chinese Lang. Process. , 2016, pp. 108–117
2016
Earlier work this paper cites.
X. Li, A. Taheri, L. Tu, and K. Gimpel, “Commonsense knowledge base completion,” in Proc. 57th Annu. Meeting Assoc. Comput. Linguistics , 2016, pp. 1445–1455
2016
Earlier work this paper cites.
N. Mostafazadeh et al. , “A corpus and cloze evaluation for deeper understanding of commonsense stories,” in Proc. Conference North Amer. Chapter Assoc. Comput. Linguistics , 2016, pp. 839–849
2016
Earlier work this paper cites.
T. Nguyen, M. Rosenberg, X. Song, J. Gao, S. Tiwary, R. Majumder, and L. Deng, “Ms marco: A human generated machine reading comprehension dataset,” in Proc. Workshop on Cognitive Comput.: Integrating neural and symbolic approaches 2016 co-located with the 30th Annu. Conf. Neural Inf. Process. Syst. , 2016. [Online]. Available: http://ceur-ws.org/Vol-1773/CoCoNIPS_2016_paper9.pdf
2016
Earlier work this paper cites.
A. Vaswani et al. , “Attention is all you need,” in Proc. Int. Conf. Neural Inf. Process. Syst. , 2017, pp. 6000–6010
2017
Earlier work this paper cites.
R. Speer, J. Chin, and C. Havasi, “Conceptnet 5.5: An open multilingual graph of general knowledge,” in Proc. 31th AAAI Conf. Artif. Intell. , 2017
2017
Earlier work this paper cites.
B. McCann, J. Bradbury, C. Xiong, and R. Socher, “Learned in translation: Contextualized word vectors,” in Proc. 30th Int. Conf. Neural Inf. Process. Syst. , 2017, pp. 6294–6305
2017
Earlier work this paper cites.
B. Xu, Y. Xu, J. Liang, C. Xie, B. Liang, W. Cui, and Y. Xiao, “Cn-dbpedia: A never-ending chinese knowledge extraction system,” in Int. Conf. Ind., Eng. and Other Appl. of Appl. Intell. Syst. Springer, 2017, pp. 428–438
2017
Earlier work this paper cites.
L. Derczynski, E. Nichols, M. van Erp, and N. Limsopatham, “Results of the wnut2017 shared task on novel and emerging entity recognition,” in Proc. 3rd Workshop Noisy User-generated Text , 2017, pp. 140–147
2017
Earlier work this paper cites.
Y. Zhang, V. Zhong, D. Chen, G. Angeli, and C. D. Manning, “Position-aware attention and supervised data improve slot filling,” in Proc. 2017 Conf. Empir. Methods Lang. Process. , 2017, pp. 35–45
2017
Earlier work this paper cites.
M. Krallinger et al. , “Overview of the biocreative vi chemical-protein interaction track,” in Proc. sixth BioCreative Challenge Eval. Workshop , vol. 1, 2017, pp. 141–146
2017
Earlier work this paper cites.
M. Joshi, E. Choi, D. S. Weld, and L. Zettlemoyer, “Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension,” in Proc. 55th Annu. Meeting Assoc. Comput. Linguistics , 2017, pp. 1601–1611
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
M. E. Peters et al. , “Deep contextualized word representations,” in Proc. Conference North Amer. Chapter Assoc. Comput. Linguistics: Hum. Lang. Technol. , New Orleans, Louisiana, 2018, pp. 2227–2237
2018
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever, “Improving language understanding by generative pre-training,” OpenAI Blog , 2018. [Online]. Available: https://cdn.openai.com/research-covers/language-unsupervised/language_understanding_paper.pdf
2018
Earlier work this paper cites.
H. Elsahar et al. , “T-rex: A large scale alignment of natural language with knowledge base triples,” in Proc. 11th Int. Conf. Lang. Resour. Eval. , 2018. [Online]. Available: https://aclanthology.org/L18-1544.pdf
2018
Earlier work this paper cites.
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. Bowman, “Glue: A multi-task benchmark and analysis platform for natural language understanding,” in Proc. EMNLP Workshop BlackboxNLP , 2018, pp. 353–355
2018
Earlier work this paper cites.
E. Choi, O. Levy, Y. Choi, and L. Zettlemoyer, “Ultra-fine entity typing,” in Proc. 56th Annu. Meeting Assoc. Comput. Linguistics , 2018, pp. 87–96
2018
Earlier work this paper cites.
X. Han et al. , “Fewrel: A large-scale supervised few-shot relation classification dataset with state-of-the-art evaluation,” in Proc. 2018 Conf. Empir. Methods Natural Lang. Process. , 2018, pp. 4803–4809
2018
Earlier work this paper cites.
T. Mihaylov, P. Clark, T. Khot, and A. Sabharwal, “Can a suit of armor conduct electricity? a new dataset for open book question answering,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2018, pp. 2381–2391
2018
Earlier work this paper cites.
T. Dettmers, P. Minervini, P. Stenetorp, and S. Riedel, “Convolutional 2d knowledge graph embeddings,” in Proc. 32th AAAI Conf. Artif. Intell. , 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
K. Zhou, S. Prabhumoye, and A. W. Black, “A dataset for document grounded conversations,” in Proc. 2018 Conf. Empir. Methods Natural Lang. Process. , 2018, pp. 708–713
2018
Earlier work this paper cites.
Y. Cheng, D. Wang, P. Zhou, and T. Zhang, “Model compression and acceleration for deep neural networks: The principles, progress, and challenges,” IEEE Signal Process. Mag. , vol. 35, no. 1, pp. 126–136, 2018
2018
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proc. Conf. North Amer. Chapter Assoc. Comput. Linguistics: Hum. Lang. Technol. , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
B. Y. Lin, X. Chen, J. Chen, and X. Ren, “Kagnet: Knowledge-aware graph networks for commonsense reasoning,” in Proc. 2019 Conf. Empir. Methods Natural Lang. Process. and 9th Int. Joint Conf. Natural Lang. Process. , 2019, pp. 2829–2839
2019
Earlier work this paper cites.
A. Bosselut, H. Rashkin, M. Sap, C. Malaviya, A. Celikyilmaz, and Y. Choi, “Comet: Commonsense transformers for automatic knowledge graph construction,” in PProc. 57th Annu. Meeting Assoc. Comput. Linguistics , 2019, pp. 4762–4779
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Sap et al. , “Atomic: An atlas of machine commonsense for if-then reasoning,” in Proc. 33th AAAI Conf. Artif. Intell. , vol. 33, no. 01, 2019, pp. 3027–3035
2019
Cited alongside, same era.
2019
Cited alongside, same era.
I. Beltagy, K. Lo, and A. Cohan, “Scibert: A pretrained language model for scientific text,” in EMNLP-IJCNLP , 2019, pp. 3613–3618
2019
Cited alongside, same era.
P. Zhong, D. Wang, and C. Miao, “Knowledge-enriched transformer for emotion detection in textual conversations,” in Proc. 2019 Conf. Empir. Methods Natural Lang. Process. and 9th Int. Joint Conf. Natural Lang. Process. , 2019, pp. 165–176
2019
Cited alongside, same era.
2020
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
M. Yasunaga, H. Ren, A. Bosselut, P. Liang, and J. Leskovec, “Qa-gnn: Reasoning with language models and knowledge graphs for question answering,” in Proc. 2021 Conf. North Amer. Chapter Assoc. Comput. Linguistics: Hum. Lang. Technol. , 2021, pp. 535–546
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Z. Zhang, X. Han, Z. Liu, X. Jiang, M. Sun, and Q. Liu, “Ernie: Enhanced language representation with informative entities,” in Proc. 57th Annu. Meeting Assoc. Comput. Linguistics , 2019, pp. 1441–1451
2019
Cited alongside, same era.
M. E. Peters et al. , “Knowledge enhanced contextual word representations,” in Proc. 2019 Conf. Empir. Methods Natural Lang. Process. and 9th Int. Joint Conf. Natural Lang. Process. , 2019, pp. 43–54
2019
Cited alongside, same era.
2019
Cited alongside, same era.
A. Yang, Q. Wang, J. Liu, K. Liu, Y. Lyu, H. Wu, Q. She, and S. Li, “Enhancing pre-trained language representations with rich knowledge for machine reading comprehension,” in Proc. 57th Annu. Meeting Assoc. Comput. Linguistics , 2019, pp. 2346–2357
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2021
Later among the works it cites.
Y. Liu, Y. Wan, L. He, H. Peng, and S. Y. Philip, “Kg-bart: Knowledge graph-augmented bart for generative commonsense reasoning,” in Proc. AAAI Conf. Artif. Intell. , vol. 35, no. 7, 2021, pp. 6418–6425
2021
Later among the works it cites.
A. Gajbhiye, N. A. Moubayed, and S. Bradley, “Exbert: An external knowledge enhanced bert for natural language inference,” in Int. Conf. Artif. Neural Netw. Springer, 2021, pp. 460–472
2021
Later among the works it cites.
T.-Y. Chang et al. , “Incorporating commonsense knowledge graph in pretrained models for social commonsense tasks,” in Proc. Deep Learn. Inside Out , 2021, pp. 74–79
2021
Later among the works it cites.
2021
Later among the works it cites.
J. D. Hwang et al. , “(comet-) atomic 2020: On symbolic and neural commonsense knowledge graphs,” in Proc. 35th AAAI Conf. Artif. Intell. , vol. 35, no. 7, 2021, pp. 6384–6392
2021
Later among the works it cites.
G. Michalopoulos, Y. Wang, H. Kaka, H. Chen, and A. Wong, “Umlsbert: Clinical domain knowledge augmentation of contextual embeddings using the unified medical language system metathesaurus,” in Proc. Conference North Amer. Chapter Assoc. Comput. Linguistics: Hum. Lang. Technol. , 2021, pp. 1744–1753
2021
Later among the works it cites.
T. Zhang, Z. Cai, C. Wang, M. Qiu, B. Yang, and X. He, “Smedbert: A knowledge-enhanced pre-trained language model with structured semantics for medical text mining,” in Proc. 59th Annu. Meeting Assoc. Comput. Linguistics and 11th Int. Joint Conf. Artif. Intell. , 2021, pp. 5882–5893
2021
Later among the works it cites.
Z. Yuan, Y. Liu, C. Tan, S. Huang, and F. Huang, “Improving biomedical pretrained language models with knowledge,” in Proc. 20th Workshop Biomed. Lang. Process , 2021, pp. 180–190
2021
Later among the works it cites.
W. Nicholas, T. Amalie, H. Haoyan, L. Sanghoon, C. Kevin, D. John, D. Alexander, P. Kristin, C. Gerbrand, and J. Anubhav, “The impact of domain-specific pre-training on named entity recognition tasks in materials science,” in http://dx.doi.org/10.2139/ssrn.3950755 , 2021, pp. 1–43
2021
Later among the works it cites.
F. Sun, F.-L. Li, R. Wang, Q. Chen, X. Cheng, and J. Zhang, “K-aid: Enhancing pre-trained language models with domain knowledge for question answering,” in Proc. 30th ACM Int. Conf. Inf. Knowl. Manage , 2021, pp. 4125–4134
2021
Later among the works it cites.
S. Xu et al. , “K-plug: Knowledge-injected pre-trained language model for natural language understanding and generation in e-commerce,” in Findings Assoc. Comput. Linguistics , 2021, pp. 1–17
2021
Later among the works it cites.
L. Zheng, N. Guha, B. R. Anderson, P. Henderson, and D. E. Ho, “When does pretraining help?: assessing self-supervised learning for law and the casehold dataset of 53, 000+ legal holdings,” in ICAIL , 2021, pp. 159–168
2021
Later among the works it cites.
C. Xiao, X. Hu, Z. Liu, C. Tu, and M. Sun, “Lawformer: A pre-trained language model for chinese legal long documents,” in AI Open , 2021, pp. 79–84
2021
Later among the works it cites.
H. Yao, Y. Chen, Q. Ye, X. Jin, and X. Ren, “Refining language models with compositional explanations,” in Proc. 34th Int. Conf. Neural Inf. Process. Syst. , vol. 34, 2021, pp. 8954–8967
2021
Later among the works it cites.
D. Guo, S. Ren, S. Lu, Z. Feng, D. Tang, S. Liu, L. Zhou, N. Duan, A. Svyatkovskiy, S. Fu, M. Tufano, S. K. Deng, C. B. Clement, D. Drain, N. Sundaresan, J. Yin, D. Jiang, and M. Zhou, “Graphcodebert: Pre-training code representations with data flow,” in ICLR , 2021, pp. 1–18
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
K. Basu, S. C. Varanasi, F. Shakerin, J. Arias, and G. Gupta, “Knowledge-driven natural language understanding of english text and its applications,” in Proc. 35th AAAI Conf. Artif. Intell. , vol. 35, no. 14, 2021, pp. 12 554–12 563
2021
Later among the works it cites.
Y. Qin et al. , “Erica: Improving entity and relation understanding for pre-trained language models via contrastive learning,” in Proc. 59th Annu. Meeting Assoc. Comput. Linguistics and 11th Int. Joint Conf. Natural Lang. Process. , 2021, pp. 3350–3363
2021
Later among the works it cites.
X. Wang et al. , “Kepler: A unified model for knowledge embedding and pre-trained language representation,” Trans. Assoc. Comput. Linguistics , vol. 9, pp. 176–194, 2021
2021
Later among the works it cites.
Y. Su et al. , “Cokebert: Contextual knowledge selection and embedding towards enhanced pre-trained language models,” AI Open , vol. 2, pp. 127–134, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
N. Bian, X. Han, B. Chen, and L. Sun, “Benchmarking knowledge-enhanced commonsense question answering via knowledge-to-text transformation,” in Proc. 35th AAAI Conf. Artif. Intell. , vol. 35, no. 14, 2021, pp. 12 574–12 582
2021
Later among the works it cites.
D. Wilmot and F. Keller, “Memory and knowledge augmented language models for inferring salience in long-form stories,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2021, pp. 851–865
2021
Later among the works it cites.
2021
Later among the works it cites.
F. Petroni et al. , “Kilt: A benchmark for knowledge intensive language tasks,” in Proc. Conference North Amer. Chapter Assoc. Comput. Linguistics: Hum. Lang. Technol. , 2021, pp. 2523–2544
2021
Later among the works it cites.
2021
Later among the works it cites.
H. Schuff, H.-Y. Yang, H. Adel, and N. T. Vu, “Does external knowledge help explainable natural language inference? automatic evaluation vs. human ratings,” in Proc. 4th BlackboxNLP Workshop Analyzing and interpreting Neural Netw for NLP , 2021, pp. 26–41
2021
Later among the works it cites.
N. De Cao, W. Aziz, and I. Titov, “Editing factual knowledge in language models,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2021, pp. 6491–6506
2021
Later among the works it cites.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
X. Zhang, A. Bosselut, M. Yasunaga, H. Ren, P. Liang, C. D. Manning, and J. Leskovec, “Greaselm: Graph reasoning enhanced language models,” in ICLR , 2022, pp. 1–16
2022
Closest in time.
P. Hosseini, D. A. Broniatowski, and M. Diab, “Knowledge-augmented language models for cause-effect relation classification,” in Proc. First Workshop on Commonsense Representation Reasoning , 2022, pp. 43–48
2022
Closest in time.
C. Yu, H. Zhang, Y. Song, and W. Ng, “Cocolm: Complex commonsense enhanced language model with discourse relations,” in Findings Assoc. Comput. Linguistics , 2022, pp. 1175–1187
2022
Closest in time.
N. Liu, Q. Hu, H. Xu, X. Xu, and M. Chen, “Med-bert: A pretraining framework for medical records named entity recognition,” IEEE Trans. Ind. Informatics , vol. 18, no. 8, pp. 5600–5608, 2022
2022
Closest in time.
2022
Closest in time.
X. Jiang, Y. Liang, W. Chen, and N. Duan, “Xlm-k: Improving cross-lingual language model pre-training with multilingual knowledge,” in Proc. 36th AAAI Conf. Artif. Intell. , vol. 36, no. 10, 2022, pp. 10 840–10 848
2022
Closest in time.
Q. Liu, D. Yogatama, and P. Blunsom, “Relational memory-augmented language models,” Trans. Assoc. Comput. Linguistics , vol. 10, pp. 555–572, 2022
2022
Closest in time.
2022
Closest in time.
W. Wang et al. , “Visually-augmented language modeling,” arXiv preprint arXiv:2205.10178 , 2022
2022
Closest in time.
T. Zhang et al. , “Dkplm: Decomposable knowledge-enhanced pre-trained language model for natural language understanding,” in Proc. 36th AAAI Conf. Artif. Intell. , vol. 36, no. 10, 2022, pp. 11 703–11 711
2022
Closest in time.
D. Yu, C. Zhu, Y. Yang, and M. Zeng, “Jaket: Joint pre-training of knowledge graph and language understanding,” in Proc. 36th AAAI Conf. Artif. Intell. , vol. 36, no. 10, 2022, pp. 11 630–11 638
2022
Closest in time.
Z. Meng, F. Liu, E. Shareghi, Y. Su, C. Collins, and N. Collier, “Rewire-then-probe: A contrastive recipe for probing biomedical knowledge of pre-trained language models,” in Proc. 60th Annu. Meeting Assoc. Comput. Linguistics , 2022, pp. 4798–4810
2022
Closest in time.
2022
Closest in time.
S. Li, M. Sridhar, C. S. Prakash, J. Cao, W. Hamza, and J. McAuley, “Instilling type knowledge in language models via multi-task qa,” in Findings Assoc. Comput. Linguistics: NAACL , 2022, pp. 594–603
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
S. Hu, N. Ding, H. Wang, Z. Liu, J. Wang, J. Li, W. Wu, and M. Sun, “Knowledgeable prompt-tuning: Incorporating knowledge into prompt verbalizer for text classification,” in ACL , 2022, pp. 2225–2240
2022
Closest in time.
X. Chen, N. Zhang, X. Xie, S. Deng, Y. Yao, C. Tan, F. Huang, L. Si, and H. Chen, “Knowprompt: Knowledge-aware prompt-tuning with synergistic optimization for relation extraction,” in WWW , 2022, pp. 2778–2788
2022
Closest in time.
H. Ye, N. Zhang, S. Deng, X. Chen, H. Chen, F. Xiong, X. Chen, and H. Chen, “Ontology-enhanced prompt-tuning for few-shot learning,” in WWW , 2022, pp. 778–787
2022
Closest in time.
B. Ryan, D. Minh-Hoang, H. Fabian, H. Yuan, M.-P. Albert, and S. Vijay, “Improving language model predictions via prompts enriched with knowledge graphs,” in ISWC , 2022, pp. 1–10
2022
Closest in time.
2022
Closest in time.
H. Tan and M. Bansal, “Vokenization: Improving language understanding with contextualized, visual-grounded supervision,” in Proc. Conf. Empir. Methods Natural Lang. Process. , 2020, pp. 2066–2080
2080
Closest in time.