Fetching the paper…
Reading the bibliography…
Large-scale pre-trained language models have demonstrated strong knowledge representation ability.
Ernie: Enhanced representation through knowledge integration
Yu Sun, Shuohuan Wang, Yukun Li, Shikun Feng, Xuyi Chen, Han Zhang, Xin Tian, Danxiang Zhu, Hao Tian, and Hua Wu. 2019 · 1904
Earlier work this paper cites.
Pre-training with whole word masking for chinese bert
Yiming Cui, Wanxiang Che, Ting Liu, Bing Qin, Ziqing Yang, Shijin Wang, and Guoping Hu. 2019 · 1906
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Kg-bert: Bert for knowledge graph completion
Liang Yao, Chengsheng Mao, and Yuan Luo. 2019 · 1909
Earlier work this paper cites.
Limit-bert: Linguistic informed multi-task bert
Junru Zhou, Zhuosheng Zhang, and Hai Zhao. 2019 · 1910
Earlier work this paper cites.
The algebra of events
Emmon Bach. 1986 · 1986
Earlier work this paper cites.
K-adapter: Infusing knowledge into pre-trained models with adapters
Ruize Wang, Duyu Tang, Nan Duan, Zhongyu Wei, Xuanjing Huang, Cuihong Cao, Daxin Jiang, Ming Zhou, et al. 2020b · 2002
Earlier work this paper cites.
Don’t stop pretraining: Adapt language models to domains and tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A. Smith. 2020 · 2004
Earlier work this paper cites.
How context affects language models’ factual predictions
Fabio Petroni, Patrick Lewis, Aleksandra Piktus, Tim Rocktäschel, Yuxiang Wu, Alexander H Miller, and Sebastian Riedel. 2020 · 2005
Earlier work this paper cites.
Conversational neuro-symbolic commonsense reasoning
Forough Arabshahi, Jennifer Lee, Mikayla Gawarecki, Kathryn Mazaitis, Amos Azaria, and Tom Mitchell. 2020 · 2006
Earlier work this paper cites.
Facts as experts: Adaptable and interpretable neural memory over symbolic knowledge
Pat Verga, Haitian Sun, Livio Baldini Soares, and William W. Cohen. 2020 · 2007
Earlier work this paper cites.
Unsupervised learning of narrative event chains
Nathanael Chambers and Dan Jurafsky. 2008 · 2008
Earlier work this paper cites.
Comet-atomic 2020: On symbolic and neural commonsense knowledge graphs
Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras, Jeff Da, Keisuke Sakaguchi, Antoine Bosselut, and Yejin Choi. 2020 · 2010
Earlier work this paper cites.
Jaket: Joint pre-training of knowledge graph and language understanding
Donghan Yu, Chenguang Zhu, Yiming Yang, and Michael Zeng. 2020 · 2010
Earlier work this paper cites.
SemEval-2012 task 7: Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Andrew Gordon, Zornitsa Kozareva, and Melissa Roemmele. 2012 · 2012
Earlier work this paper cites.
Deepwalk: online learning of social representations
Bryan Perozzi, Rami Al-Rfou, and Steven Skiena. 2014 · 2014
Earlier work this paper cites.
A corpus and cloze evaluation for deeper understanding of commonsense stories
Nasrin Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James Allen. 2016 · 2016
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Open-domain event detection using distant supervision
Jun Araki and Teruko Mitamura. 2018 · 2018
Earlier work this paper cites.
Automatic prediction of discourse connectives
Eric Malmi, Daniele Pighin, Sebastian Krause, and Mikhail Kozhevnikov. 2018 · 2018
Earlier work this paper cites.
Tackling the story ending biases in the story cloze test
Rishi Sharma, James Allen, Omid Bakhshandeh, and Nasrin Mostafazadeh. 2018 · 2018
Cited alongside, same era.
Comet: Commonsense transformers for automatic knowledge graph construction
Antoine Bosselut, Hannah Rashkin, Maarten Sap, Chaitanya Malaviya, Asli Celikyilmaz, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Event representation learning enhanced with external commonsense knowledge
Xiao Ding, Kuo Liao, Ting Liu, Zhongyang Li, and Junwen Duan. 2019 · 2019
Cited alongside, same era.
Story ending prediction by transferable bert
Zhongyang Li, Xiao Ding, and Ting Liu. 2019 · 2019
Cited alongside, same era.
KagNet: Knowledge-aware graph networks for commonsense reasoning
Specializing unsupervised pretraining models for word-level semantic similarity
Anne Lauscher, Ivan Vulić, Edoardo Maria Ponti, Anna Korhonen, and Goran Glavaš. 2020 · 2020
Closest in time.
SenseBERT: Driving some sense into BERT
Yoav Levine, Barak Lenz, Or Dagan, Ori Ram, Dan Padnos, Or Sharir, Shai Shalev-Shwartz, Amnon Shashua, and Yoav Shoham. 2020 · 2020
Closest in time.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Closest in time.
K-BERT: enabling language representation with knowledge graph
Weijie Liu, Peng Zhou, Zhe Zhao, Zhiruo Wang, Qi Ju, Haotang Deng, and Ping Wang. 2020 · 2020
Closest in time.
Integrating external event knowledge for script learning
Shangwen Lv, Fuqing Zhu, and Songlin Hu. 2020 · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bill Yuchen Lin, Xinyue Chen, Jamin Chen, and Xiang Ren. 2019 · 2019
Cited alongside, same era.
DisSent: Learning sentence representations from explicit discourse relations
Allen Nie, Erin Bennett, and Noah Goodman. 2019 · 2019
Cited alongside, same era.
An improved neural baseline for temporal relation extraction
Qiang Ning, Sanjay Subramanian, and Dan Roth. 2019 · 2019
Cited alongside, same era.
Knowledge enhanced contextual word representations
Matthew E Peters, Mark Neumann, Robert Logan, Roy Schwartz, Vidur Joshi, Sameer Singh, and Noah A Smith. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick S. H. Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander H. Miller. 2019 · 2019
Cited alongside, same era.
Atomic: An atlas of machine commonsense for if-then reasoning
Maarten Sap, Ronan Le Bras, Emily Allaway, Chandra Bhagavatula, Nicholas Lourie, Hannah Rashkin, Brendan Roof, Noah A Smith, and Yejin Choi. 2019 · 2019
Cited alongside, same era.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2019 · 2019
Cited alongside, same era.
Nasrin Mostafazadeh, Aditya Kalyanpur, Lori Moon, David Buchanan, Lauren Berkowitz, Or Biran, and Jennifer Chu-Carroll. 2020 · 2020
Closest in time.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Closest in time.
Exploiting structured knowledge in text via graph-guided representation learning
Tao Shen, Yi Mao, Pengcheng He, Guodong Long, Adam Trischler, and Weizhu Chen. 2020 · 2020
Closest in time.
SKEP: Sentiment knowledge enhanced pre-training for sentiment analysis
Hao Tian, Can Gao, Xinyan Xiao, Hao Liu, Bolei He, Hua Wu, Haifeng Wang, and Feng Wu. 2020 · 2020
Closest in time.
Connecting the dots: A knowledgeable path generator for commonsense question answering
Peifeng Wang, Nanyun Peng, Filip Ilievski, Pedro Szekely, and Xiang Ren. 2020a · 2020
Closest in time.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Closest in time.
Pretrained encyclopedia: Weakly supervised knowledge-pretrained language model
Wenhan Xiong, Jingfei Du, William Yang Wang, and Veselin Stoyanov. 2020 · 2020
Closest in time.
Aser: A large-scale eventuality knowledge graph
Hongming Zhang, Xin Liu, Haojie Pan, Yangqiu Song, and Cane Wing-Ki Leung. 2020b · 2020
Closest in time.
Temporal common sense acquisition with minimal supervision
Ben Zhou, Qiang Ning, Daniel Khashabi, and D. Roth. 2020 · 2020
Closest in time.
Discos: Bridging the gap between discourse knowledge and commonsense knowledge
Tianqing Fang, Hongming Zhang, Weiqi Wang, Y. Song, and Bin He. 2021 · 2021
Closest in time.
Paragraph-level commonsense transformers with recurrent memory
Saadia Gabriel, Chandra Bhagavatula, Vered Shwartz, Ronan Le Bras, Maxwell Forbes, and Yejin Choi. 2021 · 2021
Closest in time.
Conditional generation of temporally-ordered event sequences
Shih-Ting Lin, Nathanael Chambers, and Greg Durrett. 2021 · 2021
Closest in time.
Relational World Knowledge Representation in Contextual Language Models: A Review
Tara Safavi and Danai Koutra. 2021 · 2021
Closest in time.
Masked language modeling and the distributional hypothesis: Order word matters pre-training for little
Koustuv Sinha, Robin Jia, Dieuwke Hupkes, Joelle Pineau, Adina Williams, and Douwe Kiela. 2021 · 2021
Closest in time.
Yucheng Zhou, Tao Shen, Xiubo Geng, Guodong Long, and Daxin Jiang. 2022 · 2022
Closest in time.