Fetching the paper…
Reading the bibliography…
We introduce S2ORC, a large corpus of 81.1M English-language academic papers spanning many academic disciplines.
Co-citation in the scientific literature: A new measure of the relationship between two documents
Henry Small. 1973 · 1973
Earlier work this paper cites.
Citeseer: an automatic citation indexing system
C. L. Giles, K. D. Bollacker, and S. Lawrence. 1998 · 1998
Earlier work this paper cites.
Genia corpus—a semantically annotated corpus for bio-textmining
J-D Kim, Tomoko Ohta, Yuka Tateisi, and Jun’ichi Tsujii. 2003 · 2003
Earlier work this paper cites.
Introduction to the bio-entity recognition task at JNLPBA
Nigel Collier and Jin-Dong Kim. 2004 · 2004
Earlier work this paper cites.
Automatic classification of citation function
Simone Teufel, Advaith Siddharthan, and Dan Tidhar. 2006 · 2006
Earlier work this paper cites.
The ACL anthology reference corpus: A reference dataset for bibliographic research in computational linguistics
Steven Bird, Robert Dale, Bonnie Dorr, Bryan Gibson, Mark Joseph, Min-Yen Kan, Dongwon Lee, Brett Powley, Dragomir Radev, and Yee Fan Tan. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey E. Hinton. 2008 · 2008
Earlier work this paper cites.
Scientific paper summarization using citation summary networks
Vahed Qazvinian and Dragomir R. Radev. 2008 · 2008
Earlier work this paper cites.
Arnetminer: Extraction and mining of academic social networks
Jie Tang, Jing Zhang, Limin Yao, Juanzi Li, Li Zhang, and Zhong Su. 2008 · 2008
Earlier work this paper cites.
Grobid: Combining automatic bibliographic data recognition and term extraction for scholarship publications
Patrice Lopez. 2009 · 2009
Earlier work this paper cites.
The acl anthology network corpus
Dragomir R. Radev, Pradeep Muthukrishnan, and Vahed Qazvinian. 2009 · 2009
Earlier work this paper cites.
Context-aware citation recommendation
Qi He, Jian Pei, Daniel Kifer, Prasenjit Mitra, and Lee Giles. 2010 · 2010
Earlier work this paper cites.
Context-enhanced citation sentiment detection
Awais Athar and Simone Teufel. 2012 · 2012
Earlier work this paper cites.
Citation prediction in heterogeneous bibliographic networks
Xiao Yu, Quanquan Gu, Mianwei Zhou, and Jiawei Han. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Citation-enhanced keyphrase extraction from research papers: A supervised approach
Cornelia Caragea, Florin Adrian Bulgarov, Andreea Godea, and Sujatha Das Gollapalli. 2014 · 2014
Earlier work this paper cites.
Content-based citation analysis: The next generation of citation analysis
Ying Ding, Guo Zhang, Tamy Chambers, Min Song, Xiaolong Wang, and Cheng xiang Zhai. 2014 · 2014
Earlier work this paper cites.
Ncbi disease corpus: a resource for disease name recognition and concept normalization
Rezarta Islamaj Doğan, Robert Leaman, and Zhiyong Lu. 2014 · 2014
Earlier work this paper cites.
Citation resolution: A method for evaluating context-based citation recommendation systems
Daniel Duma and Ewan Klein. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Scientific article summarization using citation-context and article’s discourse structure
Arman Cohan and Nazli Goharian. 2015 · 2015
Cited alongside, same era.
A neural probabilistic model for context based citation recommendation
Wenyi Huang, Zhaohui Wu, Chen Liang, Prasenjit Mitra, and C. Lee Giles. 2015 · 2015
Cited alongside, same era.
Context-based collaborative filtering for citation recommendation
Haifeng Liu, Xiangjie Kong, Xiaomei Bai, Wei Wang, Teshome Megersa Bekele, and Feng Xia. 2015 · 2015
Cited alongside, same era.
Summarizing citation contexts of scientific publications
A corpus with multi-level annotations of patients, interventions and outcomes to support language processing for medical literature
Benjamin Nye, Junyi Jessy Li, Roma Patel, Yinfei Yang, Iain Marshall, Ani Nenkova, and Byron Wallace. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018a · 2018
Later among the works it cites.
Dissecting contextual word embeddings: Architecture and representation
Matthew Peters, Mark Neumann, Luke Zettlemoyer, and Wen-tau Yih. 2018b · 2018
Later among the works it cites.
A web-scale system for scientific knowledge exploration
Zhihong Shen, Hao Ma, and Kuansan Wang. 2018 · 2018
Later among the works it cites.
Machine learning vs. rules and out-of-the-box vs. retrained: An evaluation of open-source bibliographic reference and citation parsers
Dominika Tkaczyk, Andrew Collins, Paraic Sheridan, and Joeran Beel. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sandra Mitrović and Henning Müller. 2015 · 2015
Cited alongside, same era.
Identifying meaningful citations
Marco Valenzuela, Vu Ha, and Oren Etzioni. 2015 · 2015
Cited alongside, same era.
cite2vec: Citation-driven document exploration via word embeddings
Matthew Berger, Katherine McDonough, and Lee M Seversky. 2016 · 2016
Cited alongside, same era.
Biocreative V CDR task corpus: a resource for chemical disease relation extraction
Jiao Li, Yueping Sun, Robin J Johnson, Daniela Sciaky, Chih-Hsuan Wei, Robert Leaman, Allan Peter Davis, Carolyn J Mattingly, Thomas C Wiegers, and Zhiyong Lu. 2016 · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Gregory S. Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
Cited alongside, same era.
Overview of the biocreative vi chemical-protein interaction track
Martin Krallinger, Obdulia Rabal, Saber Ahmad Akhondi, Martín Pérez Pérez, Jésús López Santamaría, Gael Pérez Rodríguez, Georgios Tsatsaronis, Ander Intxaurrondo, José Antonio Baso López, Umesh Nandal, Erin M. van Buel, A. Poorna Chandrasekhar, Marleen Rodenburg, Astrid Lægreid, Marius A. Doornenbal, Julen Oyarzábal, Anália Lourenço, and Alfonso Valencia. 2017 · 2017
Cited alongside, same era.
Cad: an algorithm for citation-anchors detection in research papers
Riaz Ahmad and Muhammad Tanvir Afzal. 2018 · 2018
Cited alongside, same era.
Document co-citation analysis to enhance transdisciplinary research
Caleb M. Trujillo and Tammy M. Long. 2018 · 2018
Later among the works it cites.
SciBERT: A pretrained language model for scientific text
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Closest in time.
Structural scaffolds for citation intent classification in scientific publications
Arman Cohan, Waleed Ammar, Madeleine van Zuylen, and Field Cady. 2019 · 2019
Closest in time.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.
Extended co-citation search: Graph-based document retrieval on a co-citation network containing citation context information
Masaki Eto. 2019 · 2019
Closest in time.
A context-aware citation recommendation model with bert and graph convolutional networks
Chanwoo Jeong, Sion Jang, Hyuna Shin, Eunjeong Park, and Sungchul Choi. 2019 · 2019
Closest in time.
A scalable hybrid research paper recommender system for microsoft academic
Anshul Kanakia, Zhihong Shen, Darrin Eide, and Kuansan Wang. 2019 · 2019
Closest in time.
Linguistic knowledge and transferability of contextual representations
Nelson F. Liu, Matt Gardner, Yonatan Belinkov, Matthew E. Peters, and Noah A. Smith. 2019a · 2019
Closest in time.
ScispaCy: Fast and robust models for biomedical natural language processing
Mark Neumann, Daniel King, Iz Beltagy, and Waleed Ammar. 2019 · 2019
Closest in time.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Closest in time.
Bibliometric-enhanced arxiv: A data set for paper-based and citation-based tasks
Tarek Saier and Michael Färber. 2019 · 2019
Closest in time.
unarxive: a large scholarly data set with publications’ full-text, annotated in-text citations, and links to metadata
Tarek Saier and Michael Färber. 2020 · 2020
Closest in time.
CORD-19: The Covid-19 Open Research Dataset
Lucy Lu Wang, Kyle Lo, Yoganand Chandrasekhar, Russell Reas, Jiangjiang Yang, Darrin Eide, Kathryn Funk, Rodney Kinney, Ziyang Liu, William Merrill, Paul Mooney, Dewey Murdick, Devvret Rishi, Jerry Sheehan, Zhihong Shen, Brandon Stilson, Alex D. Wade, Kuansan Wang, Chris Wilhelm, Boya Xie, Douglas Raymond, Daniel S. Weld, Oren Etzioni, and Sebastian Kohlmeier. 2020 · 2020
Closest in time.