Fetching the paper…
Reading the bibliography…
We present ACL OCL, a scholarly corpus derived from the ACL Anthology to assist Open scientific research in the Computational Linguistics domain.
Studying the evolution of scientific topics and their relationships
Ana Sabina Uban, Cornelia Caragea, and Liviu P. Dinu. 2021 · 1922
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Minimum error rate training in statistical machine translation
Franz Josef Och. 2003 · 2003
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
EMNLP versus ACL: Analyzing NLP research over time
Sujatha Das Gollapalli and Xiaoli Li. 2015 · 2006
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, Chris Dyer, Ondřej Bojar, Alexandra Constantin, and Evan Herbst. 2007 · 2007
Earlier work this paper cites.
The ACL Anthology reference corpus: A reference dataset for bibliographic research in computational linguistics
Steven Bird, Robert Dale, Bonnie Dorr, Bryan Gibson, Mark Joseph, Min-Yen Kan, Dongwon Lee, Brett Powley, Dragomir Radev, and Yee Fan Tan. 2008 · 2008
Earlier work this paper cites.
Studying the history of ideas using topic models
David Hall, Daniel Jurafsky, and Christopher D. Manning. 2008 · 2008
Earlier work this paper cites.
Graph-based keyword extraction for single-document summarization
Marina Litvak and Mark Last. 2008 · 2008
Earlier work this paper cites.
The ACL Anthology network
Dragomir R. Radev, Pradeep Muthukrishnan, and Vahed Qazvinian. 2009 · 2009
Earlier work this paper cites.
Towards automated related work summarization
Cong Duy Vu Hoang and Min-Yen Kan. 2010 · 2010
Earlier work this paper cites.
Language detection library for java
Nakatani Shuyo. 2010 · 2010
Earlier work this paper cites.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Automatic generation of related work sections in scientific papers: An optimization approach
Yue Hu and Xiaojun Wan. 2014 · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
A neural probabilistic model for context based citation recommendation
Wenyi Huang, Zhaohui Wu, Chen Liang, Prasenjit Mitra, and C. Lee Giles. 2015 · 2015
Earlier work this paper cites.
Forecasting emerging trends from scientific literature
Kartik Asooja, Georgeta Bordea, Gabriela Vulcu, and Paul Buitelaar. 2016 · 2016
Cited alongside, same era.
Pdffigures 2.0: Mining figures from research papers
Christopher Clark and Santosh Divvala. 2016 · 2016
Cited alongside, same era.
Predicting the rise and fall of scientific topics from trends in their rhetorical framing
Vinodkumar Prabhakaran, William L. Hamilton, Dan McFarland, and Dan Jurafsky. 2016 · 2016
Cited alongside, same era.
Structural scaffolds for citation intent classification in scientific publications
Arman Cohan, Waleed Ammar, Madeleine van Zuylen, and Field Cady. 2019 · 2019
Cited alongside, same era.
Benchmarking zero-shot text classification: Datasets, evaluation and entailment approach
Wenpeng Yin, Jamaal Hay, and Dan Roth. 2019 · 2019
Cited alongside, same era.
TLDR: Extreme summarization of scientific documents
Bringing structure into summaries: a faceted summarization dataset for long scientific documents
Rui Meng, Khushboo Thaker, Lei Zhang, Yue Dong, Xingdi Yuan, Tong Wang, and Daqing He. 2021 · 2021
Later among the works it cites.
Survey of computational approaches to lexical semantic change detection
Nina Tahmasebi, Lars Borin, and Adam Jatowt. 2021 · 2021
Later among the works it cites.
Weakly-supervised text classification based on keyword graph
Lu Zhang, Jiandong Ding, Yi Xu, Yingyao Liu, and Shuigeng Zhou. 2021 · 2021
Later among the works it cites.
Automatic error analysis for document-level information extraction
Aliva Das, Xinya Du, Barry Wang, Kejian Shi, Jiayuan Gu, Thomas Porter, and Claire Cardie. 2022 · 2022
Later among the works it cites.
CSL: A large-scale Chinese scientific literature dataset
Yudong Li, Yuqing Zhang, Zhe Zhao, Linlin Shen, Weijie Liu, Weiquan Mao, and Hui Zhang. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Isabel Cachola, Kyle Lo, Arman Cohan, and Daniel Weld. 2020 · 2020
Cited alongside, same era.
YAKE! Keyword extraction from single documents using multiple local features
Ricardo Campos, Vítor Mangaravite, Arian Pasquali, Alípio Jorge, Célia Nunes, and Adam Jatowt. 2020 · 2020
Cited alongside, same era.
Autoencoding keyword correlation graph for document clustering
Billy Chiu, Sunil Kumar Sahu, Derek Thomas, Neha Sengupta, and Mohammady Mahdy. 2020 · 2020
Cited alongside, same era.
SPECTER: Document-level representation learning using citation-informed transformers
Arman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey, and Daniel Weld. 2020 · 2020
Cited alongside, same era.
SciREX: A challenge dataset for document-level information extraction
Sarthak Jain, Madeleine van Zuylen, Hannaneh Hajishirzi, and Iz Beltagy. 2020 · 2020
Cited alongside, same era.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
S2ORC: The semantic scholar open research corpus
Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney, and Daniel Weld. 2020 · 2020
Cited alongside, same era.
WANLI: Worker and AI collaboration for natural language inference dataset creation
Alisa Liu, Swabha Swayamdipta, Noah A. Smith, and Yejin Choi. 2022 · 2022
Later among the works it cites.
Learn to explain: Multimodal reasoning via thought chains for science question answering
Pan Lu, Swaroop Mishra, Tony Xia, Liang Qiu, Kai-Wei Chang, Song-Chun Zhu, Oyvind Tafjord, Peter Clark, and Ashwin Kalyan. 2022 · 2022
Later among the works it cites.
Some languages are more equal than others: Probing deeper into the linguistic disparity in the NLP world
Surangika Ranathunga and Nisansa de Silva. 2022 · 2022
Later among the works it cites.
Tarek Saier, Johan Krause, and Michael Färber. 2023 · 2022
Later among the works it cites.
Galactica: A large language model for science
Ross Taylor, Marcin Kardas, Guillem Cucurull, Thomas Scialom, Anthony Hartshorn, Elvis Saravia, Andrew Poulton, Viktor Kerkez, and Robert Stojnic. 2022 · 2022
Later among the works it cites.
What factors should paper-reviewer assignments rely on? community perspectives on issues and ideals in conference peer-review
Terne Thorn Jakobsen and Anna Rogers. 2022 · 2022
Later among the works it cites.
Incorporating hierarchy into text encoder: a contrastive learning approach for hierarchical text classification
Zihan Wang, Peiyi Wang, Lianzhe Huang, Xin Sun, and Houfeng Wang. 2022 · 2022
Later among the works it cites.
Acl-fig: A dataset for scientific figure classification
Zeba Karishma, Shaurya Rohatgi, Kavya Shrinivas Puranik, Jian Wu, and C. Lee Giles. 2023 · 2023
Closest in time.
The semantic scholar open data platform
Rodney Kinney, Chloe Anastasiades, Russell Authur, Iz Beltagy, Jonathan Bragg, Alexandra Buraczynski, Isabel Cachola, Stefan Candra, Yoganand Chandrasekhar, Arman Cohan, et al. 2023 · 2023
Closest in time.
A benchmark of pdf information extraction tools using a multi-task and multi-domain evaluation framework for academic documents
Norman Meuschke, Apurva Jagdale, Timo Spinde, Jelena Mitrović, and Bela Gipp. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.
Assessing the potential of ai-assisted pragmatic annotation: The case of apologies
Danni Yu, Luyang Li, Hang Su, and Matteo Fuoli. 2023 · 2023
Closest in time.