Fetching the paper…
Reading the bibliography…
The recent literature in text classification is biased towards short text sequences (e.g., sentences or paragraphs).
DocBERT: BERT for Document Classification
Ashutosh Adhikari, Achyudh Ram, Raphael Tang, and Jimmy Lin. 2019 · 1904
Earlier work this paper cites.
Generating long sequences with sparse transformers
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever. 2019 · 1904
Earlier work this paper cites.
RoBERTa: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Andriy Mulyar, Elliot Schumacher, Masoud Rouhizadeh, and Mark Dredze. 2019 · 1910
Earlier work this paper cites.
A Probabilistic Analysis of the Rocchio Algorithm with TFIDF for Text Categorization
Thorsten Joachims. 1997 · 1997
Earlier work this paper cites.
Longformer: The Long-Document Transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
A shared task involving multi-label classification of clinical free text
John Pestian, Chris Brew, Pawel Matykiewicz, Dj J Hovermale, Neil Johnson, K Bretonnel Cohen, and Wlodzislaw Duch. 2007 · 2007
Earlier work this paper cites.
Automatic construction of rule-based ICD-9-CM coding systems
Richárd Farkas and György Szarvas. 2008 · 2008
Earlier work this paper cites.
Efficient Transformers: A Survey
Yi Tay, Mostafa Dehghani, Dara Bahri, and Donald Metzler. 2020 · 2009
Earlier work this paper cites.
Automatic classification of diseases from free-text death certificates for real-time surveillance
Bevan Koopman, Sarvnaz Karimi, Anthony Nguyen, Rhydwyn McGuire, David Muscatello, Madonna Kemp, Donna Truran, Ming Zhang, and Sarah Thackway. 2015 · 2015
Earlier work this paper cites.
Document modeling with gated recurrent neural network for sentiment classification
Duyu Tang, Bing Qin, and Ting Liu. 2015 · 2015
Earlier work this paper cites.
MIMIC-III, a freely accessible critical care database
Alistair E W Johnson, Tom J Pollard, Lu Shen, H Lehman Li-Wei, Mengling Feng, Mohammad Ghassemi, Benjamin Moody, Peter Szolovits, Leo Anthony Celi, and Roger G Mark. 2016 · 2016
Earlier work this paper cites.
Hierarchical Attention Networks for Document Classification
Zichao Yang, Diyi Yang, Chris Dyer, Xiaodong He, Alex Smola, and Eduard Hovy. 2016 · 2016
Earlier work this paper cites.
Automatic Diagnosis Coding of Radiology Reports: A Comparison of Deep Learning and Conventional Classification Methods
Sarvnaz Karimi, Xiang Dai, Hamed Hassanzadeh, and Anthony Nguyen. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Explainable Prediction of Medical Codes from Clinical Text
James Mullenbach, Sarah Wiegreffe, Jon Duke, Jimeng Sun, and Jacob Eisenstein. 2018 · 2018
Earlier work this paper cites.
Few-Shot and Zero-Shot Multi-Label Learning for Structured Label Spaces
Anthony Rios and Ramakanth Kavuluru. 2018 · 2018
Earlier work this paper cites.
A Neural Architecture for Automated ICD Coding
Pengtao Xie, Haoran Shi, Ming Zhang, and Eric Xing. 2018 · 2018
Cited alongside, same era.
Transformer-XL: Attentive Language Models beyond a Fixed-Length Context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime Carbonell, Quoc Le, and Ruslan Salakhutdinov. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
ML-Net: multi-label classification of biomedical texts with deep neural networks
Jingcheng Du, Qingyu Chen, Yifan Peng, Yang Xiang, Cui Tao, and Zhiyong Lu. 2019 · 2019
Cited alongside, same era.
SemEval-2019 Task 4: Hyperpartisan News Detection
Johannes Kiesel, Maria Mestre, Rishabh Shukla, Emmanuel Vincent, Payam Adineh, David Corney, Benno Stein, and Martin Potthast. 2019 · 2019
Cited alongside, same era.
How to Fine-Tune BERT for Text Classification?
Neural Unsupervised Domain Adaptation in NLP—A Survey
Alan Ramponi and Barbara Plank. 2020 · 2020
Later among the works it cites.
A label attention model for ICD coding from clinical text
Thanh Vu, Dat Quoc Nguyen, and Anthony Nguyen. 2020 · 2020
Later among the works it cites.
Beyond 512 Tokens: Siamese Multi-depth Transformer-based Hierarchical Encoder for Long-Form Document Matching
Liu Yang, Mingyang Zhang, Cheng Li, Michael Bendersky, and Marc Najork. 2020 · 2020
Later among the works it cites.
Big Bird: Transformers for Longer Sequences
Manzil Zaheer, Guru Guruganesh, Avinava Dubey, Joshua Ainslie, Chris Alberti, Santiago Ontanon, Philip Pham, Anirudh Ravula, Qifan Wang, and Li Yang. 2020 · 2020
Later among the works it cites.
Rethinking Attention with Performers
Krzysztof Marcin Choromanski, Valerii Likhosherstov, David Dohan, Xingyou Song, Andreea Gane, Tamás Sarlós, Peter Hawkins, Jared Quincy Davis, Afroz Mohiuddin, Lukasz Kaiser, David Benjamin Belanger, Lucy J Colwell, and Adrian Weller. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chi Sun, Xipeng Qiu, Yige Xu, and Xuanjing Huang. 2019 · 2019
Cited alongside, same era.
Label-Specific Document Representation for Multi-Label Text Classification
Lin Xiao, Xin Huang, Boli Chen, and Liping Jing. 2019 · 2019
Cited alongside, same era.
EHR Coding with Multi-scale Feature Attention and Structured Knowledge Graph Propagation
Xiancheng Xie, Yun Xiong, Philip S Yu, and Yangyong Zhu. 2019 · 2019
Cited alongside, same era.
Xingxing Zhang, Furu Wei, and Ming Zhou. 2019 · 2019
Cited alongside, same era.
HyperCore: Hyperbolic and Co-graph Representation for Automatic ICD Coding
Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao, Shengping Liu, and Weifeng Chong. 2020 · 2020
Cited alongside, same era.
An Empirical Study on Large-Scale Multi-Label Text Classification Including Few and Zero-Shot Labels
Ilias Chalkidis, Manos Fergadiotis, Sotiris Kotitsas, Prodromos Malakasiotis, Nikolaos Aletras, and Ion Androutsopoulos. 2020 · 2020
Cited alongside, same era.
Don’t Stop Pretraining: Adapt Language Models to Domains and Tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A Smith. 2020 · 2020
Cited alongside, same era.
ERNIE-Doc: A Retrospective Long-Document Modeling Transformer
SiYu Ding, Junyuan Shang, Shuohuan Wang, Yu Sun, Hao Tian, Hua Wu, and Haifeng Wang. 2021 · 2021
Later among the works it cites.
Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
Jesse Dodge, Maarten Sap, Ana Marasović, William Agnew, Gabriel Ilharco, Dirk Groeneveld, Margaret Mitchell, and Matt Gardner. 2021 · 2021
Later among the works it cites.
Explainable automated coding of clinical notes using hierarchical label-wise attention networks and label embedding initialisation
Hang Dong, Víctor Suárez-Paniagua, William Whiteley, and Honghan Wu. 2021 · 2021
Later among the works it cites.
Limitations of Transformers on Clinical Text Classification
Shang Gao, Mohammed Alawad, M. Todd Young, John Gounley, Noah Schaefferkoetter, Hong Jun Yoon, Xiao-Cheng Wu, Eric B. Durbin, Jennifer Doherty, Antoinette Stroup, Linda Coyle, and Georgia Tourassi. 2021 · 2021
Later among the works it cites.
mDAPT: Multilingual Domain Adaptive Pretraining in a Single Model
Rasmus Kær Jørgensen, Mareike Hartmann, Xiang Dai, and Desmond Elliott. 2021 · 2021
Later among the works it cites.
On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines
Marius Mosbach, Maksym Andriushchenko, and Dietrich Klakow. 2021 · 2021
Later among the works it cites.
Towards BERT-based Automatic ICD Coding: Limitations and Opportunities
Damian Pascual, Sandro Luck, and Roger Wattenhofer. 2021 · 2021
Later among the works it cites.
Simple Local Attentions Remain Competitive for Long-Context Tasks
Wenhan Xiong, Barlas Oğuz, Anchit Gupta, Xilun Chen, Diana Liskovich, Omer Levy, Wen-tau Yih, and Yashar Mehdad. 2021 · 2021
Later among the works it cites.
LexGLUE: A Benchmark Dataset for Legal Language Understanding in English
Ilias Chalkidis, Abhik Jana, Dirk Hartung, Michael J Bommarito II, Ion Androutsopoulos, Daniel Martin Katz, and Nikolaos Aletras. 2022 · 2022
Closest in time.
Efficient Classification of Long Documents Using Transformers
Hyunji Park, Yogarshi Vyas, and Kashif Shah. 2022 · 2022
Closest in time.
Code Synonyms Do Matter: Multiple Synonyms Matching Network for Automatic ICD Coding
Zheng Yuan, Chuanqi Tan, and Songfang Huang. 2022 · 2022
Closest in time.