Fetching the paper…
Reading the bibliography…
In document classification for, e.g., legal and biomedical text, we often deal with hundreds of classes, including very infrequent ones, as well as temporal concept drift caused by the influence of real world events, e.g., policy changes, conflicts, or pandemics.
Martin Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz. 2020 · 1907
Earlier work this paper cites.
Well-read students learn better: The impact of student initialization on knowledge distillation
Iulia Turc, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 1908
Earlier work this paper cites.
Principles of risk minimization for learning theory
V. Vapnik. 1992 · 1992
Earlier work this paper cites.
Improving predictive inference under covariate shift by weighting the log-likelihood function
Hidetoshi Shimodaira. 2000 · 2000
Earlier work this paper cites.
RCV1: A New Benchmark Collection for Text Categorization Research
David D. Lewis, Yiming Yang, Tony G. Rose, and Fan Li. 2004 · 2004
Earlier work this paper cites.
Estimating class priors in domain adaptation for word sense disambiguation
Yee Seng Chan and Hwee Tou Ng. 2006 · 2006
Earlier work this paper cites.
An Evaluation of Efficient Multilabel Classification Algorithms for Large-Scale Problems in the Legal Domain
Eneldo Loza Mencia and Johannes Fürnkranzand. 2007 · 2007
Earlier work this paper cites.
Introduction to Information Retrieval
Christopher D. Manning, Prabhakar Raghavan, and Hinrich Schütze. 2009 · 2009
Earlier work this paper cites.
Stream-based translation models for statistical machine translation
Abby Levenberg, Chris Callison-Burch, and Miles Osborne. 2010 · 2010
Earlier work this paper cites.
Gradient starvation: A learning proclivity in neural networks
Mohammad Pezeshki, Sékou-Oumar Kaba, Yoshua Bengio, Aaron Courville, Doina Precup, and Guillaume Lajoie. 2020 · 2011
Earlier work this paper cites.
Statistical topic models for multi-label document classification
Timothy Rubin, America Chambers, Padhraic Smyth, and Mark Steyvers. 2012 · 2012
Earlier work this paper cites.
A survey on concept drift adaptation
João Gama, Indrundefined Žliobaitundefined, Albert Bifet, Mykola Pechenizkiy, and Abdelhamid Bouchachia. 2014 · 2014
Earlier work this paper cites.
An overview of the bioasq large-scale biomedical semantic indexing and question answering competition
George Tsatsaronis, Georgios Balikas, Prodromos Malakasiotis, Ioannis Partalas, Matthias Zschunke, Michael R Alvers, Dirk Weissenborn, Anastasia Krithara, Sergios Petridis, Dimitris Polychronopoulos, Yannis Almirantis, John Pavlopoulos, Nicolas Baskiotis, Patrick Gallinari, Thierry Artieres, Axel Ngonga, Norman Heino, Eric Gaussier, Liliana Barrio-Alvers, Michael Schroeder, Ion Androutsopoulos, and Georgios Paliouras. 2015 · 2015
Earlier work this paper cites.
Deep CORAL: Correlation Alignment for Deep Domain Adaptation
Baochen Sun and Kate Saenko. 2016 · 2016
Earlier work this paper cites.
MIMIC-III, a freely accessible critical care database
Alistair EW Johnson, David J. Stone, Leo A. Celi, and Tom J. Pollard. 2017 · 2017
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling. 2017 · 2017
Cited alongside, same era.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard S. Zemel. 2017 · 2017
Cited alongside, same era.
Examining temporality in document classification
Xiaolei Huang and Michael J. Paul. 2018 · 2018
Cited alongside, same era.
Sentiment analysis under temporal shift
Jan Lukes and Anders Søgaard. 2018 · 2018
Cited alongside, same era.
Explainable Prediction of Medical Codes from Clinical Text
James Mullenbach, Sarah Wiegreffe, Jon Duke, Jimeng Sun, and Jacob Eisenstein. 2018 · 2018
LEGAL-BERT: The muppets straight out of law school
Ilias Chalkidis, Manos Fergadiotis, Prodromos Malakasiotis, Nikolaos Aletras, and Ion Androutsopoulos. 2020b · 2020
Later among the works it cites.
Out-of-Distribution Generalization via Risk Extrapolation (REx)
David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Rémi Le Priol, and Aaron C. Courville. 2020 · 2020
Later among the works it cites.
MESA: boost ensemble imbalanced learning with meta-sampler
Zhining Liu, Pengfei Wei, Jing Jiang, Wei Cao, Jiang Bian, and Yi Chang. 2020 · 2020
Later among the works it cites.
SSMBA: Self-supervised manifold based data augmentation for improving out-of-domain robustness
Nathan Ng, Kyunghyun Cho, and Marzyeh Ghassemi. 2020 · 2020
Later among the works it cites.
Temporally-informed analysis of named entity recognition
Shruti Rijhwani and Daniel Preotiuc-Pietro. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Few-shot and zero-shot multi-label learning for structured label spaces
Anthony Rios and Ramakanth Kavuluru. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
MultiFiT: Efficient multi-lingual language model fine-tuning
Julian Eisenschlos, Sebastian Ruder, Piotr Czapla, Marcin Kadras, Sylvain Gugger, and Jeremy Howard. 2019 · 2019
Cited alongside, same era.
We need to talk about standard splits
Kyle Gorman and Steven Bedrick. 2019 · 2019
Cited alongside, same era.
Neural temporality adaptation for document classification: Diachronic word embeddings and domain adaptation models
Xiaolei Huang and Michael J. Paul. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
Distributionally Robust Neural Networks
Shiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, and Percy Liang. 2020 · 2020
Later among the works it cites.
Measure twice, cut once: Quantifying bias and fairness in deep neural networks
Cody Blakeney, Gentry Atkinson, Nathaniel Huish, Yan Yan, Vangelis Metris, and Ziliang Zong. 2021 · 2021
Later among the works it cites.
MultiEURLEX - a multi-lingual and multi-label legal document classification dataset for zero-shot cross-lingual transfer
Ilias Chalkidis, Manos Fergadiotis, and Ion Androutsopoulos. 2021 · 2021
Later among the works it cites.
Few-shot cross-lingual stance detection with sentiment-based pre-training
Momchil Hardalov, Arnav Arora, Preslav Nakov, and Isabelle Augenstein. 2021 · 2021
Later among the works it cites.
WILDS: A benchmark of in-the-wild distribution shifts
Pang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie, Marvin Zhang, Akshay Balsubramani, Weihua Hu, Michihiro Yasunaga, Richard Lanas Phillips, Irena Gao, Tony Lee, Etienne David, Ian Stavness, Wei Guo, Berton A. Earnshaw, Imran S. Haque, Sara Beery, Jure Leskovec, Anshul Kundaje, Emma Pierson, Sergey Levine, Chelsea Finn, and Percy Liang. 2021 · 2021
Later among the works it cites.
Pitfalls of static language modelling
Angeliki Lazaridou, Adhiguna Kuncoro, Elena Gribovskaya, Devang Agrawal, Adam Liska, Tayfun Terzi, Mai Gimenez, Cyprien de Masson d’Autume, Sebastian Ruder, Dani Yogatama, Kris Cao, Tomás Kociský, Susannah Young, and Phil Blunsom. 2021 · 2021
Later among the works it cites.
Overview of bioasq 2021: The ninth bioasq challenge on large-scale biomedical semantic indexing and question answering
Anastasios Nentidis, Georgios Katsimpras, Eirini Vandorou, Anastasia Krithara, Luis Gasco, Martin Krallinger, and Georgios Paliouras. 2021 · 2021
Later among the works it cites.
We need to talk about random splits
Anders Søgaard, Sebastian Ebert, Jasmijn Bastings, and Katja Filippova. 2021 · 2021
Later among the works it cites.
Fairness-aware class imbalanced learning
Shivashankar Subramanian, Afshin Rahimi, Timothy Baldwin, Trevor Cohn, and Lea Frermann. 2021 · 2051
Closest in time.