Fetching the paper…
Reading the bibliography…
Natural Language Processing algorithms have made incredible progress, but they still struggle when applied to out-of-distribution examples.
Martín Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz. 2019 · 1907
Earlier work this paper cites.
Distilbert, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Continual learning in reinforcement environments
Mark B. Ring. 1995 · 1995
Earlier work this paper cites.
Supervised and unsupervised PCFG adaptation to novel domains
Brian Roark and Michiel Bacchiani. 2003 · 2003
Earlier work this paper cites.
Mining and summarizing customer reviews
Minqing Hu and Bing Liu. 2004 · 2004
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Domain adaptation with structural correspondence learning
John Blitzer, Ryan McDonald, and Fernando Pereira. 2006 · 2006
Earlier work this paper cites.
Domain adaptation for statistical classifiers
Hal Daumé III and Daniel Marcu. 2006 · 2006
Earlier work this paper cites.
Biographies, bollywood, boom-boxes and blenders: Domain adaptation for sentiment classification
John Blitzer, Mark Dredze, and Fernando Pereira. 2007 · 2007
Earlier work this paper cites.
Self-training for enhancement and domain adaptation of statistical parsers trained on small datasets
Roi Reichart and Ari Rappoport. 2007 · 2007
Earlier work this paper cites.
Multi-domain sentiment classification
Shoushan Li and Chengqing Zong. 2008 · 2008
Earlier work this paper cites.
Transfer learning from multiple source domains via consensus regularization
Ping Luo, Fuzhen Zhuang, Hui Xiong, Yuhong Xiong, and Qing He. 2008 · 2008
Earlier work this paper cites.
Zero-shot domain adaptation: A multi-view approach
John Blitzer, Dean P. Foster, and Sham Machandranath Kakade. 2009 · 2009
Earlier work this paper cites.
Zero-shot learning with semantic output codes
Mark Palatucci, Dean Pomerleau, Geoffrey E. Hinton, and Tom M. Mitchell. 2009 · 2009
Earlier work this paper cites.
Automatic domain adaptation for parsing
David McClosky, Eugene Charniak, and Mark Johnson. 2010 · 2010
Earlier work this paper cites.
Cross-domain sentiment classification via spectral feature alignment
Sinno Jialin Pan, Xiaochuan Ni, Jian-Tao Sun, Qiang Yang, and Zheng Chen. 2010 · 2010
Earlier work this paper cites.
Sentence and expression level annotation of opinions in user-generated discourse
Cigdem Toprak, Niklas Jakob, and Iryna Gurevych. 2010 · 2010
Earlier work this paper cites.
Domain adaptation for large-scale sentiment classification: A deep learning approach
Xavier Glorot, Antoine Bordes, and Yoshua Bengio. 2011 · 2011
Earlier work this paper cites.
Conditioned natural language generation using only unconditioned language model: An exploration
Fan-Keng Sun and Cheng-I Lai. 2020 · 2011
Earlier work this paper cites.
Marginalized denoising autoencoders for domain adaptation
Minmin Chen, Zhixiang Eddie Xu, Kilian Q. Weinberger, and Fei Sha. 2012 · 2012
Earlier work this paper cites.
WILDS: A benchmark of in-the-wild distribution shifts
Pang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie, Marvin Zhang, Akshay Balsubramani, Weihua Hu, Michihiro Yasunaga, Richard Lanas Phillips, Sara Beery, Jure Leskovec, Anshul Kundaje, Emma Pierson, Sergey Levine, Chelsea Finn, and Percy Liang. 2020 · 2012
Earlier work this paper cites.
Improved parsing and POS tagging using inter-sentence consistency constraints
Alexander M. Rush, Roi Reichart, Michael Collins, and Amir Globerson. 2012 · 2012
Earlier work this paper cites.
Domain generalization via invariant feature representation
Krikamol Muandet, David Balduzzi, and Bernhard Schölkopf. 2013 · 2013
Earlier work this paper cites.
Semeval-2014 task 4: Aspect based sentiment analysis
Maria Pontiki, Dimitris Galanis, John Pavlopoulos, Harris Papageorgiou, Ion Androutsopoulos, and Suresh Manandhar. 2014 · 2014
Earlier work this paper cites.
FLORS: fast and simple domain adaptation for part-of-speech tagging
Tobias Schnabel and Hinrich Schütze. 2014 · 2014
Cited alongside, same era.
Fast easy unsupervised domain adaptation with marginalized structured dropout
Yi Yang and Jacob Eisenstein. 2014 · 2014
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel Bowman, Gabor Angeli, Christopher Potts, and Christopher Manning. 2015 · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor S. Lempitsky. 2016 · 2016
Cited alongside, same era.
How can we know what language models know
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Later among the works it cites.
BART: denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Domain robustness in neural machine translation
Mathias Müller, Annette Rios, and Rico Sennrich. 2020 · 2020
Later among the works it cites.
Continual learning with hypernetworks
Johannes von Oswald, Christian Henning, João Sacramento, and Benjamin F. Grewe. 2020 · 2020
Later among the works it cites.
MAD-X: an adapter-based framework for multi-task cross-lingual transfer
Jonas Pfeiffer, Ivan Vulic, Iryna Gurevych, and Sebastian Ruder. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ashwin K. Vijayakumar, Michael Cogswell, Ramprasaath R. Selvaraju, Qing Sun, Stefan Lee, David J. Crandall, and Dhruv Batra. 2016 · 2016
Cited alongside, same era.
Analysing how people orient to and spread rumours in social media by looking at conversational threads
Arkaitz Zubiaga, Maria Liakata, Rob Procter, Geraldine Wong Sak Hoi, and Peter Tolmie. 2016 · 2016
Cited alongside, same era.
Domain attention with an ensemble of experts
Young-Bum Kim, Karl Stratos, and Dongchan Kim. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Neural structural correspondence learning for domain adaptation
Yftah Ziser and Roi Reichart. 2017 · 2017
Cited alongside, same era.
Exploiting context for rumour detection in social media
Arkaitz Zubiaga, Maria Liakata, and Rob Procter. 2017 · 2017
Cited alongside, same era.
Multinomial adversarial networks for multi-domain text classification
Xilun Chen and Claire Cardie. 2018 · 2018
Cited alongside, same era.
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
Neural unsupervised domain adaptation in NLP - A survey
Alan Ramponi and Barbara Plank. 2020 · 2020
Later among the works it cites.
Multicqa: Zero-shot transfer of self-supervised text matching models on a massive scale
Andreas Rücklé, Jonas Pfeiffer, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
Distributionally robust neural networks
Shiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, and Percy Liang. 2020 · 2020
Later among the works it cites.
Autoprompt: Eliciting knowledge from language models with automatically generated prompts
Taylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace, and Sameer Singh. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Later among the works it cites.
Transformer based multi-source domain adaptation
Dustin Wright and Isabelle Augenstein. 2020 · 2020
Later among the works it cites.
Bertscore: Evaluating text generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Later among the works it cites.
Making pre-trained language models better few-shot learners
Tianyu Gao, Adam Fisch, and Danqi Chen. 2021 · 2021
Closest in time.
PTR: prompt tuning with rules for text classification
Xu Han, Weilin Zhao, Ning Ding, Zhiyuan Liu, and Maosong Sun. 2021 · 2021
Closest in time.
Bertese: Learning to speak to BERT
Adi Haviv, Jonathan Berant, and Amir Globerson. 2021 · 2021
Closest in time.
DILBERT: customized pre-training for domain adaptation with category shift, with an application to aspect extraction
Entony Lekhtman, Yftah Ziser, and Roi Reichart. 2021 · 2021
Closest in time.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Closest in time.
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang. 2021 · 2021
Closest in time.
How many data points is a prompt worth?
Teven Le Scao and Alexander M. Rush. 2021 · 2021
Closest in time.
Exploiting cloze-questions for few-shot text classification and natural language inference
Timo Schick and Hinrich Schütze. 2021 · 2021
Closest in time.
On calibration and out-of-domain generalization
Yoav Wald, Amir Feder, Daniel Greenfeld, and Uri Shalit. 2021 · 2021
Closest in time.
Does distributionally robust supervised learning give robust classifiers?
Weihua Hu, Gang Niu, Issei Sato, and Masashi Sugiyama. 2018 · 2042
Closest in time.