Fetching the paper…
Reading the bibliography…
In this paper, we introduce Target-Aware Weighted Training (TAWT), a weighted training algorithm for cross-task learning based on minimizing a representation-based task distance between the source and target tasks.
A model of inductive bias learning
Jonathan Baxter · 2000
Earlier work this paper cites.
Sample splitting and threshold estimation
Bruce E Hansen · 2000
Earlier work this paper cites.
Improving predictive inference under covariate shift by weighting the log-likelihood function
Hidetoshi Shimodaira · 2000
Earlier work this paper cites.
Introduction to the conll-2000 shared task: chunking
Erik F Tjong Kim Sang and Sabine Buchholz · 2000
Earlier work this paper cites.
Mirror descent and nonlinear projected subgradient methods for convex optimization
Amir Beck and Marc Teboulle · 2003
Earlier work this paper cites.
Ontonotes: the 90% solution
Eduard Hovy, Mitch Marcus, Martha Palmer, Lance Ramshaw, and Ralph Weischedel · 2006
Earlier work this paper cites.
Local rademacher complexities and oracle inequalities in risk minimization
Vladimir Koltchinskii · 2006
Earlier work this paper cites.
Instance weighting for domain adaptation in nlp
Jing Jiang and ChengXiang Zhai · 2007
Earlier work this paper cites.
Learning bounds for importance weighting
Corinna Cortes, Yishay Mansour, and Mehryar Mohri · 2010
Earlier work this paper cites.
Concentration inequalities: A nonasymptotic theory of independence
Stéphane Boucheron, Gábor Lugosi, and Pascal Massart · 2013
Earlier work this paper cites.
The Nature of Statistical Learning Theory
Vladimir Vapnik · 2013
Earlier work this paper cites.
Weak convergence and empirical processes: with applications to statistics
Jon Wellner and Aad van der Vaart · 2013
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Cited alongside, same era.
The benefit of multitask representation learning
Andreas Maurer, Massimiliano Pontil, and Bernardino Romera-Paredes · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Automated curriculum learning for neural networks
Alex Graves, Marc G Bellemare, Jacob Menick, Remi Munos, and Koray Kavukcuoglu · 2017
Cited alongside, same era.
Double/debiased machine learning for treatment and structural parameters
Victor Chernozhukov, Denis Chetverikov, Mert Demirer, Esther Duflo, Christian Hansen, Whitney Newey, and James Robins · 2018
Multi-task deep neural networks for natural language understanding
Xiaodong Liu, Pengcheng He, Weizhu Chen, and Jianfeng Gao · 2019
Later among the works it cites.
Minimax lower bounds for transfer learning with linear and one-hidden layer neural networks
MM Kalan and Z Fabian · 2020
Later among the works it cites.
Meta-learning transferable representations with a single target domain
Hong Liu, Jeff Z HaoChen, Colin Wei, and Tengyu Ma · 2020
Later among the works it cites.
What is being transferred in transfer learning?
Behnam Neyshabur, Hanie Sedghi, and Chiyuan Zhang · 2020
Later among the works it cites.
On the theory of transfer learning: The importance of task diversity
Nilesh Tripuraneni, Michael Jordan, and Chi Jin · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The natural language decathlon: Multitask learning as question answering
Bryan McCann, Nitish Shirish Keskar, Caiming Xiong, and Richard Socher · 2018
Cited alongside, same era.
Taskonomy: Disentangling task transfer learning
Amir R Zamir, Alexander Sax, William Shen, Leonidas J Guibas, Jitendra Malik, and Silvio Savarese · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
A comparison of loss weighting strategies for multi task learning in deep neural networks
Ting Gong, Tyler Lee, Cory Stephenson, Venkata Renduchintala, Suchismita Padhy, Anthony Ndirango, Gokce Keskin, and Oguz H Elibol · 2019
Cited alongside, same era.
On the value of target data in transfer learning
Steve Hanneke and Samory Kpotufe · 2019
Cited alongside, same era.
Survey on deep learning with class imbalance
Justin M Johnson and Taghi M Khoshgoftaar · 2019
Cited alongside, same era.
Thomas Wolf, Julien Chaumond, Lysandre Debut, Victor Sanh, Clement Delangue, Anthony Moi, Pierric Cistac, Morgan Funtowicz, Joe Davison, and Sam Shleifer · 2020
Later among the works it cites.
Transfer learning for nonparametric classification: Minimax rate and adaptive classifier
T Tony Cai and Hongji Wei · 2021
Closest in time.
How fine-tuning allows for effective meta-learning
Kurtland Chua, Qi Lei, and Jason D Lee · 2021
Closest in time.
Few-shot learning via learning the representation, provably
Simon Shaolei Du, Wei Hu, Sham M. Kakade, Jason D. Lee, and Qi Lei · 2021
Closest in time.
Foreseeing the benefits of incidental supervision
Hangfeng He, Mingyuan Zhang, Qiang Ning, and Dan Roth · 2021
Closest in time.
Provable meta-learning of linear representations
Nilesh Tripuraneni, Chi Jin, and Michael Jordan · 2021
Closest in time.
A survey on multi-task learning
Yu Zhang and Qiang Yang · 2021
Closest in time.