Fetching the paper…
Reading the bibliography…
Under distribution shift (DS) where the training data distribution differs from the test one, a powerful technique is importance weighting (IW) which handles DS in two separate steps: weight estimation (WE) estimates the test-over-training density ratio and weighted classification (WC) trains the classifier from weighted training data.
A stochastic approximation method
H. Robbins and S. Monro · 1951
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Improving predictive inference under covariate shift by weighting the log-likelihood function
H. Shimodaira · 2000
Earlier work this paper cites.
Learning with Kernels
B. Schölkopf and A. Smola · 2001
Earlier work this paper cites.
The class imbalance problem: A systematic study
N. Japkowicz and S. Stephen · 2002
Earlier work this paper cites.
Integrating structured biological data by kernel maximum mean discrepancy
K. M. Borgwardt, A. Gretton, M. J. Rasch, H.-P. Kriegel, B. Schölkopf, and A. J. Smola · 2006
Earlier work this paper cites.
Domain adaptation for statistical classifiers
H. Daume III and D. Marcu · 2006
Earlier work this paper cites.
Analysis of representations for domain adaptation
S. Ben-David, J. Blitzer, K. Crammer, and F. Pereira · 2007
Earlier work this paper cites.
Correcting sample selection bias by unlabeled data
J. Huang, A. Gretton, K. Borgwardt, B. Schölkopf, and A. Smola · 2007
Earlier work this paper cites.
Visualizing data using t-sne
L. v. d. Maaten and G. Hinton · 2008
Earlier work this paper cites.
Direct importance estimation for covariate shift adaptation
M. Sugiyama, T. Suzuki, S. Nakajima, H. Kashima, P. von Bünau, and M. Kawanabe · 2008
Earlier work this paper cites.
Learning from imbalanced data
H. He and E. A. Garcia · 2009
Earlier work this paper cites.
A least-squares approach to direct importance estimation
T. Kanamori, S. Hido, and M. Sugiyama · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
A survey on transfer learning
S. Pan and Q. Yang · 2009
Earlier work this paper cites.
Dataset shift in machine learning
J. Quionero-Candela, M. Sugiyama, A. Schwaighofer, and N. Lawrence · 2009
Earlier work this paper cites.
A kernel two-sample test
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. Smola · 2012
Earlier work this paper cites.
Machine learning in non-stationary environments: Introduction to covariate shift adaptation
M. Sugiyama and M. Kawanabe · 2012
Earlier work this paper cites.
Density ratio estimation in machine learning
M. Sugiyama, T. Suzuki, and T. Kanamori · 2012
Earlier work this paper cites.
Robust solutions of optimization problems affected by uncertain probabilities
A. Ben-Tal, D. Den Hertog, A. De Waegenaere, B. Melenberg, and G. Rennen · 2013
Cited alongside, same era.
Clustering unclustered data: Unsupervised binary labeling of two datasets having different class balances
M. C. du Plessis, G. Niu, and M. Sugiyama · 2013
Cited alongside, same era.
Learning with noisy labels
N. Natarajan, I. S. Dhillon, P. K. Ravikumar, and A. Tewari · 2013
Cited alongside, same era.
Classification with asymmetric label noise: Consistency and maximal denoising
C. Scott, G. Blanchard, and G. Handy · 2013
Cited alongside, same era.
Domain adaptation under target and conditional shift
K. Zhang, B. Schölkopf, K. Muandet, and Z. Wang · 2013
Cited alongside, same era.
Dropout: a simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Variance-based regularization with convex objectives
H. Namkoong and J. C. Duchi · 2017
Later among the works it cites.
Making deep neural networks robust to label noise: A loss correction approach
G. Patrini, A. Rozza, A. Krishna Menon, R. Nock, and L. Qu · 2017
Later among the works it cites.
Asymmetric tri-training for unsupervised domain adaptation
K. Saito, Y. Ushiku, and T. Harada · 2017
Later among the works it cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
H. Xiao, K. Rasul, and R. Vollgraf · 2017
Later among the works it cites.
Understanding deep learning requires rethinking generalization
C. Zhang, S. Bengio, M. Hardt, B. Recht, and O. Vinyals · 2017
Later among the works it cites.
A systematic study of the class imbalance problem in convolutional neural networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Robust learning under uncertain test distributions: Relating covariate shift to model misspecification
J. Wen, C.-N. Yu, and R. Greiner · 2014
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
D. P. Kingma and J. L. Ba · 2015
Cited alongside, same era.
Learning from corrupted binary labels via class-probability estimation
A. Menon, B. Van Rooyen, C. S. Ong, and B. Williamson · 2015
Cited alongside, same era.
Learning with symmetric label noise: The importance of being unhinged
B. Van Rooyen, A. Menon, and R. C. Williamson · 2015
Cited alongside, same era.
Domain-adversarial training of neural networks
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. March, and V. Lempitsky · 2016
Cited alongside, same era.
M. Buda, A. Maki, and M. A. Mazurowski · 2018
Later among the works it cites.
Does distributionally robust supervised learning give robust classifiers?
W. Hu, G. Niu, I. Sato, and M. Sugiyama · 2018
Later among the works it cites.
Mentornet: Learning data-driven curriculum for very deep neural networks on corrupted labels
L. Jiang, Z. Zhou, T. Leung, L.-J. Li, and L. Fei-Fei · 2018
Later among the works it cites.
Detecting and correcting for label shift with black box predictors
Z. C. Lipton, Y.-X. Wang, and A. Smola · 2018
Later among the works it cites.
Learning to reweight examples for robust deep learning
M. Ren, W. Zeng, B. Yang, and R. Urtasun · 2018
Later among the works it cites.
What is the effect of importance weighting in deep learning?
J. Byrd and Z. C. Lipton · 2019
Later among the works it cites.
Learning imbalanced datasets with label-distribution-aware margin loss
K. Cao, C. Wei, A. Gaidon, N. Arechiga, and T. Ma · 2019
Later among the works it cites.
On the minimal supervision for training any binary classifier from only unlabeled data
N. Lu, G. Niu, A. K. Menon, and M. Sugiyama · 2019
Later among the works it cites.
Are anchor points really indispensable in label-noise learning?
X. Xia, T. Liu, N. Wang, B. Han, C. Gong, G. Niu, and M. Sugiyama · 2019
Later among the works it cites.
How does disagreement help generalization against label corruption?
X. Yu, B. Han, J. Yao, G. Niu, I. W. Tsang, and M. Sugiyama · 2019
Later among the works it cites.
SIGUA: Forgetting may make learning with noisy labels more robust
B. Han, G. Niu, X. Yu, Q. Yao, M. Xu, I. W. Tsang, and M. Sugiyama · 2020
Closest in time.
Mitigating overfitting in supervised classification from two unlabeled datasets: A consistent risk correction approach
N. Lu, T. Zhang, G. Niu, and M. Sugiyama · 2020
Closest in time.
Part-dependent label noise: Towards instance-dependent label noise
X. Xia, T. Liu, B. Han, N. Wang, M. Gong, H. Liu, G. Niu, D. Tao, and M. Sugiyama · 2020
Closest in time.
A one-step approach to covariate shift adaptation
T. Zhang, I. Yamane, N. Lu, and M. Sugiyama · 2020
Closest in time.