Fetching the paper…
Reading the bibliography…
Unlike traditional supervised learning, in many settings only partial feedback is available.
Classifying treatment responders under causal effect monotonicity
Kallus, N. 2019 · 1902
Earlier work this paper cites.
Adapting neural networks for the estimation of treatment effects
Shi, C.; Blei, D. M.; and Veitch, V. 2019 · 1906
Earlier work this paper cites.
The central role of the propensity score in observational studies for causal effects
Rosenbaum, P. R.; and Rubin, D. B. 1983 · 1983
Earlier work this paper cites.
Estimation of consumer demand systems with binding non-negativity constraints
Wales, T. J.; and Woodland, A. D. 1983 · 1983
Earlier work this paper cites.
Model-based direct adjustment
Rosenbaum, P. R. 1987 · 1987
Earlier work this paper cites.
Virtual adversarial training: a regularization method for supervised and semi-supervised learning
Miyato, T.; Maeda, S.-i.; Koyama, M.; and Ishii, S. 2018 · 1993
Earlier work this paper cites.
Revenue management: Research overview and prospects
McGill, J. I.; and Van Ryzin, G. J. 1999 · 1999
Earlier work this paper cites.
Text classification from labeled and unlabeled documents using EM
Nigam, K.; McCallum, A. K.; Thrun, S.; and Mitchell, T. 2000 · 2000
Earlier work this paper cites.
Models, reasoning and inference
Pearl, J.; et al. 2000 · 2000
Earlier work this paper cites.
Semi-supervised logistic regression
Amini, M.-R.; and Gallinari, P. 2002 · 2002
Earlier work this paper cites.
A kernel method for multi-labelled classification
Elisseeff, A.; and Weston, J. 2002 · 2002
Earlier work this paper cites.
Learning multi-label scene classification
Boutell, M. R.; Luo, J.; Shen, X.; and Brown, C. M. 2004 · 2004
Earlier work this paper cites.
Variance reduction techniques for gradient estimates in reinforcement learning
Greensmith, E.; Bartlett, P. L.; and Baxter, J. 2004 · 2004
Earlier work this paper cites.
Semi-supervised learning by entropy minimization
Grandvalet, Y.; and Bengio, Y. 2005 · 2005
Earlier work this paper cites.
Causal inference using potential outcomes: Design, modeling, decisions
Rubin, D. B. 2005 · 2005
Earlier work this paper cites.
Demystifying double robustness: A comparison of alternative strategies for estimating a population mean from incomplete data
Kang, J. D.; Schafer, J. L.; et al. 2007 · 2007
Earlier work this paper cites.
Demand estimation and assortment optimization under substitution: Methodology and application
Kök, A. G.; and Fisher, M. L. 2007 · 2007
Earlier work this paper cites.
A kernel statistical test of independence
Gretton, A.; Fukumizu, K.; Teo, C. H.; Song, L.; Schölkopf, B.; and Smola, A. J. 2008 · 2008
Earlier work this paper cites.
A self-training semi-supervised SVM algorithm and its application in an EEG-based brain computer interface speller system
Li, Y.; Guan, C.; Li, H.; and Chin, Z. 2008 · 2008
Cited alongside, same era.
Recent developments in the econometrics of program evaluation
Imbens, G. W.; and Wooldridge, J. M. 2009 · 2009
Cited alongside, same era.
Empirical Bernstein bounds and sample variance penalization
Maurer, A.; and Pontil, M. 2009 · 2009
Cited alongside, same era.
A contextual-bandit approach to personalized news article recommendation
Li, L.; Chu, W.; Langford, J.; and Schapire, R. E. 2010 · 2010
Cited alongside, same era.
Theoretical analysis of self-training with deep networks on unlabeled data
Wei, C.; Shen, K.; Chen, Y.; and Ma, T. 2020 · 2010
Cited alongside, same era.
Semi-supervised self-training for decision tree classifiers
Tanha, J.; van Someren, M.; and Afsarmanesh, H. 2017 · 2017
Later among the works it cites.
Deep learning with logged bandit feedback
Joachims, T.; Swaminathan, A.; and de Rijke, M. 2018 · 2018
Later among the works it cites.
Confounding-robust policy improvement
Kallus, N.; and Zhou, A. 2018 · 2018
Later among the works it cites.
GANITE: Estimation of individualized treatment effects using generative adversarial nets
Yoon, J.; Jordon, J.; and van der Schaar, M. 2018 · 2018
Later among the works it cites.
Domain adaptation for semantic segmentation via class-balanced self-training
Zou, Y.; Yu, Z.; Kumar, B.; and Wang, J. 2018 · 2018
Later among the works it cites.
Counterfactual Randomization: Rescuing Experimental Studies from Obscured Confounding
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
LIBSVM: A library for support vector machines
Chang, C.-C.; and Lin, C.-J. 2011 · 2011
Cited alongside, same era.
Bayesian nonparametric modeling for causal inference
Hill, J. L. 2011 · 2011
Cited alongside, same era.
Identification of conditional interventional distributions
Shpitser, I.; and Pearl, J. 2012 · 2012
Cited alongside, same era.
Doubly robust policy evaluation and optimization
Dudík, M.; Erhan, D.; Langford, J.; Li, L.; et al. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2014 · 2014
Cited alongside, same era.
The power and limits of predictive approaches to observational-data-driven optimization
Bertsimas, D.; and Kallus, N. 2016 · 2016
Cited alongside, same era.
Learning representations for counterfactual inference
Johansson, F.; Shalit, U.; and Sontag, D. 2016 · 2016
Cited alongside, same era.
Forney, A.; and Bareinboim, E. 2019 · 2019
Later among the works it cites.
Unsupervised domain adaptation via calibrating uncertainties
Han, L.; Zou, Y.; Gao, R.; Wang, L.; and Metaxas, D. 2019 · 2019
Later among the works it cites.
Metalearners for estimating heterogeneous treatment effects using machine learning
Künzel, S. R.; Sekhon, J. S.; Bickel, P. J.; and Yu, B. 2019 · 2019
Later among the works it cites.
Batch Learning from Bandit Feedback through Bias Corrected Reward Imputation
Wang, L.; Bai, Y.; Bhalla, A.; and Joachims, T. 2019 · 2019
Later among the works it cites.
Confidence regularized self-training
Zou, Y.; Yu, Z.; Liu, X.; Kumar, B.; and Wang, J. 2019 · 2019
Later among the works it cites.
Cost-Effective Incentive Allocation via Structured Counterfactual Inference
Lopez, R.; Li, C.; Yan, X.; and Xiong, J. 2020 · 2020
Later among the works it cites.
Off-policy bandits with deficient support
Sachdeva, N.; Su, Y.; and Joachims, T. 2020 · 2020
Later among the works it cites.
Counterfactual cross-validation: Stable model selection procedure for causal inference models
Saito, Y.; and Yasui, S. 2020 · 2020
Later among the works it cites.
Counterfactual Learning of Continuous Stochastic Policies
Zenati, H.; Bietti, A.; Martin, M.; Diemert, E.; and Mairal, J. 2020 · 2020
Later among the works it cites.
Loss Functions for Discrete Contextual Pricing with Observational Data
Biggs, M.; Gao, R.; and Sun, W. 2021 · 2021
Closest in time.
Human-AI Collaboration with Bandit Feedback
Gao, R.; Saar-Tsechansky, M.; De-Arteaga, M.; Han, L.; Lee, M. K.; and Lease, M. 2021 · 2021
Closest in time.
Rizve, M. N.; Duarte, K.; Rawat, Y. S.; and Shah, M. 2021 · 2021
Closest in time.