Fetching the paper…
Reading the bibliography…
Instrumental variable (IV) regression is a standard strategy for learning causal relationships between confounded treatment and outcome variables from observational data by utilizing an instrumental variable, which affects the outcome only through the treatment.
The Tariff on Animal and Vegetable Oils
P. Wright · 1928
Earlier work this paper cites.
Generalized inverses in reproducing kernel spaces: An approach to regularization of linear operator equations
M. Z. Nashed and G. Wahba · 1974
Earlier work this paper cites.
Large sample properties of generalized method of moments estimators
L. P. Hansen · 1982
Earlier work this paper cites.
Neuronlike adaptive elements that can solve difficult learning control problems
A. G. Barto, R. S. Sutton, and C. W. Anderson · 1983
Earlier work this paper cites.
Lifetime earnings and the Vietnam era draft lottery: Evidence from social security administrative records
J. D. Angrist · 1990
Earlier work this paper cites.
Efficient Memory-Based Learning for Robot Control
A. W. Moore · 1990
Earlier work this paper cites.
Split-sample instrumental variables estimates of the return to schooling
J. D. Angrist and A. B. Krueger · 1995
Earlier work this paper cites.
Residual algorithms: Reinforcement learning with function approximation
L. Baird · 1995
Earlier work this paper cites.
Identification of causal effects using instrumental variables
J. D. Angrist, G. W. Imbens, and D. B. Rubin · 1996
Earlier work this paper cites.
Linear least-squares algorithms for temporal difference learning
S. J. Bradtke and A. G. Barto · 1996
Earlier work this paper cites.
Jackknife instrumental variables estimation
J. D. Angrist, G. W. Imbens, and A. B. Krueger · 1999
Earlier work this paper cites.
Instrumental variable estimation of nonparametric models
W. K. Newey and J. L. Powell · 2003
Earlier work this paper cites.
Retrospectives: Who invented instrumental variable regression?
J. H. Stock and F. Trebbi · 2003
Earlier work this paper cites.
Tree-based batch mode reinforcement learning
D. Ernst, P. Geurts, and L. Wehenkel · 2005
Earlier work this paper cites.
Semi-nonparametric IV estimation of shape-invariant engel curves
R. Blundell, X. Chen, and D. Kristensen · 2007
Earlier work this paper cites.
Linear inverse problems in structural econometrics estimation based on spectral decomposition and regularization
M. Carrasco, J.-P. Florens, and E. Renault · 2007
Cited alongside, same era.
Random features for large-scale kernel machines
A. Rahimi and B. Recht · 2008
Cited alongside, same era.
MNIST handwritten digit database
Y. LeCun and C. Cortes · 2010
Cited alongside, same era.
Nonparametric instrumental regression
S. Darolles, Y. Fan, J. P. Florens, and E. Renault · 2011
Cited alongside, same era.
Causal inference by surrogate experiments: Z-identifiability
E. Bareinboim and J. Pearl · 2012
Cited alongside, same era.
Measuring the price responsiveness of gasoline demand: Economic shape restrictions and nonparametric demand estimation
R. Blundell, J. Horowitz, and M. Parey · 2012
Cited alongside, same era.
Spectral normalization for generative adversarial networks
T. Miyato, T. Kataoka, M. Koyama, and Y. Yoshida · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
R. S. Sutton and A. G. Barto · 2018
Later among the works it cites.
Deep generalized method of moments for instrumental variable analysis
A. Bennett, N. Kallus, and T. Schnabel · 2019
Later among the works it cites.
Batch policy learning under constraints
H. Le, C. Voloshin, and Y. Yue · 2019
Later among the works it cites.
Behaviour suite for reinforcement learning
I. Osband, Y. Doron, M. Hessel, J. Aslanides, E. Sezener, A. Saraiva, K. McKinney, T. Lattimore, C. Szepesvari, S. Singh, et al · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Estimation of nonparametric conditional moment models with possibly nonsmooth generalized residuals
X. Chen and D. Pouzo · 2012
Cited alongside, same era.
Foundations of Machine Learning
M. Mohri, A. Rostamizadeh, and A. Talwalkar · 2012
Cited alongside, same era.
Instrumental variables estimation with many weak instruments using regularized jive
C. Hansen and D. Kozbur · 2014
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Cited alongside, same era.
Deep IV: A flexible approach for counterfactual prediction
J. Hartford, G. Lewis, K. Leyton-Brown, and M. Taddy · 2017
Cited alongside, same era.
Later among the works it cites.
Environment reconstruction with hidden confounders for reinforcement learning based recommendation
W. Shang, Y. Yu, Q. Li, Z. Qin, Y. Meng, and J. Ye · 2019
Later among the works it cites.
Kernel instrumental variable regression
R. Singh, M. Sahani, and A. Gretton · 2019
Later among the works it cites.
Empirical study of off-policy policy evaluation for reinforcement learning
C. Voloshin, H. M. Le, N. Jiang, and Y. Yue · 2019
Later among the works it cites.
Stabilizing generative adversarial networks: A survey
M. Wiatrak, S. V. Albrecht, and A. Nystrom · 2019
Later among the works it cites.
Acme: A research framework for distributed reinforcement learning
M. Hoffman, B. Shahriari, J. Aslanides, G. Barth-Maron, F. Behbahani, T. Norman, A. Abdolmaleki, A. Cassirer, F. Yang, K. Baumli, S. Henderson, A. Novikov, S. G. Colmenarejo, S. Cabi, C. Gulcehre, T. L. Paine, A. Cowie, Z. Wang, B. Piot, and N. de Freitas · 2020
Closest in time.
Dual IV: A single stage instrumental variable regression
K. Muandet, A. Mehrjou, S. K. Lee, and A. Raj · 2020
Closest in time.
Off-policy policy evaluation for sequential decisions under unobserved confounding
H. Namkoong, R. Keramati, S. Yadlowsky, and E. Brunskill · 2020
Closest in time.
Hyperparameter selection for offline reinforcement learning
T. L. Paine, C. Paduraru, A. Michi, C. Gulcehre, K. Zolna, A. Novikov, Z. Wang, and N. de Freitas · 2020
Closest in time.