Fetching the paper…
Reading the bibliography…
We study how to learn optimal interventions sequentially given causal information represented as a causal graph along with associated conditional distributions.
Causality: models, reasoning and inference
Pearl, J. (2000) · 2000
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
Auer, P., Cesa-Bianchi, N., and Fischer, P. (2002) · 2002
Earlier work this paper cites.
Probabilistic graphical models: principles and techniques
Koller, D. and Friedman, N. (2009) · 2009
Earlier work this paper cites.
A structured multiarmed bandit problem and the greedy policy
Mersereau, A. J., Rusmevichientong, P., and Tsitsiklis, J. N. (2009) · 2009
Earlier work this paper cites.
An empirical evaluation of thompson sampling
Chapelle, O. and Li, L. (2011) · 2011
Earlier work this paper cites.
Analysis of thompson sampling for the multi-armed bandit problem
Agrawal, S. and Goyal, N. (2012) · 2012
Earlier work this paper cites.
Combinatorial bandits
Cesa-Bianchi, N. and Lugosi, G. (2012) · 2012
Cited alongside, same era.
Experiment selection for causal discovery
Hyttinen, A., Eberhardt, F., and Hoyer, P. O. (2013) · 2013
Cited alongside, same era.
Generalized thompson sampling for sequential decision-making and causal inference
Ortega, P. A. and Braun, D. A. (2014) · 2014
Cited alongside, same era.
Improving online marketing experiments with drifting multi-armed bandits
Burtini, G., Loeppky, J. L., and Lawrence, R. (2015) · 2015
Cited alongside, same era.
Multi-armed bandit models for the optimal design of clinical trials: benefits and challenges
Villar, S. S., Bowden, J., and Wason, J. (2015) · 2015
Cited alongside, same era.
Causal bandits: Learning good interventions via causal inference
Lattimore, F., Lattimore, T., and Reid, M. D. (2016) · 2016
Cited alongside, same era.
Axis: Generating explanations at scale with learnersourcing and machine learning
Williams, J. J., Kim, J., Rafferty, A., Maldonado, S., Gajos, K. Z., Lasecki, W. S., and Heffernan, N. (2016) · 2016
Later among the works it cites.
Online learning for causal bandits
Sachidananda, V. and Brunskill, E. (2017) · 2017
Later among the works it cites.
Identifying best interventions through online importance sampling
Sen, R., Shanmugam, K., Dimakis, A. G., and Shakkottai, S. (2017) · 2017
Later among the works it cites.
From ads to interventions: Contextual bandits in mobile health
Tewari, A. and Murphy, S. A. (2017) · 2017
Later among the works it cites.
Structural causal bandits: where to intervene?
Lee, S. and Bareinboim, E. (2018) · 2018
Later among the works it cites.
Bandit algorithms
Lattimore, T. and Szepesvári, C. (2020) · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…