Fetching the paper…
Reading the bibliography…
Influence diagrams have recently been used to analyse the safety and fairness properties of AI systems.
Information Value Theory
Howard, R. A. 1966 · 1966
Earlier work this paper cites.
The economic value of analysis and computation
Matheson, J. E. 1968 · 1968
Earlier work this paper cites.
Evaluating influence diagrams
Shachter, R. D. 1986 · 1986
Earlier work this paper cites.
Causal Networks: Semantics and Expressiveness
Verma, T.; and Pearl, J. 1988 · 1988
Earlier work this paper cites.
A note about redundancy in influence diagrams
Fagiuoli, E.; and Zaffalon, M. 1998 · 1998
Earlier work this paper cites.
Bayes-Ball: The Rational Pastime (for Determining Irrelevance and Requisite Information in Belief Networks and Influence Diagrams)
Shachter, R. D. 1998 · 1998
Earlier work this paper cites.
Welldefined decision scenarios
Nielsen, T. D.; and Jensen, F. V. 1999 · 1999
Earlier work this paper cites.
Evaluating influence diagrams using LIMIDs
Nilsson, D.; and Lauritzen, S. L. 2000 · 2000
Earlier work this paper cites.
Multi-agent influence diagrams for representing and solving games
Koller, D.; and Milch, B. 2003 · 2003
Earlier work this paper cites.
Influence Diagram Retrospective
Howard, R. A.; Matheson, J. E.; Howard, R. A.; and Matheson, J. E. 2005 · 2005
Earlier work this paper cites.
Ignorable Information in Multi-Agent Scenarios
Milch, B.; and Koller, D. 2008 · 2008
Cited alongside, same era.
Sensitivity analysis: a review of recent advances
Borgonovo, E.; and Plischke, E. 2016 · 2016
Cited alongside, same era.
Cooperative Inverse Reinforcement Learning
Hadfield-Menell, D.; Dragan, A.; Abbeel, P.; and Russell, S. J. 2016 · 2016
Cited alongside, same era.
Decisions and Dependence in Influence Diagrams
Shachter, R. D. 2016 · 2016
Cited alongside, same era.
Counterfactual Fairness
Kusner, M. J.; Loftus, J. R.; Russell, C.; and Silva, R. 2017 · 2017
Cited alongside, same era.
Should robots be obedient?
Milli, S.; Hadfield-Menell, D.; Dragan, A.; and Russell, S. J. 2017 · 2017
Cited alongside, same era.
Characterizing optimal mixed policies: Where to intervene and what to observe
Lee, S.; and Bareinboim, E. 2020 · 2020
Later among the works it cites.
Causal imitation learning with unobserved confounders
Zhang, J.; Kumor, D.; and Bareinboim, E. 2020 · 2020
Later among the works it cites.
User Tampering in Reinforcement Learning Recommender Systems
Evans, C.; and Kasirzadeh, A. 2021 · 2021
Later among the works it cites.
PyCID: A Python Library for Causal Influence Diagrams
Fox, J.; Everitt, T.; Carey, R.; Langlois, E.; Abate, A.; and Wooldridge, M. 2021 · 2021
Later among the works it cites.
Equilibrium Refinements for Multi-Agent Influence Diagrams: Theory and Practice
Hammond, L.; Fox, J.; Everitt, T.; Abate, A.; and Wooldridge, M. 2021 · 2021
Later among the works it cites.
How RL Agents Behave when their Actions are Modified
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Everitt, T.; Kumar, R.; Krakovna, V.; and Legg, S. 2019 · 2019
Cited alongside, same era.
Pitfalls in learning a reward function online
Armstrong, S.; Orseau, L.; Leike, J.; and Legg, S. 2020 · 2020
Cited alongside, same era.
Asymptotically Unambitious Artificial General Intelligence
Cohen, M. K.; Vellambi, B. N.; and Hutter, M. 2020 · 2020
Cited alongside, same era.
AGI Agent Safety by Iteratively Improving the Utility Function
Holtman, K. 2020 · 2020
Cited alongside, same era.
Agent Incentives: A Causal Perspective
Everitt, T.; Carey, R.; Langlois, E.; Ortega, P. A.; and Legg, S. 2021a
Cited in the paper.
Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective
Everitt, T.; Hutter, M.; Kumar, R.; and Krakovna, V. 2021b
Cited in the paper.
Langlois, E.; and Everitt, T. 2021 · 2021
Later among the works it cites.
Submodel Decomposition Bounds for Influence Diagrams
Lee, J.; Marinescu, R.; and Dechter, R. 2021 · 2021
Later among the works it cites.
Why Fair Labels Can Yield Unfair Predictions: Graphical Conditions for Introduced Unfairness
Ashurst, C.; Carey, R.; Chiappa, S.; and Everitt, T. 2022 · 2022
Closest in time.
Path-Specific Objectives for Safer Agent Incentives
Farquhar, S.; Carey, R.; and Everitt, T. 2022 · 2022
Closest in time.