Fetching the paper…
Reading the bibliography…
Agents are systems that optimize an objective function in an environment.
“Modeling AGI Safety Frameworks with Causal Influence Diagrams”
Tom Everitt, Ramana Kumar, Victoria Krakovna and Shane Legg · 1906
Earlier work this paper cites.
Tom Everitt and Marcus Hutter · 1908
Earlier work this paper cites.
“The Foundations of Statistics”
Leonard Savage · 1954
Earlier work this paper cites.
“Information Value Theory”
Ronald Howard · 1966
Earlier work this paper cites.
“Sex Bias in Graduate Admissions: Data from Berkeley”
P Bickel, E Hammel and J O’Connell · 1975
Earlier work this paper cites.
“Causal Decision Theory”
Brian Skyrms · 1982
Earlier work this paper cites.
“Influence Diagrams”
Ronald Howard and James Matheson · 1984
Earlier work this paper cites.
“The Intentional Stance”
Daniel Dennett · 1987
Earlier work this paper cites.
“Causal Networks: Semantics and Expressiveness”
Thomas Verma and Judea Pearl · 1988
Earlier work this paper cites.
“On the Logic of Causal Models”
Dan Geiger and Judea Pearl · 1990
Earlier work this paper cites.
“From inuence to relevance to knowledge”
Ronald Howard · 1990
Earlier work this paper cites.
“Using influence diagrams to value information and control”
James Matheson · 1990
Earlier work this paper cites.
“Decision-Theoretic Foundations for Causal Reasoning”
David Heckerman and Ross Shachter · 1995
Earlier work this paper cites.
“Strong Completeness and Faithfulness in Bayesian Networks”
Christopher Meek · 1995
Earlier work this paper cites.
“A note about redundancy in influence diagrams”
Enrico Fagiuoli and Marco Zaffalon · 1998
Earlier work this paper cites.
“Bayes-Ball: The Rational Pastime (for Determining Irrelevance and Requisite Information in Belief Networks and Influence Diagrams)”
Ross Shachter · 1998
Earlier work this paper cites.
“Welldefined decision scenarios”
Thomas Nielsen and Finn Jensen · 1999
Earlier work this paper cites.
“Markov perfect equilibrium. I. Observable actions”
Eric Maskin and Jean Tirole · 2000
Earlier work this paper cites.
“Representing and Solving Decision Problems with Limited Information”
Steffen. Lauritzen and Dennis Nilsson · 2001
Earlier work this paper cites.
“Influence Diagrams for Causal Modelling and Inference”
A Dawid · 2002
Earlier work this paper cites.
“Causal Models, Value of Intervention, and Search for Opportunities”
Tsai-ching Lu and Marek Druzdzel · 2002
Earlier work this paper cites.
“Multi-agent influence diagrams for representing and solving games”
Daphne Koller and Brian Milch · 2003
Earlier work this paper cites.
“Describing and Valuing Interventions That Observe or Control Decision Situations”
David Matheson and James Matheson · 2005
Earlier work this paper cites.
“Elements of Information Theory”
Thomas. Cover and Joy. Thomas · 2006
Cited alongside, same era.
“Interventions and Causal Inference”
Frederick Eberhardt and Richard Scheines · 2007
Cited alongside, same era.
“Universal Intelligence: A definition of machine intelligence”
Shane Legg and Marcus Hutter · 2007
Cited alongside, same era.
“Gödel Machines: Self-Referential Universal Problem Solvers Making Provably Optimal Self-Improvements”
Jürgen Schmidhuber · 2007
Cited alongside, same era.
“Networks of influence diagrams: A formalism for representing agents’ beliefs and decision-making processes”
Ya’akov Gal and Avi Pfeffer · 2008
Cited alongside, same era.
“Ignorable Information in Multi-Agent Scenarios”, 2008
Brian Milch and Daphne Koller · 2008
Cited alongside, same era.
“Safely interruptible agents”
Laurent Orseau and Stuart Armstrong · 2016
Later among the works it cites.
“Causal Decision Theory”
Paul Weirich · 2016
Later among the works it cites.
“Good and safe uses of AI Oracles”, 2017, pp. 1–11
Stuart Armstrong · 2017
Later among the works it cites.
“Low Impact Artificial Intelligences”, 2017
Stuart Armstrong and Benjamin Levinstein · 2017
Later among the works it cites.
“Guidelines for Artificial Intelligence Containment”, 2017
James Babcock, Janos Kramar and Roman. Yampolskiy · 2017
Later among the works it cites.
“Exposing the Probabilistic Causal Structure of Discrimination”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“The Basic AI Drives”
Stephen Omohundro · 2008
Cited alongside, same era.
“Causality: Models, Reasoning, and Inference”
Judea Pearl · 2009
Cited alongside, same era.
“Pearl Causality and the Value of Control”
Ross Shachter and David Heckerman · 2010
Cited alongside, same era.
“Self-modification and mortality in artificial agents”
Laurent Orseau and Mark Ring · 2011
Cited alongside, same era.
“Delusion, Survival, and Intelligent Agents”
Mark Ring and Laurent Orseau · 2011
Cited alongside, same era.
“Thinking inside the box: Controlling and using an oracle AI”
Stuart Armstrong, Anders Sandberg and Nick Bostrom · 2012
Cited alongside, same era.
Francesco Bonchi, Sara Hajian, Bud Mishra and Daniele Ramazzotti · 2017
Later among the works it cites.
“Reinforcement Learning with Corrupted Reward Signal”
Tom Everitt, Victoria Krakovna, Laurent Orseau, Marcus Hutter and Shane Legg · 2017
Later among the works it cites.
“On Formalizing Fairness in Prediction with Machine Learning”, 2017
Pratik Gajane and Mykola Pechenizkiy · 2017
Later among the works it cites.
Dylan Hadfield-Menell, Anca Dragan, Pieter Abbeel and Stuart Russell · 2017
Later among the works it cites.
“Avoiding Discrimination through Causal Reasoning”
Niki Kilbertus et al · 2017
Later among the works it cites.
Matt. Kusner, Joshua. Loftus, Chris Russell and Ricardo Silva · 2017
Later among the works it cites.
Jan Leike et al · 2017
Later among the works it cites.
“A Game-Theoretic Analysis of the Off-Switch Game”
Tobias Wängberg, Mikael Böörs, Elliot Catt, Tom Everitt and Marcus Hutter · 2017
Later among the works it cites.
“Anti-discrimination learning: a causal modeling-based framework”
Lu Zhang and Xintao Wu · 2017
Later among the works it cites.
“The Measure and Mismeasure of Fairness: A Critical Review of Fair Machine Learning”, 2018
Sam Corbett-Davies and Sharad Goel · 2018
Later among the works it cites.
“Towards Safe Artificial General Intelligence”, 2018
Tom Everitt · 2018
Later among the works it cites.
“The Alignment Problem for Bayesian History-Based Reinforcement Learners”, 2018
Tom Everitt and Marcus Hutter · 2018
Later among the works it cites.
“AGI Safety Literature Review”
Tom Everitt, Gary Lea and Marcus Hutter · 2018
Later among the works it cites.
Joel Lehman et al · 2018
Later among the works it cites.
“Reinforcement Learning: An Introduction”
Richard Sutton and Andrew Barto · 2018
Later among the works it cites.
“Path-Specific Counterfactual Fairness”
Silvia Chiappa · 2019
Closest in time.
“Penalizing side effects using stepwise relative reachability”
Victoria Krakovna, Laurent Orseau, Miljan Martic and Shane Legg · 2019
Closest in time.
“Pitfalls in learning a reward function online”
Stuart Armstrong, Laurent Orseau, Jan Leike and Shane Legg · 2020
Closest in time.