Fetching the paper…
Reading the bibliography…
In many real-world applications of reinforcement learning (RL), deployed policies have varied impacts on different stakeholders, creating challenges in reaching consensus on how to effectively aggregate their preferences.
Interpersonal Comparability and Social Choice Theory
Kevin W. S. Roberts · 1980
Earlier work this paper cites.
Restless bandits: Activity allocation in a changing world
P. Whittle · 1988
Earlier work this paper cites.
The complexity of optimal queuing network control
Christos H Papadimitriou and John N Tsitsiklis · 1999
Earlier work this paper cites.
Hierarchical reinforcement learning with the maxq value function decomposition
Thomas G Dietterich · 2000
Earlier work this paper cites.
Handbook of Means and Their Inequalities
P.S Bullen · 2003
Earlier work this paper cites.
Fair Division and Collective Welfare
Hervé Moulin · 2003
Earlier work this paper cites.
Simultaneous optimization via approximate majorization for concave profits or convex costs
Ashish Goel and Adam Meyerson · 2006
Earlier work this paper cites.
All-norms and all-l_p-norms approximation algorithms
Daniel Golovin, Anupam Gupta, Amit Kumar, and Kanat Tangwongsan · 2008
Earlier work this paper cites.
Responsive elastic computing
Julien Perez, Cécile Germain-Renaud, Balázs Kégl, and Charles Loomis · 2009
Earlier work this paper cites.
A survey of multi-objective sequential decision-making
Diederik M. Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley · 2013
Earlier work this paper cites.
Policy gradient approaches for multi-objective sequential decision making
Simone Parisi, Matteo Pirotta, Nicola Smacchia, Luca Bascetta, and Marcello Restelli · 2014
Earlier work this paper cites.
Multi-objective reinforcement learning using sets of pareto dominating policies
Kristof Van Moffaert and Ann Nowé · 2014
Earlier work this paper cites.
Joint trajectory and communication design for multi-uav enabled wireless networks
Qingqing Wu, Yong Zeng, and Rui Zhang · 2017
Earlier work this paper cites.
Approximation algorithms for minimum norm and ordered optimization problems
Deeparnab Chakrabarty and Chaitanya Swamy · 2019
Cited alongside, same era.
Multi-objective multi-agent decision making: a utility-based analysis and survey
Roxana Rădulescu, Patrick Mannion, Diederik M. Roijers, and Ann Nowé · 2019
Cited alongside, same era.
A generalized algorithm for multi-objective reinforcement learning and policy adaptation
Runzhe Yang, Xingyuan Sun, and Karthik Narasimhan · 2019
Cited alongside, same era.
Bringing fairness to actor-critic reinforcement learning for network utility optimization
Jingdi Chen, Yimeng Wang, and Tian Lan · 2021
Cited alongside, same era.
Multi-objective reinforcement learning with non-linear scalarization
Mridul Agarwal, Vaneet Aggarwal, and Tian Lan · 2022
Cited alongside, same era.
Welfare and fairness in multi-objective reinforcement learning
Socially fair reinforcement learning, 2023
Debmalya Mandal and Jiarui Gan · 2023
Later among the works it cites.
Expanding impact of mobile health programs: Saheli for maternal and child care
Shresth Verma, Gargi Singh, Aditya Mate, Paritosh Verma, Sruthi Gorantla, Neha Madhiwalla, Aparna Hegde, Divy Thakkar, Manish Jain, Milind Tambe, et al · 2023
Later among the works it cites.
Policy aggregation
Parand A. Alamdari, Soroush Ebadian, and Ariel D. Procaccia · 2024
Later among the works it cites.
Armman: Advancing reduction in mortality and morbidity of mothers, children, and neonates, 2024
ARMMAN · 2024
Later among the works it cites.
MaxMin-RLHF: Alignment with diverse human preferences
Souradip Chakraborty, Jiahao Qiu, Hui Yuan, Alec Koppel, Dinesh Manocha, Furong Huang, Amrit Bedi, and Mengdi Wang · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zimeng Fan, Nianli Peng, Muhang Tian, and Brandon Fain · 2022
Cited alongside, same era.
A practical guide to multi-objective reinforcement learning and planning
Conor F. Hayes, Roxana Rădulescu, Eugenio Bargiacchi, Johan Källström, Matthew Macfarlane, Mathieu Reymond, Timothy Verstraeten, Luisa M. Zintgraf, Richard Dazeley, Fredrik Heintz, Enda Howley, Athirai A. Irissappane, Patrick Mannion, Ann Nowé, Gabriel Ramos, Marcello Restelli, Peter Vamplew, and Diederik M. Roijers · 2022
Cited alongside, same era.
Pareto conditioned networks
Mathieu Reymond, Eugenio Bargiacchi, and Ann Nowé · 2022
Cited alongside, same era.
Sample-efficient multi-objective learning via generalized policy improvement prioritization
Lucas N. Alegre, Diederik M. Roijers, Ann Nowé, Ana L. C. Bazzan, and Bruno C. da Silva · 2023
Cited alongside, same era.
Revisiting fair-pac learning and the axioms of cardinal welfare
Cyrus Cousins · 2023
Cited alongside, same era.
Welfare and fairness in multi-objective reinforcement learning
Ziming Fan, Nianli Peng, Muhang Tian, and Brandon Fain · 2023
Cited alongside, same era.
Which lp norm is the fairest? Approximations for fair facility location across all "p"
Swati Gupta, Jai Moondra, and Mohit Singh · 2023
Cited alongside, same era.
Cyrus Cousins, Kavosh Asadi, Elita Lobo, and Michael Littman · 2024
Later among the works it cites.
Data-driven solution portfolios
Marina Drygala, Silvio Lattanzi, Andreas Maggiori, Miltiadis Stouras, Ola Svensson, and Sergei Vassilvitskii · 2024
Later among the works it cites.
Achieving fairness in multi-agent MDP using reinforcement learning
Peizhong Ju, Arnob Ghosh, and Ness Shroff · 2024
Later among the works it cites.
Learning social welfare functions
Kanad Shrikar Pardeshi, Itai Shapira, Ariel D. Procaccia, and Aarti Singh · 2024
Later among the works it cites.
RLHF from heterogeneous feedback via personalization and preference aggregation, 2024
Chanwoo Park, Mingyang Liu, Dingwen Kong, Kaiqing Zhang, and Asuman Ozdaglar · 2024
Later among the works it cites.
Fair deep reinforcement learning with generalized gini welfare functions
Guanbao Yu, Umer Siddique, and Paul Weng · 2024
Later among the works it cites.
Provable multi-party reinforcement learning with diverse human feedback, 2024
Huiying Zhong, Zhun Deng, Weijie J. Su, Zhiwei Steven Wu, and Linjun Zhang · 2024
Later among the works it cites.
Balancing notions of equity: Trade-offs between fair portfolio sizes and achievable guarantees
Swati Gupta, Jai Moondra, and Mohit Singh · 2025
Closest in time.