Fetching the paper…
Reading the bibliography…
Restless multi-armed bandits (RMAB) have demonstrated success in optimizing resource allocation for large beneficiary populations in public health settings.
Restless bandits: Activity allocation in a changing world
Peter Whittle · 1988
Earlier work this paper cites.
On an index policy for restless bandits
Richard R Weber and Gideon Weiss · 1990
Earlier work this paper cites.
Accountability for reasonableness: Establishing a fair process for priority setting is easier than agreeing on principles, 2000
Norman Daniels · 2000
Earlier work this paper cites.
Priority setting of health interventions: the need for multi-criteria decision analysis
Rob Baltussen and Louis Niessen · 2006
Earlier work this paper cites.
Some indexable families of restless bandit problems
Kevin D Glazebrook, Diego Ruiz-Hernandez, and Christopher Kirkbride · 2006
Earlier work this paper cites.
Where do rewards come from
Satinder Singh, Richard L Lewis, and Andrew G Barto · 2009
Earlier work this paper cites.
General notions of indexability for queueing control and asset management
Kevin D Glazebrook, David J Hodge, and Chris Kirkbride · 2011
Earlier work this paper cites.
From efficacy to equity: Literature review of decision criteria for resource allocation and healthcare decisionmaking
Lalla Aïda Guindo, Monika Wagner, Rob Baltussen, Donna Rindress, Janine van Til, Paul Kind, and Mireille M Goetghebeur · 2012
Earlier work this paper cites.
Improving health outcomes through better capacity allocation in a community-based chronic care model
Sarang Deo, Seyed Iravani, Tingting Jiang, Karen Smilowitz, and Stephen Samuelson · 2013
Earlier work this paper cites.
Transforming our world: the 2030 agenda for sustainable development, 2015
Department of Economic United Nations and Social Affairs Sustainable Development · 2015
Earlier work this paper cites.
Strategies towards ending preventable maternal mortality (epmm)
World Health Organization et al · 2015
Earlier work this paper cites.
Reinforcement Learning from Demonstration through Shaping
Tim Brys, Anna Harutyunyan, Halit Bener Suay, Sonia Chernova, and Matthew E Taylor · 2015
Earlier work this paper cites.
Disparities in the use of a mhealth medication adherence promotion intervention for low-income adults with type 2 diabetes
Lyndsay A Nelson, Shelagh A Mulvaney, Tebeb Gebretsadik, Yun-Xian Ho, Kevin B Johnson, and Chandra Y Osborn · 2016
Earlier work this paper cites.
Deep reward shaping from demonstrations
Ahmed Hussein, Eyad Elyan, Mohamed Medhat Gaber, and Chrisina Jayne · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Equity in healthcare resource allocation decision making: a systematic review
Haylee Lane, Mitchell Sarkies, Jennifer Martin, and Terry Haines · 2017
Earlier work this paper cites.
Allocation of health care resources: Principles for decision-making
Hellen Ransom and John M Olsson · 2017
Earlier work this paper cites.
High-quality health systems in the sustainable development goals era: time for a revolution
Margaret E Kruk, Anna D Gage, Catherine Arsenault, Keely Jordan, Hannah H Leslie, Sanam Roder-DeWan, Olusoji Adeyi, Pierre Barker, Bernadette Daelmans, Svetlana V Doubova, et al · 2018
Earlier work this paper cites.
Strategies to improve treatment coverage in community-based public health programs: A systematic review of the literature
Katrina Deardorff, Arianna Rubin Means, Kristjana Hrönn Ásbjörnsdóttir, and Judd L. Walson · 2018
Earlier work this paper cites.
Ensuring fairness in machine learning to advance health equity
Alvin Rajkomar, Michaela Hardt, Michael D Howell, Greg Corrado, and Marshall H Chin · 2018
Earlier work this paper cites.
Prioritizing hepatitis c treatment in us prisons
Turgay Ayer, Can Zhang, Anthony Bonifonte, Anne C Spaulding, and Jagpreet Chhatwal · 2019
Earlier work this paper cites.
Communicating about infectious disease threats: Insights from public health information officers
Yan Jin, Lucinda Austin, Santosh Vijaykumar, Hyoyeun Jun, and Glen Nowak · 2019
Earlier work this paper cites.
Transformers as soft reasoners over language, 2020
Peter Clark, Oyvind Tafjord, and Kyle Richardson · 2020
Cited alongside, same era.
Mobile health technology to improve maternal health awareness in tribal populations: mobile for mothers
Avishek Choudhury, Onur Asan, and Murari M Choudhury · 2021
Cited alongside, same era.
Annual report 2020-2021, 2021
ARMMAN · 2021
Cited alongside, same era.
Neurwin: Neural whittle index network for restless bandits via deep rl
Khaled Nakhleh, Santosh Ganji, Ping-Chun Hsieh, I Hou, Srinivas Shakkottai, et al · 2021
Cited alongside, same era.
Deep reinforcement learning at the edge of the statistical precipice
Rishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron C Courville, and Marc Bellemare · 2021
Cited alongside, same era.
Lm-nav: Robotic navigation with large pre-trained models of language, vision, and action
Dhruv Shah, Błażej Osiński, Sergey Levine, et al · 2023
Later among the works it cites.
Chatgpt for robotics: Design principles and model abilities
Sai Vemprala, Rogerio Bonatti, Arthur Bucker, and Ashish Kapoor · 2023
Later among the works it cites.
Chatgpt applications in medical, dental, pharmacy, and public health education: A descriptive study highlighting the advantages and limitations
Malik Sallam, Nesreen A Salim, Muna Barakat, and B Ala’a · 2023
Later among the works it cites.
Nlp for maternal healthcare: Perspectives and guiding principles in the age of llms
AAKANKSHA NAIK, CARLA S ALVARADO, LUCY LU WANG, and IRENE CHEN · 2023
Later among the works it cites.
Robust planning over restless groups: engagement interventions for a large-scale maternal telehealth program
Jackson A Killian, Arpita Biswas, Lily Xu, Shresth Verma, Vineet Nair, Aparna Taneja, Aparna Hegde, Neha Madhiwalla, Paula Rodriguez Diaz, Sonja Johnson-Yu, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gail Weiss, Yoav Goldberg, and Eran Yahav · 2021
Cited alongside, same era.
Assessing the impact of mobile-based intervention on health literacy among pregnant women in urban india, 2019
ARMMAN · 2022
Cited alongside, same era.
Informal care in times of a public health crisis: Objective burden, subjective burden and quality of life of caregivers in the netherlands during the covid-19 pandemic
Leonoor Gräler, Leonarda G. M. Bremmers, Pieter Bakx, Job van Exel, and Marianne E van Bochove · 2022
Cited alongside, same era.
Pre-trained language models for interactive decision-making
Shuang Li, Xavier Puig, Chris Paxton, Yilun Du, Clinton Wang, Linxi Fan, Tao Chen, De-An Huang, Ekin Akyürek, Anima Anandkumar, et al · 2022
Cited alongside, same era.
Field study in deploying restless multi-armed bandits: Assisting non-profits in improving maternal and child health
Aditya Mate, Lovish Madaan, Aparna Taneja, Neha Madhiwalla, Shresth Verma, Gargi Singh, Aparna Hegde, Pradeep Varakantham, and Milind Tambe · 2022
Cited alongside, same era.
Pandemics and technology engagement: New evidence from m-health intervention during covid-19 in india
Sawan Rathi, Anindya S Chakrabarti, Chirantan Chatterjee, and Aparna Hegde · 2022
Cited alongside, same era.
Restless and uncertain: Robust policies for restless bandits via deep multi-agent reinforcement learning
Jackson A Killian, Lily Xu, Arpita Biswas, and Milind Tambe · 2022
Cited alongside, same era.
Later among the works it cites.
Scalable decision-focused learning in restless multi-armed bandits with application to maternal and child health
Kai Wang, Shresth Verma, Aditya Mate, Sanket Shah, Aparna Taneja, Neha Madhiwalla, Aparna Hegde, and Milind Tambe · 2023
Later among the works it cites.
The perils of trial-and-error reward design: misdesign through overfitting and invalid task specifications
Serena Booth, W Bradley Knox, Julie Shah, Scott Niekum, Peter Stone, and Alessandro Allievi · 2023
Later among the works it cites.
Guiding pretraining in reinforcement learning with large language models
Yuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas, Trevor Darrell, Pieter Abbeel, Abhishek Gupta, and Jacob Andreas · 2023
Later among the works it cites.
Auto mc-reward: Automated dense reward design with large language models for minecraft
Hao Li, Xue Yang, Zhaokai Wang, Xizhou Zhu, Jie Zhou, Yu Qiao, Xiaogang Wang, Hongsheng Li, Lewei Lu, and Jifeng Dai · 2023
Later among the works it cites.
Increasing impact of mobile health programs: Saheli for maternal and child care
Shresth Verma, Gargi Singh, Aditya Mate, Paritosh Verma, Sruthi Gorantla, Neha Madhiwalla, Aparna Hegde, Divy Thakkar, Manish Jain, Milind Tambe, et al · 2023
Later among the works it cites.
Restless multi-armed bandits for maternal and child health: Results from decision-focused learning
Shresth Verma, Aditya Mate, Kai Wang, Neha Madhiwalla, Aparna Hegde, Aparna Taneja, and Milind Tambe · 2023
Later among the works it cites.
Tracr: Compiled transformers as a laboratory for interpretability, 2023
David Lindner, János Kramár, Sebastian Farquhar, Matthew Rahtz, Thomas McGrath, and Vladimir Mikulik · 2023
Later among the works it cites.
Eureka: Human-level reward design via coding large language models
Yecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang, Osbert Bastani, Dinesh Jayaraman, Yuke Zhu, Linxi Fan, and Anima Anandkumar · 2024
Closest in time.
Niclas Boehmer, Yunfan Zhao, Guojun Xiong, Paula Rodriguez-Diaz, Paola Del Cueto Cibrian, Joseph Ngonzi, Adeline Boatin, and Milind Tambe · 2024
Closest in time.
Reflexion: Language agents with verbal reinforcement learning
Noah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao · 2024
Closest in time.
Towards a pretrained model for restless bandits via multi-arm generalization
Yunfan Zhao, Nikhil Behari, Edward Hughes, Edwin Zhang, Dheeraj Nagaraj, Karl Tuyls, Aparna Taneja, and Milind Tambe · 2024
Closest in time.
Improving the prediction of individual engagement in recommendations using cognitive models
Roderick Seow, Yunfan Zhao, Duncan Wood, Milind Tambe, and Cleotilde Gonzalez · 2024
Closest in time.
The bandit whisperer: Communication learning for restless bandits
Yunfan Zhao, Tonghan Wang, Dheeraj Nagaraj, Aparna Taneja, and Milind Tambe · 2024
Closest in time.
Online restless multi-armed bandits with long-term fairness constraints
Shufan Wang, Guojun Xiong, and Jian Li · 2024
Closest in time.
Group fairness in predict-then-optimize settings for restless bandits
Shresth Verma, Yunfan Zhao, Sanket Shah, Niclas Boehmer, Aparna Taneja, and Milind Tambe · 2024
Closest in time.
Balancing act: Prioritization strategies for llm-designed restless bandit rewards
Shresth Verma, Niclas Boehmer, Lingkai Kong, and Milind Tambe · 2024
Closest in time.