Fetching the paper…
Reading the bibliography…
This paper studies privacy-preserving exploration in Markov Decision Processes (MDPs) with linear representation.
Weighted sums of certain dependent random variables
Kazuoki Azuma · 1967
Earlier work this paper cites.
Stochastic optimal control: the discrete-time case
Dimitir P Bertsekas and Steven Shreve · 2004
Earlier work this paper cites.
Adaptive estimation of a quadratic functional of a density by model selection
Béatrice Laurent · 2005
Earlier work this paper cites.
Calibrating noise to sensitivity in private data analysis
Cynthia Dwork, Frank McSherry, Kobbi Nissim, and Adam D. Smith · 2006
Earlier work this paper cites.
Private and continual release of statistics
T.-H. Hubert Chan, Elaine Shi, and Dawn Song · 2010
Earlier work this paper cites.
Differential privacy under continual observation
Cynthia Dwork, Moni Naor, Toniann Pitassi, and Guy N. Rothblum · 2010
Earlier work this paper cites.
Boosting and differential privacy
Cynthia Dwork, Guy N. Rothblum, and Salil Vadhan · 2010
Earlier work this paper cites.
A multiplicative weights mechanism for privacy-preserving data analysis
Moritz Hardt and Guy N. Rothblum · 2010
Earlier work this paper cites.
Improved algorithms for linear stochastic bandits
Yasin Abbasi-yadkori, Dávid Pál, and Csaba Szepesvári · 2011
Earlier work this paper cites.
Topics in random matrix theory
Terence Tao · 2011
Earlier work this paper cites.
Matrix Theory: Basic Results and Techniques
Fuzhen Zhang · 2011
Earlier work this paper cites.
The algorithmic foundations of differential privacy
Cynthia Dwork, Aaron Roth, et al · 2014
Earlier work this paper cites.
Rappor: Randomized aggregatable privacy-preserving ordinal response
Úlfar Erlingsson, Vasyl Pihur, and Aleksandra Korolova · 2014
Earlier work this paper cites.
Mechanism design in large games: incentives and privacy
Michael J. Kearns, Mallesh M. Pai, Aaron Roth, and Jonathan R. Ullman · 2014
Cited alongside, same era.
The composition theorem for differential privacy, 2015
Peter Kairouz, Sewoong Oh, and Pramod Viswanath · 2015
Cited alongside, same era.
(nearly) optimal differentially private stochastic multi-arm bandits
Nikita Mishra and Abhradeep Thakurta · 2015
Cited alongside, same era.
Nearly minimax optimal reinforcement learning for linear mixture markov decision processes
Dongruo Zhou, Quanquan Gu, and Csaba Szepesvari · 2015
Cited alongside, same era.
Deep learning with differential privacy
Martin Abadi, Andy Chu, Ian Goodfellow, H Brendan McMahan, Ilya Mironov, Kunal Talwar, and Li Zhang · 2016
Cited alongside, same era.
Private matchings and allocations
Justin Hsu, Zhiyi Huang, Aaron Roth, Tim Roughgarden, and Zhiwei Steven Wu · 2016
An optimal private stochastic-mab algorithm based on optimal private stopping rule
Touqir Sajed and Or Sheffet · 2019
Later among the works it cites.
Model-based reinforcement learning with value-targeted regression
Alex Ayoub, Zeyu Jia, Csaba Szepesvári, Mengdi Wang, and Lin Yang · 2020
Later among the works it cites.
Near-linear time gaussian process optimization with adaptive batching and resparsification
Daniele Calandriello, Luigi Carratino, Alessandro Lazaric, Michal Valko, and Lorenzo Rosasco · 2020
Later among the works it cites.
Local differentially private regret minimization in reinforcement learning
Evrard Garcelon, Vianney Perchet, Ciara Pike-Burke, and Matteo Pirotta · 2020
Later among the works it cites.
Model-based reinforcement learning with value-targeted regression
Zeyu Jia, Lin Yang, Csaba Szepesvari, and Mengdi Wang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Algorithms for differentially private multi-armed bandits
Aristide C. Y. Tossou and Christos Dimitrakakis · 2016
Cited alongside, same era.
The us census bureau adopts differential privacy
John M Abowd · 2018
Cited alongside, same era.
Is q-learning provably efficient?
Chi Jin, Zeyuan Allen-Zhu, Sebastien Bubeck, and Michael I Jordan · 2018
Cited alongside, same era.
Differentially private contextual linear bandits
Roshan Shariff and Or Sheffet · 2018
Cited alongside, same era.
The privacy blanket of the shuffle model
Borja Balle, James Bell, Adrià Gascón, and Kobbi Nissim · 2019
Cited alongside, same era.
Distributed differential privacy via shuffling
Albert Cheu, Adam D. Smith, Jonathan R. Ullman, David Zeber, and Maxim Zhilyaev · 2019
Cited alongside, same era.
Provably efficient reinforcement learning with linear function approximation
Chi Jin, Zhuoran Yang, Zhaoran Wang, and Michael I. Jordan · 2020
Later among the works it cites.
Private reinforcement learning with pac and regret guarantees
Giuseppe Vietri, Borja de Balle Pigem, Akshay Krishnamurthy, and Steven Wu · 2020
Later among the works it cites.
Locally differentially private (contextual) bandits learning
Kai Zheng, Tianle Cai, Weiran Huang, Zhenguo Li, and Liwei Wang · 2020
Later among the works it cites.
A provably efficient algorithm for linear markov decision process with low switching cost, 2021
Minbo Gao, Tianle Xie, Simon S. Du, and Lin F. Yang · 2021
Closest in time.
Homomorphically encrypted linear contextual bandit
Evrard Garcelon, Vianney Perchet, and Matteo Pirotta · 2021
Closest in time.
Locally differentially private reinforcement learning for linear mixture markov decision processes
Chonghua Liao, Jiafan He, and Quanquan Gu · 2021
Closest in time.
Provably efficient reinforcement learning with linear function approximation under adaptivity constraints, 2021
Tianhao Wang, Dongruo Zhou, and Quanquan Gu · 2021
Closest in time.