Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) trains an agent from experiences interacting with the environment.
N. Carlini, S. Chien, M. Nasr, S. Song, A. Terzis, and F. Tramèr, “Membership inference attacks from first principles,” in 2022 IEEE Symposium on Security and Privacy (SP) , 2022, pp. 1897–1914
1914
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in Proceedings of The 33rd International Conference on Machine Learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
F. E. Grubbs, “Sample criteria for testing outlying observations,” The Annals of Mathematical Statistics , vol. 21, no. 1, pp. 27–58, 1950
1950
Earlier work this paper cites.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Advances in neural information processing systems , vol. 1, 1988
1988
Earlier work this paper cites.
E. Greensmith, P. L. Bartlett, and J. Baxter, “Variance reduction techniques for gradient estimates in reinforcement learning,” J. Mach. Learn. Res. , vol. 5, pp. 1471–1530, 2004
2004
Earlier work this paper cites.
C. Dwork, F. McSherry, K. Nissim, and A. D. Smith, “Calibrating noise to sensitivity in private data analysis,” in TCC , 2006, pp. 265–284
2006
Earlier work this paper cites.
A. Gretton, K. M. Borgwardt, M. J. Rasch, B. Schölkopf, and A. J. Smola, “A kernel two-sample test,” J. Mach. Learn. Res. , vol. 13, pp. 723–773, 2012
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
D. Silver, G. Lever, N. Heess, T. Degris, D. Wierstra, and M. A. Riedmiller, “Deterministic policy gradient algorithms,” in ICML , vol. 32, 2014, pp. 387–395
2014
Earlier work this paper cites.
K. Sohn, H. Lee, and X. Yan, “Learning structured output representation using deep conditional generative models,” in Advances in Neural Information Processing Systems , 2015, pp. 3483–3491
2015
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson et al. , “Openai gym,” arXiv preprint arXiv:1606.01540 , 2016
2016
Earlier work this paper cites.
Z. Izzo, M. Anne Smart, K. Chaudhuri et al. , “Approximate data deletion from machine learning models,” in Proceedings of The 24th International Conference on Artificial Intelligence and Statistics , vol. 130. PMLR, 2021, pp. 2008–2016
2016
Earlier work this paper cites.
R. Shokri, M. Stronati, C. Song, and V. Shmatikov, “Membership inference attacks against machine learning models,” in 2017 IEEE Symposium on Security and Privacy (SP) , 2017, pp. 3–18
2017
Earlier work this paper cites.
S. Tosatto, M. Pirotta, C. D’Eramo, and M. Restelli, “Boosted fitted q-iteration,” in Proceedings of the 34th International Conference on Machine Learning , vol. 70. PMLR, 2017, pp. 3434–3443
2017
Earlier work this paper cites.
S. Fujimoto, H. Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in International Conference on Machine Learning . PMLR, 2018, pp. 1587–1596
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel et al. , “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning , 2018, pp. 1861–1870
2018
Earlier work this paper cites.
M. Lejeune, “California consumer privacy act - erste ansätze einer annäherung zu prinzipien der DSGVO in den USA,” Comput. und Recht , vol. 34, no. 9, pp. 569–576, 2018
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Earlier work this paper cites.
A. Kumar, J. Fu, M. Soh, G. Tucker, and S. Levine, “Stabilizing off-policy q-learning via bootstrapping error reduction,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
Y. Ma, X. Zhang, W. Sun, and J. Zhu, “Policy poisoning in batch reinforcement learning and control,” in Advances in Neural Information Processing Systems , 2019, pp. 14 543–14 553
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
X. Pan, W. Wang, X. Zhang, B. Li, J. Yi, and D. Song, “How you act tells a lot: Privacy-leaking attack on deep reinforcement learning.” in AAMAS , vol. 19, no. 2019, 2019, pp. 368–376
2019
Earlier work this paper cites.
V. M. Panaretos and Y. Zemel, “Statistical aspects of wasserstein distances,” Annual review of statistics and its application , vol. 6, pp. 405–431, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
N. AlHinai, “Introduction to biomedical signal processing and artificial intelligence,” in Biomedical signal processing and artificial intelligence in healthcare . Elsevier, 2020, pp. 1–28
2020
Earlier work this paper cites.
D. Georgiou and C. Lambrinoudakis, “Compatibility of a security policy for a cloud-based healthcare system with the EU general data protection regulation (GDPR),” Inf. , vol. 11, no. 12, p. 586, 2020
2020
Earlier work this paper cites.
A. Golatkar and A. Achille, “Forgetting outside the box: Scrubbing deep networks of information accessible from input-output observations,” in Computer Vision - ECCV 2020 - 16th European Conference, Glasgow, UK, August 23-28, 2020 , vol. 12374, 2020, pp. 383–398
2020
Earlier work this paper cites.
A. Golatkar, A. Achille, and S. Soatto, “Eternal sunshine of the spotless net: Selective forgetting in deep networks,” in 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR , 2020, pp. 9301–9309
2020
Cited alongside, same era.
C. Guo, T. Goldstein, A. Hannun, and L. Van Der Maaten, “Certified data removal from machine learning models,” in Proceedings of the 37th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, vol. 119, 2020, pp. 3832–3842
2020
Cited alongside, same era.
A. Kumar, A. Zhou, G. Tucker, and S. Levine, “Conservative q-learning for offline reinforcement learning,” in NeurIPS , 2020
2020
Cited alongside, same era.
S. Levine, A. Kumar, G. Tucker, and J. Fu, “Offline reinforcement learning: Tutorial, review, and perspectives on open problems,” 2020
2020
Cited alongside, same era.
Y. Li, Y. Bai, Y. Jiang, Y. Yang, S.-T. Xia, and B. Li, “Untargeted backdoor watermark: Towards harmless and stealthy dataset copyright protection,” Advances in Neural Information Processing Systems , vol. 35, pp. 13 238–13 250, 2022
2022
Later among the works it cites.
T. Seno and M. Imai, “d3rlpy: An offline deep reinforcement learning library,” Journal of Machine Learning Research , vol. 23, no. 315, pp. 1–20, 2022
2022
Later among the works it cites.
B. Singh, R. Kumar, and V. P. Singh, “Reinforcement learning in robotic applications: a comprehensive survey,” Artif. Intell. Rev. , 2022
2022
Later among the works it cites.
A. Thudi, G. Deza, V. Chandrasekaran, and N. Papernot, “Unrolling SGD: understanding factors influencing machine unlearning,” in 7th IEEE European Symposium on Security and Privacy, EuroS&P 2022, Genoa, Italy, June 6-10, 2022 . IEEE, 2022, pp. 303–319
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Andrychowicz, A. Raichuk, P. Stańczyk et al. , “What matters for on-policy deep actor-critic methods? a large-scale study,” in The International Conference on Learning Representations (ICLR) , 2021
2021
Cited alongside, same era.
F. Boenisch, “A systematic review on model watermarking for neural networks,” Frontiers in big Data , vol. 4, p. 729663, 2021
2021
Cited alongside, same era.
L. Bourtoule, V. Chandrasekaran, C. A. Choquette-Choo et al. , “Machine unlearning,” in 2021 IEEE Symposium on Security and Privacy (SP) , 2021, pp. 141–159
2021
Cited alongside, same era.
J. Brophy and D. Lowd, “Machine unlearning for random forests,” in Proceedings of the 38th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, M. Meila and T. Zhang, Eds., vol. 139, 18–24 Jul 2021, pp. 1092–1104
2021
Cited alongside, same era.
M. Chen, Z. Zhang, T. Wang, M. Backes, M. Humbert, and Y. Zhang, “When machine unlearning jeopardizes privacy,” in Proceedings of the 2021 ACM SIGSAC conference on computer and communications security , 2021, pp. 896–911
2021
Cited alongside, same era.
M. Fatemi, T. W. Killian, J. Subramanian et al. , “Medical dead-ends and learning to identify high-risk states and treatments,” in Advances in Neural Information Processing Systems , 2021, pp. 4856–4870
2021
Cited alongside, same era.
J. Fu, A. Kumar, O. Nachum, G. Tucker, and S. Levine, “D4rl: Datasets for deep data-driven reinforcement learning,” 2021
2021
Cited alongside, same era.
S. Fujimoto and S. S. Gu, “A minimalist approach to offline reinforcement learning,” in Advances in Neural Information Processing Systems , vol. 34. Curran Associates, Inc., 2021, pp. 20 132–20 145
2021
Cited alongside, same era.
S. Verma, J. Fu, S. Yang, and S. Levine, “CHAI: A chatbot AI for task-oriented dialogue with offline reinforcement learning,” in Proceedings of the North American Chapter of the Association for Computational Linguistics, NAACL , 2022
2022
Later among the works it cites.
H. Yan, X. Li, Z. Guo et al. , “ARCANE: an efficient architecture for exact machine unlearning,” in Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence, IJCAI 2022 , pp. 4006–4013
2022
Later among the works it cites.
X. Zhan, H. Xu, Y. Zhang, X. Zhu, H. Yin, and Y. Zheng, “Deepthermal: Combustion optimization for thermal power generating units using offline reinforcement learning,” in AAAI , 2022
2022
Later among the works it cites.
Q. Zhang, J. Liu, Y. Dai et al. , “Multi-task fusion via reinforcement learning for long-term user satisfaction in recommender systems,” in The 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining , 2022, pp. 4510–4520
2022
Later among the works it cites.
V. S. Chundawat, A. K. Tarun, M. Mandal, and M. Kankanhalli, “Zero-shot machine unlearning,” IEEE Transactions on Information Forensics and Security , vol. 18, pp. 2345–2354, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
H. Emerson, M. Guy, and R. McConville, “Offline reinforcement learning for safer blood glucose control in people with type 1 diabetes,” J. Biomed. Informatics , 2023
2023
Later among the works it cites.
Z. Li, F. Nie, Q. Sun et al. , “Boosting offline reinforcement learning for autonomous driving with hierarchical latent skills,” 2023
2023
Later among the works it cites.
M. Nambiar, S. Ghosh, P. Ong, Y. E. Chan, Y. M. Bee, and P. Krishnaswamy, “Deep offline reinforcement learning for real-world treatment optimization applications,” in Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining, KDD , 2023
2023
Later among the works it cites.
R. F. Prudencio, M. R. O. A. Maximo, and E. L. Colombini, “A survey on offline reinforcement learning: Taxonomy, review, and open problems,” IEEE Transactions on Neural Networks and Learning Systems , 2023
2023
Later among the works it cites.
P. Sodhi, F. Wu, E. R. Elenberg, K. Q. Weinberger, and R. McDonald, “On the effectiveness of offline RL for dialogue response generation,” in International Conference on Machine Learning, ICML , vol. 202. PMLR, 2023, pp. 32 088–32 104
2023
Later among the works it cites.
S. Wang, X. Chen, D. Jannach, and L. Yao, “Causal decision transformer for recommender systems via offline reinforcement learning,” 2023
2023
Later among the works it cites.
A. Warnecke, L. Pirch, C. Wressnegger, and K. Rieck, “Machine unlearning of features and labels,” in 30th Annual Network and Distributed System Security Symposium, NDSS 2023, San Diego, California, USA, February 27 - March 3, 2023 , 2023
2023
Later among the works it cites.
Y. Wu, J. McMahan, X. Zhu, and Q. Xie, “Reward poisoning attacks on offline multi-agent reinforcement learning,” in Thirty-Seventh AAAI Conference on Artificial Intelligence, AAAI , B. Williams, Y. Chen, and J. Neville, Eds. AAAI Press, 2023, pp. 10 426–10 434
2023
Later among the works it cites.
Y. Yang and ufuk topcu, “Value-based membership inference attack on actor-critic reinforcement learning,” 2023. [Online]. Available: https://openreview.net/forum?id=wKIxJKTDmX-
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2024
Closest in time.
2024
Closest in time.
H. Xu, T. Zhu, L. Zhang, W. Zhou, and P. S. Yu, “Machine unlearning: A survey,” ACM Comput. Surv. , vol. 56, no. 1, pp. 9:1–9:36, 2024
2024
Closest in time.
J. Yao, E. Chien, M. Du, X. Niu, T. Wang, Z. Cheng, and X. Yue, “Machine unlearning of pre-trained large language models,” 2024
2024
Closest in time.
2024
Closest in time.
S. Fujimoto, D. Meger, and D. Precup, “Off-policy deep reinforcement learning without exploration,” in International Conference on Machine Learning , 2019, pp. 2052–2062
2062
Closest in time.