Back to results

University of Technology Sydney

Differential Privacy in Reinforcement Learning

Abstract

dc:description.abstract

Reinforcement learning is a principled AI framework for autonomously experience-driven learning. The primary goal of reinforcement learning is to train autonomous agents to learn the optimal behaviors for their interactive environments. Deep reinforcement learning promotes a higher-level understanding of the visual world in the field of reinforcement learning by combining deep learning models and reinforcement learning algorithms. Since reinforcement learning is achieving great success in an increasing number of application fields that may involve huge amounts of private information, the security of policies and privacy preservation in reinforcement learning have given rise to widespread concerns. In addition, deep reinforcement learning policies parameterized by neural networks have been demonstrated to be vulnerable to adversarial attacks in supervised learning settings. Privacy leakage also occurs in multi-agent reinforcement learning systems where agents’ actions or behaviors are directly exposed to other agents. To address these multiple privacy concerns in reinforcement learning, we apply differential privacy in variant scenarios of reinforcement learning. In this thesis, we introduce our differentially private methods in those diverse scenarios to preserve privacy, including the multi-agent advising framework, multi-agent planning framework, the deep reinforcement learning context, machine learning classifiers and multi-agent game theoretic framework, respectively. We have provided detailed theoretical analysis and comprehensive experimental results to demonstrate that our methods can guarantee privacy preservation as well as the utility of reinforcement learning in diverse scenario in different chapters.

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Shen, Sheng

Rights

dc:rights
Statement dc:rights
  • info:eu-repo/semantics/openAccess
  • The author owns the copyright in this thesis including all reproduction and reuse rights for the work. The work may not be altered without the permission of the copyright owner. Attribution is essential when quoting or paraphrasing from this thesis.
  • © 2022 Sheng Shen
  • au.edu.uts.lib/ppc
Language dc:language.iso
en_US

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/10453/168914
OAI identifier oai:identifier
oai:opus.lib.uts.edu.au:10453/168914

Chain of custody

source
Harvested from
University of Technology Sydney
Base URL
opus.lib.uts.edu.au/oai/request
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
related terms
citation

Shen, Sheng. Differential Privacy in Reinforcement Learning. 2022. http://hdl.handle.net/10453/168914