Papers
Topics
Authors
Recent
Search
2000 character limit reached

Differentially Private Exploration in Reinforcement Learning with Linear Representation

Published 2 Dec 2021 in cs.LG | (2112.01585v2)

Abstract: This paper studies privacy-preserving exploration in Markov Decision Processes (MDPs) with linear representation. We first consider the setting of linear-mixture MDPs (Ayoub et al., 2020) (a.k.a.\ model-based setting) and provide an unified framework for analyzing joint and local differential private (DP) exploration. Through this framework, we prove a O~(K<sup>3/4/ϵ)\widetilde{O}(K<sup>{3/4}/\sqrt{\epsilon}) regret bound for (ϵ,δ)(\epsilon,\delta)-local DP exploration and a O~(K/ϵ)\widetilde{O}(\sqrt{K/\epsilon}) regret bound for (ϵ,δ)(\epsilon,\delta)-joint DP. We further study privacy-preserving exploration in linear MDPs (Jin et al., 2020) (a.k.a.\ model-free setting) where we provide a O~(K<sup>35/ϵ<sup>25)\widetilde{O}\left(K<sup>{\frac{3}{5}}/\epsilon<sup>{\frac{2}{5}}\right) regret bound for (ϵ,δ)(\epsilon,\delta)-joint DP, with a novel algorithm based on low-switching. Finally, we provide insights into the issues of designing local DP algorithms in this model-free setting.

Citations (10)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.