Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning

Sajad Khodadadian; Pranay Sharma; Gauri Joshi; Siva Theja Maguluri

Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning

Sajad Khodadadian, Pranay Sharma, Gauri Joshi, Siva Theja Maguluri

TL;DR

This paper proposes a general framework for federated stochastic approximation with Markovian noise and heterogeneity, and applies this framework to federated reinforcement learning algorithms, examining the convergence of federated on-policy TD, off-policy TD, and Q-learning.

Abstract

Since reinforcement learning algorithms are notoriously data-intensive, the task of sampling observations from the environment is usually split across multiple agents. However, transferring these observations from the agents to a central location can be prohibitively expensive in terms of communication cost, and it can also compromise the privacy of each agent's local behavior policy. Federated reinforcement learning is a framework in which $N$ agents collaboratively learn a global model, without sharing their individual data and policies. This global model is the unique fixed point of the average of $N$ local operators, corresponding to the $N$ agents. Each agent maintains a local copy of the global model and updates it using locally sampled data. In this paper, we show that by careful collaboration of the agents in solving this joint fixed point problem, we can find the global model $N$ times faster, also known as linear speedup. We first propose a general framework for federated stochastic approximation with Markovian noise and heterogeneity, showing linear speedup in convergence. We then apply this framework to federated reinforcement learning algorithms, examining the convergence of federated on-policy TD, off-policy TD, and $Q$-learning.

Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning

TL;DR

Abstract

agents collaboratively learn a global model, without sharing their individual data and policies. This global model is the unique fixed point of the average of

local operators, corresponding to the

agents. Each agent maintains a local copy of the global model and updates it using locally sampled data. In this paper, we show that by careful collaboration of the agents in solving this joint fixed point problem, we can find the global model

times faster, also known as linear speedup. We first propose a general framework for federated stochastic approximation with Markovian noise and heterogeneity, showing linear speedup in convergence. We then apply this framework to federated reinforcement learning algorithms, examining the convergence of federated on-policy TD, off-policy TD, and

-learning.

Paper Structure (41 sections, 43 theorems, 260 equations, 5 algorithms)

This paper contains 41 sections, 43 theorems, 260 equations, 5 algorithms.

Introduction
Related Work
Heterogeneous Federated Stochastic Approximation with Markovian Noise
Single node
Generalized Heterogeneous Federated Stochastic Approximation
Main result
Single agent Reinforcement Learning
Temporal Difference Learning
Control Problem and --learning
On-policy Federated Reinforcement Learning
Off-Policy Federated Reinforcement Learning with Tabular Representation
Federated Off-Policy TD-learning
Off-policy TD-learning
Federated off-policy tabular TD-learning
Federated --learning
...and 26 more sections

Key Result

Theorem 3.1

Consider the federated heterogeneous stochastic approximation Algorithm alg:fed_stoch_app with $c_{FSAM}= 1 - \frac{\alpha {\varphi}_2}{2}\in (0,1)$ (${\varphi}_2$ is defined in Equation eq:def:constants in the appendix), and synchronization frequency $K$. Denote ${\boldsymbol \theta_t}=\frac{1}{N}\ where $\mathcal{C}_i$, $i=1,2,3,4, 5$ are some constants which are specified precisely in Appendix

Theorems & Definitions (84)

Theorem 3.1
Remark
Corollary 3.1.1
Theorem 5.1
Remark
Theorem 6.1
Theorem 6.2
Theorem 7.1
Theorem B.1
Corollary B.1.1
...and 74 more

Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning

TL;DR

Abstract

Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning

Authors

TL;DR

Abstract

Table of Contents

Key Result

Theorems & Definitions (84)