Distributed Reinforcement Learning via Gossip
From MaRDI portal
Abstract: We consider the classical TD(0) algorithm implemented on a network of agents wherein the agents also incorporate the updates received from neighboring agents using a gossip-like mechanism. The combined scheme is shown to converge for both discounted and average cost problems.
Cited in
(7)- Distributed consensus-based multi-agent temporal-difference learning
- Decentralized fused-learner architectures for Bayesian reinforcement learning
- Distributed Deep Learning on Heterogeneous Computing Resources Using Gossip Communication
- Scalable Reinforcement Learning for Multiagent Networked Systems
- Distributed L2-gain control of large-scale systems under gossip communication protocol
- Model-free adaptive control design for nonlinear discrete-time processes with reinforcement learning techniques
- Finite-time performance of distributed temporal-difference learning with linear function approximation
This page was built for publication: Distributed Reinforcement Learning via Gossip
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5282396)