Distributed consensus-based multi-agent temporal-difference learning
From MaRDI portal
Recommendations
- Distributed multi-agent temporal-difference learning with full neighbor information
- scientific article; zbMATH DE number 1974064
- Multi-agent off-policy actor-critic algorithm for distributed multi-task reinforcement learning
- Distributed learning and cooperative control for multi-agent systems
- Consensus-based iterative learning of heterogeneous agents with application to distributed optimization
- Consensus of discrete-time multi-agent system based on q-learning
- ${{\cal Q} {\cal D}}$-Learning: A Collaborative Distributed Strategy for Multi-Agent Reinforcement Learning Through ${\rm Consensus} + {\rm Innovations}$
- Multiagent Fully Decentralized Value Function Learning With Linear Convergence Rates
- The dynamics of multi-agent reinforcement learning
- Distributed Reinforcement Learning via Gossip
Cites work
- scientific article; zbMATH DE number 1600999 (Why is no real title available?)
- scientific article; zbMATH DE number 1321699 (Why is no real title available?)
- scientific article; zbMATH DE number 1972910 (Why is no real title available?)
- scientific article; zbMATH DE number 3215568 (Why is no real title available?)
- ${{\cal Q} {\cal D}}$-Learning: A Collaborative Distributed Strategy for Multi-Agent Reinforcement Learning Through ${\rm Consensus} + {\rm Innovations}$
- An emphatic approach to the problem of off-policy temporal-difference learning
- Application of reinforcement learning to wireless sensor networks: models and algorithms
- Asymptotic Properties of Distributed and Communicating Stochastic Approximation Algorithms
- Asynchronous Distributed Blind Calibration of Sensor Networks Under Noisy Measurements
- Consensus based overlapping decentralized estimation with missing observations and communication faults
- Consensus-based decentralized real-time identification of large-scale systems
- Distributed Policy Evaluation Under Multiple Behavior Strategies
- Distributed Reinforcement Learning via Gossip
- Distributed Stochastic Approximation: Weak Convergence and Network Design
- Distributed asynchronous deterministic and stochastic gradient optimization algorithms
- Distributed model based event-triggered control for synchronization of multi-agent systems
- Distributed time synchronization for networks with random delays and measurement noise
- Off-policy learning with eligibility traces: a survey
- Optimal dynamic formation control of multi-agent systems in constrained environments
- Weak convergence properties of constrained emphatic temporal-difference learning with constant and slowly diminishing stepsize
Cited in
(11)- Multi-agent natural actor-critic reinforcement learning algorithms
- Multiagent Fully Decentralized Value Function Learning With Linear Convergence Rates
- Distributed multi-agent temporal-difference learning with full neighbor information
- Learning to agree over large state spaces
- Finite-time convergence rates of distributed local stochastic approximation
- Multi-agent off-policy actor-critic algorithm for distributed multi-task reinforcement learning
- Decentralized concurrent learning with coordinated momentum and restart
- Adaptive distributed tracking control for Markov jump multiagent systems with a non-strict leader
- Finite-time error bounds for distributed linear stochastic approximation
- Finite-time performance of distributed temporal-difference learning with linear function approximation
- Distributed entropy-regularized multi-agent reinforcement learning with policy consensus
This page was built for publication: Distributed consensus-based multi-agent temporal-difference learning
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6164031)