Distributed policy evaluation via inexact ADMM in multi-agent reinforcement learning
From MaRDI portal
Recommendations
- Fully asynchronous policy evaluation in distributed reinforcement learning over networks
- Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis
- Distributed multi-agent temporal-difference learning with full neighbor information
- Multi-agent reinforcement learning using ordinal action selection and approximate policy iteration
- Cooperative multi-agent reinforcement learning with constraint-reduced DCOP
Cited in
(19)- Normalizing flow policies for multi-agent systems
- Fully asynchronous policy evaluation in distributed reinforcement learning over networks
- Adaptive output regulation for cyber-physical systems under time-delay attacks
- A solution strategy for distributed uncertain economic dispatch problems via scenario theory
- Distributed multi-agent temporal-difference learning with full neighbor information
- Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis
- Scalable Reinforcement Learning for Multiagent Networked Systems
- Byzantine-Resilient Decentralized Policy Evaluation With Linear Function Approximation
- Why the `selfish' optimizing agents could solve the decentralized reinforcement learning problems
- Distributed regularized online optimization using forward-backward splitting
- Multi-agent off-policy actor-critic algorithm for distributed multi-task reinforcement learning
- Distributed reinforcement learning for coordinate multi-robot foraging
- Entropy regularized actor-critic based multi-agent deep reinforcement learning for stochastic games
- Approximated multi-agent fitted Q iteration
- Distributed entropy-regularized multi-agent reinforcement learning with policy consensus
- Multi-agent reinforcement learning via distributed MPC as a function approximator
- Achieving collective welfare in multi-agent reinforcement learning via suggestion sharing
- Distributed constrained optimization over unbalanced graphs and delayed gradient
- Linear convergence of event-triggered distributed optimization with metric subregularity condition
This page was built for publication: Distributed policy evaluation via inexact ADMM in multi-agent reinforcement learning
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4995742)