Logit-Q dynamics for efficient learning in stochastic teams
From MaRDI portal
Cites work
- \({\mathcal Q}\)-learning
- 10.1162/153244303765208377
- 10.1162/1532443041827880
- Achieving Pareto Optimality Through Distributed Learning
- Aspiration learning in coordination games
- Asynchronous stochastic approximation with differential inclusions
- Best-response dynamics in zero-sum stochastic games
- Comparison of perturbation bounds for the stationary distribution of a Markov chain
- Decentralized Learning for Optimality in Stochastic Dynamic Teams and Games With Local Control and Global State Information
- Dynamic programming and optimal control. Vol. 1.
- Fictitious play in zero-sum stochastic games
- Game of thrones: fully distributed learning for multiplayer bandits
- Game-Theoretic Learning and Distributed Optimization in Memoryless Multi-Agent Systems
- scientific article; zbMATH DE number 1405930 (Why is no real title available?)
- scientific article; zbMATH DE number 3205836 (Why is no real title available?)
- Individual learning in normal form games: Some laboratory results
- Learning efficient Nash equilibria in distributed systems
- Markov chains and mixing times. With a chapter on ``Coupling from the past by James G. Propp and David B. Wilson.
- Multi-agent reinforcement learning: a selective overview of theories and algorithms
- Perturbation bounds for the stationary distributions of Markov chains
- Revisiting log-linear learning: asynchrony, completeness and payoff-based implementation
- Stochastic Games
This page was built for publication: Logit-Q dynamics for efficient learning in stochastic teams
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6925770)