Learning in games by random sampling

From MaRDI portal





The authors study repeated interactions among a fixed set of ``low rationality players who have status quo actions, randomly sample other actions, and change their status quo if the sampled action yields a higher payoff. This behavior generates a random process, the better-reply dynamics. Long run behaviour leads to Nash equilibrium in games with the weak finite improvement property, including finite, supermodular games and generic, continuous, two-player, quasi-concave games. If the players make mistakes and if several players can sample at the same time, the resulting better-reply dynamics with simultaneous sampling converges to the Pareto optimal Nash equilibrium in common interest games.




Cited in
(34)








This page was built for publication: Learning in games by random sampling

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5938633)