Reward is enough
From MaRDI portal
Recommendations
Cites work
- A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
- Computer Go
- Equilibrium points in n -person games
- Existence and Uniqueness of Equilibrium Points for Concave N-Person Games
- scientific article; zbMATH DE number 1149416 (Why is no real title available?)
- scientific article; zbMATH DE number 2161894 (Why is no real title available?)
- scientific article; zbMATH DE number 783783 (Why is no real title available?)
- scientific article; zbMATH DE number 1408945 (Why is no real title available?)
- scientific article; zbMATH DE number 3092990 (Why is no real title available?)
- If multi-agent learning is the answer, what is the question?
- Reinforcement learning. An introduction
- Transfer learning for reinforcement learning domains: a survey
- Universal artificial intelligence. Sequential decisions based on algorithmic probability.
Cited in
(8)- Rewarding effort
- Finding intrinsic rewards by embodied evolution and constrained reinforcement learning
- On our best behaviour
- A deep multi-agent reinforcement learning approach to solve dynamic job shop scheduling problem
- Formalization of methods for the development of autonomous artificial intelligence systems
- Reward tampering problems and solutions in reinforcement learning: a causal influence diagram perspective
- Multi-target tracking based on two-layer reinforcement learning optimization
- Data-Driven Knowledge Transfer in Batch Q * Learning
This page was built for publication: Reward is enough
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2238710)