Grid-mapping pseudo-count constraint for offline reinforcement learning
From MaRDI portal
Cites work
- \({\mathcal Q}\)-learning
- A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
- Regularisation of neural networks by enforcing Lipschitz continuity
- Reinforcement learning. An introduction
- The collected works of Wassily Hoeffding. Ed. by N. I. Fisher and P. K. Sen
This page was built for publication: Grid-mapping pseudo-count constraint for offline reinforcement learning
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6890012)