Search or split: policy gradient with adaptive policy space
From MaRDI portal
Cites work
- Curriculum learning for reinforcement learning domains: a framework and survey
- Global optimality guarantees for policy gradient methods
- scientific article; zbMATH DE number 1375577 (Why is no real title available?)
- scientific article; zbMATH DE number 1753152 (Why is no real title available?)
- Overlapping layered learning
- Simple statistical gradient-following algorithms for connectionist reinforcement learning
- Smoothing policies and safe policy gradients
This page was built for publication: Search or split: policy gradient with adaptive policy space
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6903643)