A human-centered data-driven planner-actor-critic architecture via logic programming
From MaRDI portal
Recommendations
Cites work
- A synthesis of automated planning and reinforcement learning for efficient, robust decision-making
- Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning
- Generality in artificial intelligence
- Mobile robot planning using action language \({\mathcal {BC}}\) with an abstraction hierarchy
- Natural actor-critic algorithms
- Planning and acting in partially observable stochastic domains
- Recent advances in hierarchical reinforcement learning
- Reinforcement learning. An introduction
- Robot task planning and explanation in open and uncertain worlds
- Simple statistical gradient-following algorithms for connectionist reinforcement learning
- Some properties of system descriptions of \(\mathcal{AL}_d\)
- The fast downward planning system
This page was built for publication: A human-centered data-driven planner-actor-critic architecture via logic programming
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5020557)