Hyperbolically Discounted Temporal Difference Learning
From MaRDI portal
Recommendations
- Differential Temporal Difference Learning
- On average versus discounted reward temporal-difference learning
- Temporal difference learning with incremental nearest neighbors in continuous spaces
- Q-learning and enhanced policy iteration in discounted dynamic programming
- TD-regularized actor-critic methods
- New Versions of Gradient Temporal-Difference Learning
- Multikernel recursive least-squares temporal difference learning
- Practical issues in temporal difference learning
Cites work
Cited in
(5)- Estimating scale-invariant future in continuous time
- New Versions of Gradient Temporal-Difference Learning
- Relative loss bounds for temporal-difference learning
- Immediate return preference emerged from a synaptic learning rule for return maximization
- Internal-Time Temporal Difference Model for Neural Value-Based Decision Making
This page was built for publication: Hyperbolically Discounted Temporal Difference Learning
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3568377)