Deep approximate policy iteration
From MaRDI portal
Cites work
- 10.1162/1532443041827907
- A Bernstein-type inequality for some mixing processes and dynamical systems with an application to learning
- A generalization error for Q-learning
- Adaptive approximation and generalization of deep neural network with intrinsic dimensionality
- An introduction to deep reinforcement learning
- Analysis of classification-based policy iteration algorithms
- Deep approximate policy iteration
- Deep network approximation characterized by number of neurons
- Deep Network Approximation for Smooth Functions
- Deep neural networks for estimation and inference
- Deep nonparametric regression on approximate manifolds: nonasymptotic error bounds with polynomial prefactors
- Empirical processes and random projections
- End-to-end training of deep visuomotor policies
- Equivalence of approximation by convolutional neural networks and fully-connected networks
- Error bounds for approximations with deep ReLU networks
- EXPONENTIAL INEQUALITIES AND FUNCTIONAL ESTIMATIONS FOR WEAK DEPENDENT DATA: APPLICATIONS TO DYNAMICAL SYSTEMS
- Finite-time bounds for fitted value iteration
- scientific article; zbMATH DE number 3148886 (Why is no real title available?)
- scientific article; zbMATH DE number 1944026 (Why is no real title available?)
- scientific article; zbMATH DE number 1753152 (Why is no real title available?)
- scientific article; zbMATH DE number 6542806 (Why is no real title available?)
- Learning from dependent observations
- Learning near-optimal policies with Bellman-residual minimization based fitted policy iteration and a single sample path
- Linear least-squares algorithms for temporal difference learning
- Neural Network Learning
- Nonlinear approximation via compositions
- Nonparametric regression using deep neural networks with ReLU activation function
- On deep learning as a remedy for the curse of dimensionality in nonparametric regression
- On the rate of convergence of image classifiers based on convolutional neural networks
- Optimal approximation of piecewise smooth functions using deep ReLU neural networks
- Probabilistic machine learning. An introduction
- Rates of convergence for empirical processes of stationary mixing sequences
- Regularized policy iteration with nonparametric function spaces
- Reinforcement learning. An introduction
- Stability bounds for stationary -mixing and -mixing processes
- Theory of deep convolutional neural networks: downsampling
- Universality of deep convolutional neural networks
- Whitney's extension problem for \(C^m\)
This page was built for publication: Deep approximate policy iteration
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6974372)