Policy optimization for reinforcement learning in continuous time and space
From MaRDI portal
Cites work
- \({\mathcal Q}\)-learning
- Accuracy and Stability of Numerical Algorithms
- Comparison between \(W_2\) distance and \(\dot{H}^{-1}\) norm, and localization of Wasserstein distance
- Continuous‐time mean–variance portfolio selection: A reinforcement learning framework
- Controlled Markov processes and viscosity solutions
- Deep hedging
- Exploratory HJB equations and their convergence
- Global convergence of policy gradient methods to (almost) locally optimal policies
- Hitting, occupation and inverse local times of one-dimensional diffusions: Martingale and excursion approaches
- scientific article; zbMATH DE number 51724 (Why is no real title available?)
- scientific article; zbMATH DE number 1325009 (Why is no real title available?)
- scientific article; zbMATH DE number 7626721 (Why is no real title available?)
- scientific article; zbMATH DE number 7307478 (Why is no real title available?)
- Mimicking an Itō process by a solution of a stochastic differential equation
- MM optimization algorithms
- New perturbation analyses for the Cholesky factorization
- On the perturbation of LU and Cholesky factors
- On the Perturbation of LU, Cholesky, and QR Factorizations
- Policy gradient in continuous time
- Policy iteration for exploratory Hamilton-Jacobi-Bellman equations
- Policy iterations for reinforcement learning problems in continuous time and space -- fundamental theory and methods
- Probability
- Q-learning for continuous-time linear systems: A model-free infinite horizon optimal control approach
- Queueing network controls via deep reinforcement learning
- Reinforcement learning. An introduction
- Stochastic differential equations. An introduction with applications.
- Uniqueness of the solution to the Vlasov--Poisson system with bounded density
This page was built for publication: Policy optimization for reinforcement learning in continuous time and space
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q7293483)