Convergence rate comparison of two data-driven algorithms to stochastic LQR problems
From MaRDI portal
Cites work
- ${\cal H}$-Representation and Applications to Generalized Lyapunov Equations and Linear Stochastic Systems
- \(\mathrm{H}_\infty\) control of linear discrete-time systems: off-policy reinforcement learning
- A data-driven α -policy iteration algorithm for optimal leader-following consensus of discrete-time multi-agent systems
- Adaptive optimal control for continuous-time linear systems based on policy iteration
- Distributed Reinforcement Learning for Decentralized Linear Quadratic Control: A Derivative-Free Policy Optimization Approach
- scientific article; zbMATH DE number 53439 (Why is no real title available?)
- scientific article; zbMATH DE number 1321699 (Why is no real title available?)
- scientific article; zbMATH DE number 1095138 (Why is no real title available?)
- scientific article; zbMATH DE number 6931762 (Why is no real title available?)
- scientific article; zbMATH DE number 6159604 (Why is no real title available?)
- scientific article; zbMATH DE number 3225500 (Why is no real title available?)
- Inverse linear-quadratic discrete-time finite-horizon optimal control for indistinguishable homogeneous agents: a convex optimization approach
- Iterative matrix bounds and computational solutions to the discrete algebraic Riccati equation
- LQR controller design for affine LPV systems using reinforcement learning
- Model-free H_ control of Itô stochastic system via off-policy reinforcement learning
- Model-free Q-learning designs for linear discrete-time zero-sum games with application to H^ control
- Model-free optimal control of discrete-time systems with additive and multiplicative noises
- Neural network approach to continuous-time direct adaptive optimal control for partially unknown nonlinear systems
- On a Matrix Riccati Equation of Stochastic Control
- Optimal control
- Optimal control of discrete-time switched linear systems
- Optimal output tracking control of linear discrete-time systems with unknown dynamics by adaptive dynamic programming and output feedback
- Pareto optimal strategy for linear stochastic systems with \(H_\infty\) constraint in finite horizon
- Policy iteration reinforcement learning method for continuous-time linear-quadratic mean-field control problems
- Solution to Delayed Forward and Backward Stochastic Difference Equations and Its Applications
- Solution to stochastic LQ control problem for Itô systems with state delay or input delay
- Stability Analysis of Optimal Adaptive Control Using Value Iteration With Approximation Errors
- Stochastic H₂/H_ control: a Nash game approach
- Stochastic Linear Quadratic Regulators with Indefinite Control Weight Costs
- Value iteration and adaptive dynamic programming for data-driven adaptive optimal control design
This page was built for publication: Convergence rate comparison of two data-driven algorithms to stochastic LQR problems
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q7315738)