Concentration of Contractive Stochastic Approximation and Reinforcement Learning
From MaRDI portal
Recommendations
- A concentration bound for contractive stochastic approximation
- On the convergence of reinforcement learning
- Concentration bounds for stochastic approximations
- Concentration bounds for temporal difference learning with linear function approximation: the case of batch data and uniform sampling
- Machine Learning: ECML 2004
- On convergence rates of game theoretic reinforcement learning algorithms
- Regret bounds for reinforcement learning via Markov chain concentration
- The O.D.E. Method for Convergence of Stochastic Approximation and Reinforcement Learning
Cites work
- scientific article; zbMATH DE number 5957196 (Why is no real title available?)
- scientific article; zbMATH DE number 19232 (Why is no real title available?)
- scientific article; zbMATH DE number 49674 (Why is no real title available?)
- scientific article; zbMATH DE number 51132 (Why is no real title available?)
- A concentration bound for contractive stochastic approximation
- A concentration bound for stochastic approximation via Alekseev's formula
- An Invariant Measure Approach to the Convergence of Stochastic Approximations with State Dependent Noise
- An analysis of temporal-difference learning with function approximation
- Approximate Dynamic Programming
- Comparison of perturbation bounds for the stationary distribution of a Markov chain
- Concentration bounds for temporal difference learning with linear function approximation: the case of batch data and uniform sampling
- Convergence results for single-step on-policy reinforcement-learning algorithms
- Exact formula for sensitivity analysis of Markov chains
- Exponential inequalities for martingales and asymptotic properties of the free energy of directed polymers in a random environment
- Markov Chains and Stochastic Stability
- On the Lock-in Probability of Stochastic Approximation
- On the convergence, lock-in probability, and sample complexity of stochastic approximation
- Simplified description of slow Markov walks. II
- Simulation-based optimization of Markov reward processes
- Stochastic Approximation for Nonexpansive Maps: Application to Q-Learning Algorithms
- \({\mathcal Q}\)-learning
Cited in
(5)- Concentration of contractive stochastic approximation: additive and multiplicative noise
- The O.D.E. Method for Convergence of Stochastic Approximation and Reinforcement Learning
- Revisiting stochastic gradient descent for strongly convex objectives: tight uniform-in-time bounds
- Stochastic approximation and reinforcement learning: the interface and a little beyond
- Reinforcement learning in non-Markovian environments
This page was built for publication: Concentration of Contractive Stochastic Approximation and Reinforcement Learning
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5870773)