Algorithms for reinforcement learning.
From MaRDI portal
active learningactor-critic methodsbias-variance tradeofffunction approximationleast-squares methodsMarkov decision processesnatural gradientonline learningoverfittingPAC-learningplanningpolicy gradientQ-learningreinforcement learningsimulationstochastic approximationstochastic gradient methodstemporal difference learning
Recommendations
Cited in
(65)- Reinforcement learning agents
- Continuous-action planning for discounted infinite-horizon nonlinear optimal control with Lipschitz values
- Online spatio-temporal matching in stochastic and dynamic domains
- A unified framework for stochastic optimization
- Markov decision processes with sequential sensor measurements
- Crowd computing as a cooperation problem: An evolutionary approach
- Fundamental design principles for reinforcement learning algorithms
- Decision making under uncertainty and reinforcement learning. Theory and algorithms
- On learning and branching: a survey
- Preference-based reinforcement learning: evolutionary direct policy search using a preference-based racing algorithm
- A unified DC programming framework and efficient DCA based approaches for large scale batch reinforcement learning
- A systematic study on meta-heuristic approaches for solving the graph coloring problem
- Reinforcement learning theory, algorithms and its application
- Editorial: some recent advances in learning and adaptation for uncertain feedback control systems
- Adaptive playouts for online learning of policies during Monte Carlo tree search
- Efficient model-based reinforcement learning for approximate online optimal control
- TEXPLORE: temporal difference reinforcement learning for robots and time-constrained domains
- Hypervolume indicator and dominance reward based multi-objective Monte-Carlo tree search
- Robust adaptive dynamic programming for linear and nonlinear systems: an overview
- Minimax PAC bounds on the sample complexity of reinforcement learning with a generative model
- Dynamic treatment regimes: technical challenges and applications
- Model selection in reinforcement learning
- scientific article; zbMATH DE number 1950579 (Why is no real title available?)
- Asymptotic analysis of value prediction by well-specified and misspecified models
- scientific article; zbMATH DE number 6982305 (Why is no real title available?)
- Abstraction from demonstration for efficient reinforcement learning in high-dimensional domains
- Reinforcement learning. An introduction
- Efficient augmentation and relaxation learning for individualized treatment rules using observational data
- Non-parametric policy search with limited information loss
- scientific article; zbMATH DE number 836011 (Why is no real title available?)
- Bayesian exploration for approximate dynamic programming
- Statistical reinforcement learning. Modern machine learning approaches
- Finite-time performance of distributed temporal-difference learning with linear function approximation
- scientific article; zbMATH DE number 7626721 (Why is no real title available?)
- Closed-form Approximations in Multi-asset Market Making
- Computational Benefits of Intermediate Rewards for Goal-Reaching Policy Learning
- A Reinforcement Learning Neural Network for Robotic Manipulator Control
- Deep exploration via randomized value functions
- On convergence of value iteration for a class of total cost Markov decision processes
- Undiscounted reinforcement learning algorithm based on performance potentials
- Empirical Q-value iteration
- Least squares policy iteration with instrumental variables vs. direct policy search: comparison against optimal benchmarks using energy storage
- A Two-Timescale Stochastic Algorithm Framework for Bilevel Optimization: Complexity Analysis and Application to Actor-Critic
- Optimal activation of halting multi‐armed bandit models
- Formalization of methods for the development of autonomous artificial intelligence systems
- Deep reinforcement trading with predictable returns
- Approximate Q Learning for Controlled Diffusion Processes and Its Near Optimality
- Modern Bayesian experimental design
- Investigating the properties of neural network representations in reinforcement learning
- Convergence of entropy-regularized natural policy gradient with linear function approximation
- Structure in machine learning
- UVIP: model-free approach to evaluate reinforcement learning algorithms
- Optimal sub-Gaussian variance proxy for truncated Gaussian and exponential random variables
- A characterization method of terminal ingredients for nonlinear MPC using value-based reinforcement learning
- Continual learning as computationally constrained reinforcement learning
- An L^2 analysis of reinforcement learning in high dimensions with kernel and neural network approximation
- Interpolating between BSDEs and PINNs: deep learning for elliptic and parabolic boundary value problems
- Prospect utility with hyperbolic tangent function
- Learning algorithms for verification of Markov decision processes
- Reinforcement learning via nonparametric smoothing in a continuous-time stochastic setting with noisy data
- Near-continuous time reinforcement learning for continuous state-action spaces
- Proximal algorithms and temporal difference methods for solving fixed point problems
- A convex optimization approach to dynamic programming in continuous state and action spaces
- Reinforcement learning algorithms with function approximation: recent advances and applications
- Adaptive representations for reinforcement learning.
This page was built for publication: Algorithms for reinforcement learning.
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3588852)