Bayesian Reinforcement Learning with Exploration
From MaRDI portal
Recommendations
- Bayesian exploration for approximate dynamic programming
- Bayesian reinforcement learning: a survey
- Bayesian Incentive-Compatible Bandit Exploration
- Bayesian policy gradient and actor-critic algorithms
- scientific article; zbMATH DE number 795283
- Bayesian optimistic Kullback-Leibler exploration
- Bayesian exploration: incentivizing exploration in Bayesian games
Cited in
(32)- k-Certainty Exploration Method: an action selector to identify the environment in reinforcement learning
- Bayesian policy reuse
- A model for system uncertainty in reinforcement learning
- Reinforcement learning with immediate rewards and linear hypotheses
- Exploration of multi-state environments: Local measures and back-propagation of uncertainty
- Bayesian optimistic Kullback-Leibler exploration
- Bayesian reinforcement learning: a survey
- Bayesian policy gradient and actor-critic algorithms
- Dual control for approximate Bayesian reinforcement learning
- Scalable and efficient Bayes-adaptive reinforcement learning based on Monte-Carlo tree search
- Using trajectory data to improve Bayesian optimization for reinforcement learning
- Cover tree Bayesian reinforcement learning
- Uncertainty Propagation for Efficient Exploration in Reinforcement Learning
- Towards min max generalization in reinforcement learning
- Self-Optimizing and Pareto-Optimal Policies in General Environments based on Bayes-Mixtures
- A Monte-Carlo AIXI approximation
- Reinforcement learning: a comparison of UCB versus alternative adaptive policies
- Bayesian Incentive-Compatible Bandit Exploration
- Asymptotic Learnability of Reinforcement Problems with Arbitrary Dependence
- Model selection in reinforcement learning
- scientific article; zbMATH DE number 2063052 (Why is no real title available?)
- Near-optimal PAC bounds for discounted MDPs
- scientific article; zbMATH DE number 1453042 (Why is no real title available?)
- Robust reinforcement learning with Bayesian optimisation and quadrature
- Deep exploration via randomized value functions
- Rationality, optimism and guarantees in general reinforcement learning
- A Bayesian reinforcement learning approach in Markov games for computing near-optimal policies
- Reducing reinforcement learning to KWIK online regression
- Finding the optimal exploration-exploitation trade-off online through Bayesian risk estimation and minimization
- Exploration and incentives in reinforcement learning
- Uncertainty quantification and exploration for reinforcement learning
- On the possibility of learning in reactive environments with arbitrary dependence
This page was built for publication: Bayesian Reinforcement Learning with Exploration
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2938731)