Bayesian Reinforcement Learning with Exploration
From MaRDI portal
Recommendations
- Bayesian exploration for approximate dynamic programming
- Bayesian reinforcement learning: a survey
- Bayesian Incentive-Compatible Bandit Exploration
- Bayesian policy gradient and actor-critic algorithms
- scientific article; zbMATH DE number 795283
- Bayesian optimistic Kullback-Leibler exploration
- Bayesian exploration: incentivizing exploration in Bayesian games
Cited in
(29)- Cover tree Bayesian reinforcement learning
- A Bayesian reinforcement learning approach in Markov games for computing near-optimal policies
- A model for system uncertainty in reinforcement learning
- Finding the optimal exploration-exploitation trade-off online through Bayesian risk estimation and minimization
- Model selection in reinforcement learning
- Scalable and efficient Bayes-adaptive reinforcement learning based on Monte-Carlo tree search
- Using trajectory data to improve Bayesian optimization for reinforcement learning
- Uncertainty Propagation for Efficient Exploration in Reinforcement Learning
- Towards min max generalization in reinforcement learning
- Rationality, optimism and guarantees in general reinforcement learning
- Exploration and incentives in reinforcement learning
- Reinforcement learning: a comparison of UCB versus alternative adaptive policies
- A Monte-Carlo AIXI approximation
- Uncertainty quantification and exploration for reinforcement learning
- Dual control for approximate Bayesian reinforcement learning
- Reducing reinforcement learning to KWIK online regression
- Reinforcement learning with immediate rewards and linear hypotheses
- Bayesian reinforcement learning: a survey
- Exploration of multi-state environments: Local measures and back-propagation of uncertainty
- Near-optimal PAC bounds for discounted MDPs
- scientific article; zbMATH DE number 2063052 (Why is no real title available?)
- On the possibility of learning in reactive environments with arbitrary dependence
- Self-Optimizing and Pareto-Optimal Policies in General Environments based on Bayes-Mixtures
- Bayesian policy gradient and actor-critic algorithms
- Bayesian policy reuse
- Bayesian Incentive-Compatible Bandit Exploration
- Asymptotic Learnability of Reinforcement Problems with Arbitrary Dependence
- Deep exploration via randomized value functions
- Bayesian optimistic Kullback-Leibler exploration
This page was built for publication: Bayesian Reinforcement Learning with Exploration
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2938731)