scientific article; zbMATH DE number 5957269
From MaRDI portal
Publication:3093261
Recommendations
- Reinforcement learning trees
- Improving reinforcement learning by using sequence trees
- Tree-based reinforcement learning for estimating optimal dynamic treatment regimes
- Cover tree Bayesian reinforcement learning
- Decision tree algorithm with reinforcement learning strategy
- Model-free reinforcement learning for branching Markov decision processes
- Scalable and efficient Bayes-adaptive reinforcement learning based on Monte-Carlo tree search
Cited in
(60)- Learning near-optimal policies with Bellman-residual minimization based fitted policy iteration and a single sample path
- Learning output reference model tracking for higher-order nonlinear systems with unknown dynamics
- A deep reinforcement learning framework for continuous intraday market bidding
- Challenges of real-world reinforcement learning: definitions, benchmarks and analysis
- Multi-agent reinforcement learning: a selective overview of theories and algorithms
- Batch policy learning in average reward Markov decision processes
- Data-driven switching modeling for MPC using regression trees and random forests
- Recovery of simultaneous low rank and two-way sparse coefficient matrices, a nonconvex approach
- Fitted Q-iteration by functional networks for control problems
- Scalable transfer learning in heterogeneous, dynamic environments
- A unified DC programming framework and efficient DCA based approaches for large scale batch reinforcement learning
- Efficient approximate dynamic programming based on design and analysis of computer experiments for infinite-horizon optimization
- Estimating optimal shared-parameter dynamic regimens with application to a multistage depression clinical trial
- Reinforcement learning strategies for clinical trials in nonsmall cell lung cancer
- Cover tree Bayesian reinforcement learning
- Towards min max generalization in reinforcement learning
- Extreme state aggregation beyond Markov decision processes
- Making friends on the fly: cooperating with new teammates
- Bounds for Multistage Stochastic Programs Using Supervised Learning Strategies
- Batch mode reinforcement learning based on the synthesis of artificial trajectories
- Model selection in reinforcement learning
- Abstraction from demonstration for efficient reinforcement learning in high-dimensional domains
- scientific article; zbMATH DE number 7626792 (Why is no real title available?)
- Bandit Theory: Applications to Learning Healthcare Systems and Clinical Trials
- The QLBS Q-Learner goes NuQLear: fitted Q iteration, inverse RL, and option portfolios
- Reinforcement learning trees
- Epoch-incremental reinforcement learning algorithms
- Quadratic approximate dynamic programming for input-affine systems
- Machine Learning: ECML 2004
- Hessian matrix distribution for Bayesian policy gradient reinforcement learning
- Learning when-to-treat policies
- Iteratively extending time horizon reinforcement learning.
- Extremely randomized trees
- Extremely randomized trees
- Optimized ensemble value function approximation for dynamic programming
- Tutorial on Amortized Optimization
- Target Network and Truncation Overcome the Deadly Triad in \(\boldsymbol{Q}\)-Learning
- Approximated multi-agent fitted Q iteration
- Evolving interpretable decision trees for reinforcement learning
- Exploiting action impact regularity and exogenous state variables for offline reinforcement learning
- On sparse representation for optimal individualized treatment selection with penalized outcome weighted learning
- Deep spectral Q-learning with application to mobile health
- A Q-learning algorithm for Markov decision processes with continuous state spaces
- Minimax weight learning for absorbing MDPs
- Reinforcement learning
- Value Enhancement of Reinforcement Learning via Efficient and Robust Trust Region Optimization
- Super-learning of an optimal dynamic treatment rule
- Deep spatial Q-learning for infectious disease control
- Generalized fitted Q-iteration with clustered data
- Offline reinforcement learning in large state spaces: algorithms and guarantees
- On the statistical complexity for offline and low-adaptive reinforcement learning with structures
- Optimizing return distributions with distributional dynamic programming
- Recent advances in causal machine learning and dynamic policy learning
- An L^2 analysis of reinforcement learning in high dimensions with kernel and neural network approximation
- Testing stationarity and change point detection in reinforcement learning
- Deep controlled learning for inventory control
- Stage-aware learning for dynamic treatments
- Regularized feature selection in reinforcement learning
- Reinforcement learning algorithms with function approximation: recent advances and applications
- Approximate dynamic programming with a fuzzy parameterization
This page was built for publication:
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3093261)