Q-learning for estimating optimal dynamic treatment rules from observational data
From MaRDI portal
Recommendations
- Q- and A-learning methods for estimating optimal dynamic treatment regimes
- New statistical learning methods for estimating optimal dynamic treatment regimes
- Robust estimation of optimal dynamic treatment regimes for sequential treatment decisions
- A smoothed Q‐learning algorithm for estimating optimal dynamic treatment regimes
- Robust Q-learning
Cites work
- A generalization error for Q-learning
- Causal effect models for realistic individualized treatment and intention to treat rules
- Estimating optimal dynamic regimes correcting bias under the null
- Inference for non-regular parameters in optimal dynamic treatment regimes
- Inference for optimal dynamic treatment regimes using an adaptive m-out-of-n bootstrap scheme
- Optimal Dynamic Treatment Regimes
- Optimal Structural Nested Models for Optimal Sequential Decisions
- Regret-Regression for Optimal Dynamic Treatment Regimes
- Reinforcement learning strategies for clinical trials in nonsmall cell lung cancer
- The central role of the propensity score in observational studies for causal effects
Cited in
(30)- Tree-based reinforcement learning for estimating optimal dynamic treatment regimes
- Simulation-based optimization of radiotherapy: agent-based modeling and reinforcement learning
- Q- and A-learning methods for estimating optimal dynamic treatment regimes
- Using decision lists to construct interpretable and parsimonious treatment regimes
- Interactive model building for Q-learning
- Penalized Q-learning for dynamic treatment regimens
- Incorporating patient preferences into estimation of optimal individualized treatment rules
- Efficient augmentation and relaxation learning for individualized treatment rules using observational data
- Causal Rule Sets for Identifying Subgroups with Enhanced Treatment Effects
- Estimation and optimization of composite outcomes
- Proper inference for value function in high-dimensional Q-learning for dynamic treatment regimes
- Adaptive contrast weighted learning for multi-stage multi-treatment decision-making
- A smoothed Q‐learning algorithm for estimating optimal dynamic treatment regimes
- Stochastic tree search for estimating optimal dynamic treatment regimes
- Selecting and ranking individualized treatment rules with unmeasured confounding
- Robust Q-learning
- Deep advantage learning for optimal dynamic treatment regime
- Dynamic treatment regimes with interference
- Synthetic learner: model-free inference on treatments over time
- Evaluating the use of generalized dynamic weighted ordinary least squares for individualized HIV treatment strategies
- Deep reinforcement learning for personalized treatment recommendation
- The optimal dynamic treatment rule superlearner: considerations, performance, and application to criminal justice interventions
- A sequential, multiple assignment, randomized trial design with a tailoring function
- Asymptotic inference for multi-stage stationary treatment policy with variable selection
- When the ends do not justify the means: learning who is predicted to have harmful indirect effects
- Estimation of optimal dynamic treatment assignment rules under policy constraints
- Identifying optimal dosage regimes under safety constraints: an application to long term opioid treatment of chronic pain
- Dynamic treatment regimes on dyadic networks
- Q-learning via deep learning-based Buckley-James method for non-linear censored data
- Counterfactual Q-learning via the linear Buckley-James method for longitudinal survival data
This page was built for publication: Q-learning for estimating optimal dynamic treatment rules from observational data
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2856564)