Generalized fitted Q-iteration with clustered data
From MaRDI portal
Cites work
- \({\mathcal Q}\)-learning
- A multiagent reinforcement learning framework for off-policy evaluation in two-sided markets
- Basic properties of strong mixing conditions. A survey and some open questions
- Batch policy learning in average reward Markov decision processes
- Bias from the use of generalized estimating equations to analyze incomplete longitudinal binary data
- Constructing dynamic treatment regimes over indefinite time horizons
- Distributional Off-Policy Evaluation in Reinforcement Learning
- Dynamic Causal Effects Evaluation in A/B Testing with a Reinforcement Learning Framework
- Efficient evaluation of natural stochastic policies in off-line reinforcement learning
- Estimating dynamic treatment regimes in mobile health using V-learning
- Estimating Optimal Infinite Horizon Dynamic Treatment Regimes via pT-Learning
- Finite-time bounds for fitted value iteration
- Generalized TD learning
- scientific article; zbMATH DE number 5957269 (Why is no real title available?)
- Inference and missing data
- Longitudinal Data Analysis
- Longitudinal Data Analysis
- Longitudinal data analysis using generalized linear models
- Model selection for generalized estimating equations accommodating dropout missingness
- Model selection of generalized estimating equations with multiply imputed longitudinal data
- Off-Policy Confidence Interval Estimation with Confounded Markov Decision Process
- Off-policy estimation of long-term average outcomes with applications to mobile health
- Off-Policy Evaluation in Doubly Inhomogeneous Environments
- Off-policy evaluation in partially observed Markov decision processes under sequential ignorability
- Online Bootstrap Inference For Policy Evaluation In Reinforcement Learning
- Optimal policy evaluation using kernel-based temporal difference methods
- Optimal Treatment Regimes: A Review and Empirical Comparison
- Partially observed Markov decision processes. From filtering to controlled sensing
- Penalized generalized estimating equations for high-dimensional longitudinal data analysis
- Policy evaluation for temporal and/or spatial dependent experiments
- Projected state-action balancing weights for offline reinforcement learning
- Randomization inference when N equals one
- Reinforcement learning for individual optimal policy from heterogeneous data
- Reinforcement learning. An introduction
- Robustness of generalized estimating equation (GEE) tests of significance against misspecification of the error structure model
- Settling the sample complexity of model-based offline reinforcement learning
- Statistical analysis with missing data
- Statistical inference of the value function for reinforcement learning in infinite-horizon settings
- Statistical methods for dynamic treatment regimes. Reinforcement learning, causal inference, and personalized medicine
- Statistically Efficient Advantage Learning for Offline Reinforcement Learning in Infinite Horizons
- Testing stationarity and change point detection in reinforcement learning
- Toward theoretical understandings of robust Markov decision processes: sample complexity and asymptotics
- Value Enhancement of Reinforcement Learning via Efficient and Robust Trust Region Optimization
This page was built for publication: Generalized fitted Q-iteration with clustered data
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6852119)