Statistical complexity and optimal algorithms for nonlinear ridge bandits
From MaRDI portal
Cites work
- 10.1162/153244303321897663
- A class of measures of informativity of observation channels
- Asymptotic theory of sequential estimation: Differential geometrical approach
- Asymptotically efficient adaptive allocation rules
- Bandit algorithms
- Bandit problems with infinitely many arms
- Bypassing the Monster: A Faster and Simpler Optimal Algorithm for Contextual Bandits Under Realizability
- Explore first, exploit next: the true shape of regret in bandit problems
- Exponential lower bounds for planning in MDPs with linearly-realizable optimal action-value functions
- scientific article; zbMATH DE number 3638998 (Why is no real title available?)
- Improved regret for zeroth-order adversarial bandit convex optimisation
- Information-Based Complexity, Feedback and Dynamics in Convex Programming
- Kernel-based methods for bandit convex optimization
- Multi-armed bandit models for the optimal design of clinical trials: benefits and challenges
- Nonparametric and semiparametric models.
- On the Number of Iterations of Piyavskii's Global Optimization Algorithm
- Online convex optimization in the bandit setting: gradient descent without a gradient
- Optimal reconstruction of a function from its projections
- Optimum Character of the Sequential Probability Ratio Test
- Reinforcement learning. An introduction
- Riemannian metrics on convex sets with applications to Poincaré and log-Sobolev inequalities
- Sequential Design of Experiments
- SEQUENTIAL DISCRIMINATION OF HYPOTHESES WITH CONTROL OF OBSERVATIONS
- Sequential estimation of the largest normal mean when the variance is known
- Sequential Experimental Design Procedures
- Some aspects of the sequential design of experiments
- Statistical complexity and optimal algorithms for nonlinear ridge bandits
- The Capacity of Channels With Feedback
- The Nonstochastic Multiarmed Bandit Problem
Cited in
(2)
This page was built for publication: Statistical complexity and optimal algorithms for nonlinear ridge bandits
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q7035874)