scientific article; zbMATH DE number 5348356
From MaRDI portal
Publication:3527701
Research exposition (monographs, survey articles) pertaining to dynamical systems and ergodic theory (37-02) Applications of dynamical systems (37N99) Functional limit theorems; invariance principles (60F17) Research exposition (monographs, survey articles) pertaining to statistics (62-02) Stochastic approximation (62L20) Applications of statistics in engineering and industry; control charts (62P30)
Recommendations
- Stochastic approximation. A dynamical systems viewpoint.
- Stochastic approximation. A dynamical systems viewpoint
- A Dynamical System Approach to Stochastic Approximations
- scientific article; zbMATH DE number 2052583
- Approximation of random dynamical systems with discrete time by stochastic differential equations: I. Theory
- scientific article; zbMATH DE number 1302168
- On the Investigation of Stochastic Dynamic Systems by Successive Approximation
- scientific article; zbMATH DE number 1944272
- A dynamical approximation for stochastic partial differential equations
Cited in
(only showing first 100 items - show all)- Robust adaptive Metropolis algorithm with coerced acceptance rate
- Reinforcement learning, sequential Monte Carlo and the EM algorithm
- An incremental off-policy search in a model-free Markov decision process using a single sample path
- An online prediction algorithm for reinforcement learning with linear function approximation using cross entropy method
- Convergence and efficiency of adaptive importance sampling techniques with partial biasing
- A convergence analysis of the perturbed compositional gradient flow: averaging principle and normal deviations
- An adaptive learning model with foregone payoff information
- Asymptotic behavior of truncated stochastic approximation procedures
- Approachability in Stackelberg stochastic games with vector costs
- Negatively reinforced balanced urn schemes
- Q-learning for Markov decision processes with a satisfiability criterion
- Simulation optimization of risk measures with adaptive risk levels
- Error bounds for constant step-size \(Q\)-learning
- Convergence and convergence rate of stochastic gradient search in the case of multiple and non-isolated extrema
- Stochastic approximation on Riemannian manifolds
- Conservative set valued fields, automatic differentiation, stochastic gradient methods and deep learning
- Inexact stochastic subgradient projection method for stochastic equilibrium problems with nonmonotone bifunctions: application to expected risk minimization in machine learning
- Incremental without replacement sampling in nonconvex optimization
- Fast incremental expectation maximization for finite-sum optimization: nonasymptotic convergence
- Financial replicator dynamics: emergence of systemic-risk-averting strategies
- Reference points and learning
- Unified reinforcement Q-learning for mean field game and control problems
- Pseudo-perturbation-based broadcast control of multi-agent systems
- Weak convergence of dynamical systems in two timescales
- A concentration bound for contractive stochastic approximation
- Prospect-theoretic Q-learning
- Simultaneous perturbation Newton algorithms for simulation optimization
- Optimal stochastic extragradient schemes for pseudomonotone stochastic variational inequality problems and their variants
- Stochastic subgradient method converges on tame functions
- Convergence results on stochastic adaptive learning
- Non-asymptotic error bounds for constant stepsize stochastic approximation for tracking mobile agents
- Incremental constraint projection methods for variational inequalities
- Stochastic approximation to understand simple simulation models
- A stability criterion for two timescale stochastic approximation schemes
- Markovian stochastic approximation with expanding projections
- Broadcast control of multi-agent systems
- Tuning positive feedback for signal detection in noisy dynamic environments
- Two-timescale stochastic gradient descent in continuous time with applications to joint online parameter estimation and optimal sensor placement
- Stochastic first-order methods with random constraint projection
- Convergence of Markovian stochastic approximation with discontinuous dynamics
- Nonlinear gossip
- A constrained optimization perspective on actor-critic algorithms and application to network routing
- Multiscale Q-learning with linear function approximation
- scientific article; zbMATH DE number 434712 (Why is no real title available?)
- Estimating the position of a moving object based on test disturbance of camera position
- On best-response dynamics in potential games
- Linear Convergence of Comparison-based Step-size Adaptive Randomized Search via Stability of Markov Chains
- Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals
- Learning to control a structured-prediction decoder for detection of HTTP-layer DDoS attackers
- Oja's algorithm for graph clustering, Markov spectral decomposition, and risk sensitive control
- Some Examples of Stochastic Approximation in Communications
- Distributed caching over heterogeneous mobile networks
- Stochastic approximation with long range dependent and heavy tailed noise
- Stability and delay of distributed scheduling algorithms for networks of conflicting queues
- Robust adaptive dynamic programming for linear and nonlinear systems: an overview
- Truncated stochastic approximation with moving bounds: convergence
- Stochastic fictitious play with continuous action sets
- Stochastic averaging and stochastic extremum seeking
- scientific article; zbMATH DE number 1043533 (Why is no real title available?)
- An online actor-critic algorithm with function approximation for constrained Markov decision processes
- scientific article; zbMATH DE number 1972910 (Why is no real title available?)
- On stochastic gradient and subgradient methods with adaptive steplength sequences
- Stabilization of stochastic approximation by step size adaptation
- Gradient estimation with simultaneous perturbation and compressive sensing
- Risk-constrained reinforcement learning with percentile risk criteria
- ASTRO-DF: a class of adaptive sampling trust-region algorithms for derivative-free stochastic optimization
- Stochastic Methods for Composite and Weakly Convex Optimization Problems
- Distributed stochastic approximation with local projections
- On sampling rates in simulation-based recursions
- A stochastic Kaczmarz algorithm for network tomography
- Reinforcement learning behaviors in sponsored search
- Rejoinder to ‘Reinforcement learning behaviors in sponsored search’
- Newton-based stochastic optimization using \(q\)-Gaussian smoothed functional algorithms
- On learning dynamics underlying the evolution of learning rules
- A Dynamical System Approach to Stochastic Approximations
- Stochastic Approximation Methods for Systems over an Infinite Horizon
- Optimal survey schemes for stochastic gradient descent with applications to \(M\)-estimation
- A Stochastic Subgradient Method for Nonsmooth Nonconvex Multilevel Composition Optimization
- Finite-time performance of distributed temporal-difference learning with linear function approximation
- On the fast convergence of random perturbations of the gradient flow
- scientific article; zbMATH DE number 7387192 (Why is no real title available?)
- Convergence of Recursive Stochastic Algorithms Using Wasserstein Divergence
- Distributed Bregman-distance algorithms for min-max optimization
- On gradient-based learning in continuous games
- A biologically plausible neural network for multichannel canonical correlation analysis
- Some limit properties of Markov chains induced by recursive stochastic algorithms
- scientific article; zbMATH DE number 7626794 (Why is no real title available?)
- scientific article; zbMATH DE number 7625165 (Why is no real title available?)
- Revisiting SIR in the age of COVID-19: explicit solutions and control problems
- Stochastic multilevel composition optimization algorithms with level-independent convergence rates
- A finite memory interacting Pólya contagion network and its approximating dynamical systems
- Stochastic recursive inclusions with non-additive iterate-dependent Markov noise
- Constant step stochastic approximations involving differential inclusions: stability, long-run convergence and applications
- A Q-Learning Algorithm for Discrete-Time Linear-Quadratic Control with Random Parameters of Unknown Distribution: Convergence and Stabilization
- Pathological subgradient dynamics
- Single Observation Adaptive Search for Continuous Simulation Optimization
- Full gradient DQN reinforcement learning: a provably convergent scheme
- Distributed algorithms for Internet-of-Things-enabled prosumer markets: a control theoretic perspective
- Adaptive learning algorithm convergence in passive and reactive environments
- An inertial Newton algorithm for deep learning
This page was built for publication:
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3527701)