Use of Stochastic Automata for Parameter Self-Optimization with Multimodal Performance Criteria
From MaRDI portal
Cited in
(23)- Probabilistic automata
- Learning behavior of stochastic automata in the last stage of learning
- Choice of optimal subset of numbers using a learning automaton
- Reinforcement learning with internal expectation for the random neural network
- epsilon-optimality of a general class of learning algorithms
- When can the two-armed bandit algorithm be trusted?
- Optimal non-linear reinforcement schemes for stochastic automata
- A learning automata based algorithm for optimization of continuous complex functions
- Achieving Unbounded Resolution inFinitePlayer Goore Games Using Stochastic Automata, and Its Applications
- How Fast Is the Bandit?
- Distributed dynamic reinforcement of efficient outcomes in multiagent coordination and network formation
- An application of the stochastic automaton to the investment game
- Theoretical considerations of the parameter self-optimization by stochastic automata
- On ergodic two-armed bandits
- Convergence in models with bounded expected relative hazard rates
- Nonconvergence to saddle boundary points under perturbed reinforcement learning
- Regret bounds for Narendra-Shapiro bandit algorithms
- A strategy for controlling nonlinear systems using a learning automaton
- Stochastic automata and learning systems.
- Learning automata algorithms for pattern classification.
- Combinatorial optimization by stochastic automata
- A cooperative game of a pair of learning automata
- On conditional optimality of a class of learning automata in random environments
This page was built for publication: Use of Stochastic Automata for Parameter Self-Optimization with Multimodal Performance Criteria
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5575971)