Reinforcement learning based algorithms for average cost Markov decision processes (Q2643632): Difference between revisions

From MaRDI portal
ReferenceBot (talk | contribs)
Changed an Item
Import241208061232 (talk | contribs)
Normalize DOI.
 
Property / DOI
 
Property / DOI: 10.1007/s10626-006-0003-y / rank
Normal rank
 
Property / DOI
 
Property / DOI: 10.1007/S10626-006-0003-Y / rank
 
Normal rank

Latest revision as of 12:37, 19 December 2024

scientific article
Language Label Description Also known as
English
Reinforcement learning based algorithms for average cost Markov decision processes
scientific article

    Statements

    Reinforcement learning based algorithms for average cost Markov decision processes (English)
    0 references
    27 August 2007
    0 references
    actor-critic algorithms
    0 references
    two timescale stochastic approximation
    0 references
    Markov decision processes
    0 references
    policy iteration
    0 references
    simultaneous perturbation stochastic approximation
    0 references
    normalized Hadamard matrices
    0 references
    reinforcement learning
    0 references
    TD-learning
    0 references

    Identifiers