On the convergence of reinforcement learning with Monte Carlo exploring starts (Q2665181)
From MaRDI portal
!
This is the item page for this Wikibase entity, intended for internal use and editing purposes. Please use the normal view instead:
scientific article; zbMATH DE number 7429801
| Language | Label | Description | Also known as |
|---|---|---|---|
| default for all languages | No label defined |
||
| English | On the convergence of reinforcement learning with Monte Carlo exploring starts |
scientific article; zbMATH DE number 7429801 |
Statements
On the convergence of reinforcement learning with Monte Carlo exploring starts (English)
0 references
18 November 2021
0 references
reinforcement learning
0 references
Markov decision processes
0 references
stochastic control
0 references
Monte Carlo exploring starts
0 references
optimistic policy iteration
0 references
convergence
0 references
stochastic shortest path problem
0 references
0.8110118508338928
0 references
0.7585216164588928
0 references
0.7343509197235107
0 references
0.7227733731269836
0 references
0.7225967049598694
0 references