A Generalized Minimax Q-Learning Algorithm for Two-Player Zero-Sum Stochastic Games (Q6076027)
From MaRDI portal
!
This is the item page for this Wikibase entity, intended for internal use and editing purposes. Please use the normal view instead:
scientific article; zbMATH DE number 7740982
| Language | Label | Description | Also known as |
|---|---|---|---|
| default for all languages | No label defined |
||
| English | A Generalized Minimax Q-Learning Algorithm for Two-Player Zero-Sum Stochastic Games |
scientific article; zbMATH DE number 7740982 |
Statements
A Generalized Minimax Q-Learning Algorithm for Two-Player Zero-Sum Stochastic Games (English)
0 references
21 September 2023
0 references
minimax Q-learning
0 references
successive relaxation
0 references
two-player zero-sum games
0 references