Matteo Papini
From MaRDI portal
List of research outcomes
This list is not complete and representing at the moment only items from zbMATH Open and arXiv. We are working on additional sources - please check back here soon!
| Publication | Date of Publication | Type |
|---|---|---|
| Search or split: policy gradient with adaptive policy space Machine Learning | 2025-11-10 | Paper |
| Importance-weighted offline learning done right | 2025-03-06 | Paper |
| Online learning with off-policy feedback | 2025-02-24 | Paper |
| Smoothing policies and safe policy gradients Machine Learning | 2023-06-12 | Paper |
| Importance sampling techniques for policy optimization | 2020-10-05 | Paper |
Research outcomes over time
This page was built for person: Matteo Papini