Matteo Papini

From MaRDI portal



List of research outcomes

This list is not complete and representing at the moment only items from zbMATH Open and arXiv. We are working on additional sources - please check back here soon!

PublicationDate of PublicationType
Search or split: policy gradient with adaptive policy space
Machine Learning
2025-11-10Paper
Importance-weighted offline learning done right2025-03-06Paper
Online learning with off-policy feedback2025-02-24Paper
Smoothing policies and safe policy gradients
Machine Learning
2023-06-12Paper
Importance sampling techniques for policy optimization2020-10-05Paper


Research outcomes over time


This page was built for person: Matteo Papini