Han Shen

From MaRDI portal



List of research outcomes

This list is not complete and representing at the moment only items from zbMATH Open and arXiv. We are working on additional sources - please check back here soon!

PublicationDate of PublicationType
On penalty-based bilevel gradient descent method
Mathematical Programming. Series A. Series B
2025-12-11Paper
Principled penalty-based methods for bilevel reinforcement learning and RLHF
Journal of Machine Learning Research (JMLR)
2025-12-09Paper
Towards understanding asynchronous advantage actor-critic: convergence and linear speedup
IEEE Transactions on Signal Processing
2024-09-12Paper
Byzantine-Resilient Decentralized Policy Evaluation With Linear Function Approximation
IEEE Transactions on Signal Processing
2022-09-23Paper
Byzantine-Resilient Decentralized TD Learning with Linear Function Approximation2020-09-23Paper


Research outcomes over time


This page was built for person: Han Shen