Han Shen
From MaRDI portal
List of research outcomes
This list is not complete and representing at the moment only items from zbMATH Open and arXiv. We are working on additional sources - please check back here soon!
| Publication | Date of Publication | Type |
|---|---|---|
| On penalty-based bilevel gradient descent method Mathematical Programming. Series A. Series B | 2025-12-11 | Paper |
| Principled penalty-based methods for bilevel reinforcement learning and RLHF Journal of Machine Learning Research (JMLR) | 2025-12-09 | Paper |
| Towards understanding asynchronous advantage actor-critic: convergence and linear speedup IEEE Transactions on Signal Processing | 2024-09-12 | Paper |
| Byzantine-Resilient Decentralized Policy Evaluation With Linear Function Approximation IEEE Transactions on Signal Processing | 2022-09-23 | Paper |
| Byzantine-Resilient Decentralized TD Learning with Linear Function Approximation | 2020-09-23 | Paper |
Research outcomes over time
This page was built for person: Han Shen