Sparsity in long-time control of neural ODEs

From MaRDI portal
Publication:6099693



Abstract: We consider the neural ODE and optimal control perspective of supervised learning, with ell1-control penalties, where rather than only minimizing a final cost (the emph{empirical risk}) for the state, we integrate this cost over the entire time horizon. We prove that any optimal control (for this cost) vanishes beyond some positive stopping time. When seen in the discrete-time context, this result entails an emph{ordered} sparsity pattern for the parameters of the associated residual neural network: ordered in the sense that these parameters are all 0 beyond a certain layer. Furthermore, we provide a polynomial stability estimate for the empirical risk with respect to the time horizon. This can be seen as a emph{turnpike property}, for nonsmooth dynamics and functionals with ell1-penalties, and without any smallness assumptions on the data, both of which are new in the literature.




Cites work









This page was built for publication: Sparsity in long-time control of neural ODEs

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6099693)