Why rectified linear activation functions? Why max-pooling? A possible explanation
From MaRDI portal
Recommendations
- Optimization under uncertainty explains empirical success of deep learning heuristics
- MorphoActivation: generalizing ReLU activation function by mathematical morphology
- Nonlinear approximation and (deep) ReLU networks
- Deep vs. shallow networks: an approximation theory perspective
- Spline representation and redundancies of one-dimensional ReLU neural network models
Cites work
Cited in
(4)- Optimization under uncertainty explains empirical success of deep learning heuristics
- Dying ReLU and initialization: theory and numerical examples
- MorphoActivation: generalizing ReLU activation function by mathematical morphology
- Replacing pooling functions in convolutional neural networks by linear combinations of increasing functions
This page was built for publication: Why rectified linear activation functions? Why max-pooling? A possible explanation
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2101281)