Integral representations of shallow neural network with rectified power unit activation function.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 36166980.
- Also identified by DOI 10.1016/j.neunet.2022.09.005.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
In this paper we characterize the set of functions that can be represented by infinite width neural networks with RePU activation function max(0,x)<sup>p</sup>, when the network coefficients are regularized by an ℓ<sup>2/p</sup> (quasi)norm. Compared to the more well-known ReLU activation function (which corresponds to p=1), the RePU activation functions exhibit a greater degree of smoothness which makes them preferable in several applications. Our main result shows that such representations are possible for a given function if and only if the function is κ-order Lipschitz and its R-norm is finite. This extends earlier work on this topic that has been restricted to the case of the ReLU activation function and coefficient bounds with respect to the ℓ<sup>2</sup> norm. Since for q<2, ℓ<sup>q</sup> regularizations are known to promote sparsity, our results also shed light on the ability to obtain sparse neural network representations.
Medical subject headings
- Neural Networks, Computer