REVIEW 2 cited by
Expressivity and Approximation Properties of Deep Neural Networks with ReLU$^k$ Activation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
In this paper, we investigate the expressivity and approximation properties of deep neural networks employing the ReLU$^k$ activation function for $k \geq 2$. Although deep ReLU networks can approximate polynomials effectively, deep ReLU$^k$ networks have the capability to represent higher-degree polynomials precisely. Our initial contribution is a comprehensive, constructive proof for polynomial representation using deep ReLU$^k$ networks. This allows us to establish an upper bound on both the size and count of network parameters. Consequently, we are able to demonstrate a suboptimal approximation rate for functions from Sobolev spaces as well as for analytic functions. Additionally, through an exploration of the representation power of deep ReLU$^k$ networks for shallow networks, we reveal that deep ReLU$^k$ networks can approximate functions from a range of variation spaces, extending beyond those generated solely by the ReLU$^k$ activation function. This finding demonstrates the adaptability of deep ReLU$^k$ networks in approximating functions within various variation spaces.
Forward citations
Cited by 2 Pith papers
-
Digital Twin Channel-Aided CSI Prediction: An Environment-Based Subspace Extraction Approach for Achieving Low Overhead and High Robustness
Using a digital-twin-derived channel subspace basis as a prior, the method predicts full spatial-frequency CSI from partial pilots, claiming up to 50% pilot overhead reduction and one-coherence-time-ahead prediction i...
-
Variational formulation based on duality to solve partial differential equations: Use of B-splines and machine learning approximants
The authors extend their prior dual (Lagrange multiplier) variational framework to steady and transient convection-diffusion and heat equations, discretizing with B-splines or RePU networks, and report numerical accur...
Discussion (0). Continue with ORCID to comment.