stat.MLSep 17, 2026
SaveError bounds in Sobolev norms for approximations with norm constrained ReLU neural networks
Abstract
Recent studies have shown that smooth functions can be well approximated by ReLU neural networks with path norm constraint on the weights. We extend these results from uniform approximation to approximation in Sobolev norm. Specifically, we analyze how well Sobolev functions in can be approximated by neural networks with width , depth and path norm bounded by , when the approximation error is measured in the -norm. For shallow networks with depth , we derive the approximation error bound , when the smoothness index satisfies and the input is -dimensional. For deep networks, we remove the restriction on the smoothness by showing that the approximation bound holds if the width and depth are sufficiently large.
Explore similar work
This paper studies approximation by shallow ReLU networks, , together with their generalization behavior under path-norm control. For the -type integral spaces , , spherical harmonic analysis yields approximation bounds for shallow networks. In particular, when is the uniform measure and , the approximation rate is for and for , where . Approximation bounds for Sobolev spaces , , are obtained through embeddings into spectral Barron spaces. For nonparametric regression with sub-Gaussian noise, path-norm-regularized shallow ReLU networks achieve minimax-optimal rates over and over , with matching lower bounds up to logarithmic factors.
Approximation and learning of anisotropic and mixed smooth functions by deep ReLU neural networks
This paper studies how efficiently deep ReLU neural networks can approximate and learn smooth functions. When the error is measured in norm and the approximator is a network with width and depth , recent works have proven the supper approximation rate for Besov space under the Sobolev embedding condition . In order to overcome the curse of dimensionality in this rate, we extent this result to anisotropic and mixed smooth function classes. We establish the approximation rate for anisotropic Besov space with anisotropic smoothness under the embedding condition , where the mean smoothness . For mixed smooth Besov space with mixed smoothness , we show that the approximation rate holds up to logarithmic factors. Using these results, we also derive approximation bounds for the composition of anisotropic Besov functions. As an application, it is shown that deep ReLU neural networks can achieve minimax optimal rates up to logarithmic factors for a wide range of smooth function classes.
Approximating Smooth Functionals with ReLU Networks
We study the uniform approximation of smooth scalar-valued functionals on an infinite-dimensional separable Hilbert space by ReLU neural networks. A key feature in deep learning for functional data is the varying importance of different coordinates/dimensions. Representing the functional input in a basis expansion, we quantify the importance of each coordinate through both the magnitude of its corresponding basis score and the directional sensitivity of the target functional. Our analysis combines coordinate truncation, anisotropic partitioning, local Taylor approximation, and ReLU network realization, while allowing unrestricted interactions among the retained coordinates. We establish a general nonasymptotic upper bound for the uniform approximation error and a complementary pseudo-dimension-based lower bound for the worst-case approximation error. Under generalized exponential coordinate decay , with , the upper and lower bounds match at the leading order, which is stretched-exponential in the logarithm of the network size budget, and thus yield the nearly optimal approximation rate. This is the first work to characterize neural network approximation error for infinite-dimensional functional inputs explicitly through the joint dimensional decay of coordinate magnitudes and directional sensitivities.