stat.MLFeb 4, 2025

Networks with Finite VC Dimension: Pro and Contra

Authors: Vera Kurkova, Marcello Sanguineti

Organizations: Institute of Computer Science of the Czech Academy of Sciences Pod Vodárenskou věží, 2 - 18207 Prague, Czech Republic · DIBRIS University of Genova Via Opera Pia, 13 - 16145 Genova, Italy

Abstract

Approximation and learning of classifiers of large data sets by neural networks in terms of high-dimensional geometry and statistical learning theory are investigated. The influence of the VC dimension of sets of input-output functions of networks on approximation capabilities is compared with its influence on consistency in learning from samples of data. It is shown that, whereas finite VC dimension is desirable for uniform convergence of empirical errors, it may not be desirable for approximation of functions drawn from a probability distribution modeling the likelihood that they occur in a given type of application. Based on the concentration-of-measure properties of high dimensional geometry, it is proven that both errors in approximation and empirical errors behave almost deterministically for networks implementing sets of input-output functions with finite VC dimensions in processing large data sets. Practical limitations of the universal approximation property, the trade-offs between the accuracy of approximation and consistency in learning from data, and the influence of depth of networks with ReLU units on their accuracy and consistency are discussed.

Figures & tables

Explore similar work

CardsList
  1. Approximating Smooth Functionals with ReLU Networks

    Sep 14, 2026Shuhao JiaoReLU Neural NetworksNeural Network Approximation Theory

  2. Approximation Theory for Neural Networks: Old and New

    May 20, 2026Soumendu Sundar Mukherjee, Himasish TalukdarUniversal ApproximationNeural Network Approximation Theory

  3. Do Neural Networks Really Beat the Curse of Dimensionality? A Bit-Complexity View

    Aug 2, 2026Tong Mao, Jinchao XuNeural Network Approximation Theory