Finite Sample Complexity Guarantees for Feedforward Neural Networks via o-Minimality
Abstract
We prove that, in a precise sense, a broad class of feedforward neural networks (FFNN) learn (have finite sample complexity) in the PAC model: every fixed finite feedforward architecture whose layers are definable in an o-minimal structure has finite sample complexity in the agnostic PAC setting, even with unbounded parameters. Quantitative versions are obtained for general FFNNs, possibly using multivariate non-linearities, which are Pfaffian arranged in a general computational graph (DAG). This covers standard fixed-size MLPs, CNNs, GNNs, and transformers with fixed sequence length, together with the operations and layers typically used in such architectures, including linear projections, residual connections, attention mechanisms, pooling layers, normalization layers, and admissible positional encodings. Hence, distribution-free learnability for modern non-recurrent architectures is not an exceptional property of particular activations or architecture-specific VC arguments, but a consequence of tame feedforward computation. Our results reposition finite-sample PAC learnability as a baseline rather than a differentiator: they shift the focus of architectural comparison toward inductive biases, symmetries and geometric priors, scalability, and optimization behaviour.
Then back it, or bet against it.
Related papers
Open the market on this paper to see 7 more related papers.