Scinovex
articleTop 1% cited

Regularization Theory and Neural Networks Architectures

Neural Computation · 1995 · Vol. 7(2) · pp. 219–269
Federico GirosiMichael JonesTomaso Poggio

Abstract

We had previously shown that regularization principles lead to approximation schemes that are equivalent to networks with one layer of hidden units, called regularization networks. In particular, standard smoothness functionals lead to a subclass of regularization networks, the well known radial basis functions approximation schemes. This paper shows that regularization networks encompass a much broader range of approximation schemes, including many of the popular general additive models and some of the neural networks. In particular, we introduce new classes of smoothness functionals that lead to different classes of basis functions. Additive splines as well as some tensor product splines can be obtained from appropriate classes of smoothness functionals. Furthermore, the same generalization that extends radial basis functions (RBF) to hyper basis functions (HBF) also leads from additive models to ridge approximation models, containing as special cases Breiman's hinge functions, some forms of projection pursuit regression, and several types of neural networks. We propose to use the term generalized regularization networks for this broad class of approximation schemes that follow from an extension of regularization. In the probabilistic interpretation of regularization, the different classes of basis functions correspond to different classes of prior probabilities on the approximating function spaces, and therefore to different types of smoothness assumptions. In summary, different multilayer networks with one hidden layer, which we collectively call generalized regularization networks, correspond to different classes of priors and associated smoothness functionals in a classical regularization principle. Three broad classes are (1) radial basis functions that can be generalized to hyper basis functions, (2) some tensor product splines, and (3) additive splines that can be generalized to schemes of the type of ridge approximation, hinge functions, and several perceptron-like neural networks with one hidden layer.

Neural Networks and ApplicationsImage and Signal Denoising MethodsMedical Image Segmentation TechniquesRegularization (linguistics)MathematicsBasis functionRadial basis functionApplied mathematicsArtificial neural networkRegularization perspectives on support vector machinesInverse problemAlgorithmComputer science
Citations
1,349
FWCI
50.55
field-weighted impact
References
157
Percentile
100%
vs. same field & year
Citations per year
Cited by
A Survey on Deep Learning for Data-Driven Soft Sensors
IEEE Transactions on Industrial Informatics · 2021 · 563 citations
Shape matching and object recognition using shape contexts
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2002 · 6,295 citations
An overview of statistical learning theory
IEEE Transactions on Neural Networks · 1999 · 6,176 citations
Three learning phases for radial-basis-function networks
Neural Networks · 2001 · 524 citations
Point Set Registration: Coherent Point Drift
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2010 · 2,703 citations
References
Theory and applications of the multiquadric-biharmonic method 20 years of discovery 1968–1988
Computers & Mathematics with Applications · 1990 · 847 citations
Theory of reproducing kernels
Transactions of the American Mathematical Society · 1950 · 5,387 citations
The self-organizing map
Proceedings of the IEEE · 1990 · 8,115 citations
Modeling by shortest data description
Automatica · 1978 · 5,959 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.