Scinovex
articleTop 10% cited

A Connection Between Score Matching and Denoising Autoencoders

Neural Computation · 2011 · Vol. 23(7) · pp. 1661–1674
Pascal Vincent

Abstract

Denoising autoencoders have been previously shown to be competitive alternatives to restricted Boltzmann machines for unsupervised pretraining of each layer of a deep architecture. We show that a simple denoising autoencoder training criterion is equivalent to matching the score (with respect to the data) of a specific energy-based model to that of a nonparametric Parzen density estimator of the data. This yields several useful insights. It defines a proper probabilistic model for the denoising autoencoder technique, which makes it in principle possible to sample from them or rank examples by their energy. It suggests a different way to apply score matching that is related to learning to denoise and does not require computing second derivatives. It justifies the use of tied weights between the encoder and decoder and suggests ways to extend the success of denoising autoencoders to a larger family of energy-based models.

Generative Adversarial Networks and Image SynthesisAdvanced Neural Network ApplicationsLattice Boltzmann Simulation StudiesAutoencoderRestricted Boltzmann machineArtificial intelligenceNoise reductionPattern recognition (psychology)EstimatorComputer scienceMatching (statistics)Probabilistic logicUnsupervised learning

MeSH terms

Computer SimulationModels, Statistical
Citations
933
FWCI
10.74
field-weighted impact
References
29
Percentile
99%
vs. same field & year
Citations per year
Cited by
Neural Networks and Deep Learning
Machine Learning · 2015 · 920 citations
Representation Learning: A Review and New Perspectives
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2013 · 12,724 citations
Diffusion Models: A Comprehensive Survey of Methods and Applications
ACM Computing Surveys · 2023 · 1,297 citations
References
Advances in neural information processing systems 7
Neurocomputing · 1997 · 22,296 citations
Hybrid Monte Carlo
Physics Letters B · 1987 · 3,787 citations
Training Products of Experts by Minimizing Contrastive Divergence
Neural Computation · 2002 · 4,959 citations
A Fast Learning Algorithm for Deep Belief Nets
Neural Computation · 2006 · 16,253 citations
Representation Learning: A Review and New Perspectives
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2013 · 12,724 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.