Variational learning in nonlinear Gaussian belief networks

被引：48

作者：

Frey, BJ ^{[1
]}

Hinton, GE

机构：

[1] Univ Illinois, Beckman Inst, Urbana, IL 61801 USA

[2] UCL, Gatsby Computat Neurosci Unit, London WC1N 3AR, England

来源：

NEURAL COMPUTATION | 1999年 / 11卷 / 01期

关键词：

D O I：

10.1162/089976699300016872

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We view perceptual tasks such as vision and speech recognition as inference problems where the goal is to estimate the posterior distribution over latent variables (e.g., depth in stereo vision) given the sensory input. The recent flurry of research in independent component analysis exemplifies the importance of inferring the continuous-valued latent variables of input data. The latent variables found by this method are linearly related to the input, but perception requires nonlinear inferences such as classification and depth estimation. In this article, we present a unifying framework for stochastic neural networks with nonlinear latent variables. Nonlinear units are obtained by passing the outputs of linear gaussian units through various nonlinearities. We present a general variational method that maximizes a lower bound on the likelihood of a training set and give results on two visual feature extraction problems. We also show how the variational method can be used for pattern classification and compare the performance of these nonlinear networks with other methods on the problem of handwritten digit recognition.

引用

页码：193 / 213

页数：21

共 27 条

[1]

AMARI S, 1996, ADV NEURAL INFORMATI

[2]

AMARI SI, 1985, DIFFERNTIAL GEOMETRI

[3]

[Anonymous], GRAPHICAL MODELS MAC

[4] SELF-ORGANIZING NEURAL NETWORK THAT DISCOVERS SURFACES IN RANDOM-DOT STEREOGRAMS [J].