Stochastic Ratio Matching of RBMs for Sparse High-Dimensional Inputs

Yann Dauphin; Yoshua Bengio

2013 NIPS NeurIPS 2013

Stochastic Ratio Matching of RBMs for Sparse High-Dimensional Inputs

Abstract

Sparse high-dimensional data vectors are common in many application domains where a very large number of rarely non-zero features can be devised. Unfortunately, this creates a computational bottleneck for unsupervised feature learning algorithms such as those based on auto-encoders and RBMs, because they involve a reconstruction step where the whole input vector is predicted from the current feature values. An algorithm was recently developed to successfully handle the case of auto-encoders, based on an importance sampling scheme stochastically selecting which input elements to actually reconstruct during training for each particular example. To generalize this idea to RBMs, we propose a stochastic ratio-matching algorithm that inherits all the computational advantages and unbiasedness of the importance sampling scheme. We show that stochastic ratio matching is a good estimator, allowing the approach to beat the state-of-the-art on two bag-of-word text classification benchmarks (20 Newsgroups and RCV1), while keeping computational cost linear in the number of non-zeros.

🌉 Interdisciplinary Bridge — Deep Learning and Machine Learning

📈 Trend Setter — Autoencoders

🧭 Keyword Pioneer — stochastic ratio matching

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Speech & Audio

🐣 Hot Topic Early Bird — importance sampling

Authors

Yann Dauphin , Yoshua Bengio

Topics

Machine Learning > Core Methods > Representation Learning Machine Learning > Learning Types > Unsupervised Learning Deep Learning > Architectures > Autoencoders Deep Learning > Architectures > Neural Networks Machine Learning > Core Methods > Feature Learning Deep Learning > Learning Types > Unsupervised Learning

Keywords

unsupervised learning feature learning unsupervised feature learning importance sampling sparse data stochastic ratio matching ratio matching restricted boltzmann machine sparse datum

Download PDF

Related papers

Latent Structured Active Learning 2013

On Flat versus Hierarchical Classification in Large-Scale Taxonomies 2013

Generalized Method-of-Moments for Rank Aggregation 2013

Third-Order Edge Statistics: Contour Continuation, Curvature, and Cortical Connections 2013

Accelerated Mini-Batch Stochastic Dual Coordinate Ascent 2013