A Multiplicative Model for Learning Distributed Text-Based Attribute Representations

Ryan Kiros; Richard Zemel; Ruslan Salakhutdinov

2014 NIPS NeurIPS 2014

A Multiplicative Model for Learning Distributed Text-Based Attribute Representations

Abstract

In this paper we propose a general framework for learning distributed representations of attributes: characteristics of text whose representations can be jointly learned with word embeddings. Attributes can correspond to a wide variety of concepts, such as document indicators (to learn sentence vectors), language indicators (to learn distributed language representations), meta-data and side information (such as the age, gender and industry of a blogger) or representations of authors. We describe a third-order model where word context and attribute vectors interact multiplicatively to predict the next word in a sequence. This leads to the notion of conditional word similarity: how meanings of words change when conditioned on different attributes. We perform several experimental tasks including sentiment classification, cross-lingual document classification, and blog authorship attribution. We also qualitatively evaluate conditional word neighbours and attribute-conditioned text generation.

🌉 Interdisciplinary Bridge — Deep Learning and Machine Learning and Natural Language Processing

📈 Trend Setter — Text Classification

🧭 Keyword Pioneer — cross-lingual classification

🐣 Hot Topic Early Bird — text classification

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio

Authors

Ryan Kiros , Richard Zemel , Ruslan Salakhutdinov

Topics

Machine Learning > Core Methods > Embedding Learning Natural Language Processing > Applications > Text Classification Natural Language Processing > Resources & Methods > Text Representation Deep Learning > Learning Types > Representation Learning Deep Learning > Models > Language Models

Keywords

text classification sentiment classification authorship attribution distributed representation word embedding cross-lingual classification attribute representation multiplicative model

Download PDF

Related papers

Information-based learning by agents in unbounded state spaces 2014

Stochastic Gradient Descent, Weighted Sampling, and the Randomized Kaczmarz algorithm 2014

Partition-wise Linear Models 2014

Active Regression by Stratification 2014

Cone-Constrained Principal Component Analysis 2014