Learning multi-faceted representations of individuals from heterogeneous evidence using neural networks
arXiv:1510.05198
Abstract
Inferring latent attributes of people online is an important social computing task, but requires integrating the many heterogeneous sources of information available on the web. We propose learning individual representations of people using neural nets to integrate rich linguistic and network evidence gathered from social media. The algorithm is able to combine diverse cues, such as the text a person writes, their attributes (e.g. gender, employer, education, location) and social relations to other people. We show that by integrating both textual and network evidence, these representations offer improved performance at four important tasks in social media inference on Twitter: predicting (1) gender, (2) occupation, (3) location, and (4) friendships for users. Our approach scales to large datasets and the learned representations can be used as general features in and have the potential to benefit a large number of downstream tasks including link prediction, community detection, or probabilistic reasoning over social networks.
References in corpus (3)
Cited by in corpus (8)
- A Tutorial on Network Embeddings
- Quantifying Mental Health from Social Media with Neural User Embeddings
- Overcoming Language Variation in Sentiment Analysis with Social Attention
- Does Yoga Make You Happy? Analyzing Twitter User Happiness using Textual and Temporal Information
- Analysis of Twitter Users' Lifestyle Choices using Joint Embedding Model
- Toward Socially-Infused Information Extraction: Embedding Authors, Mentions, and Entities
- Evidence Transfer for Improving Clustering Tasks Using External Categorical Evidence
- Learning Representations of Social Media Users