Extracting insights from social media with large-scale matrix approximations

V Sindhwani,E Ting,R Lawrence,A Ghoting

doi:10.1147/jrd.2011.2163281

Abstract

Social media platforms such as blogs, Twitter® accounts, and online discussion sites are large-scale forums where every individual can potentially voice an influential public opinion. According to recent surveys, a massive number of Internet users are turning to such forums to collect recommendations and reviews for products and services, and to shape their individual choices and stances by the commentary of the online community as a whole. The unsupervised extraction of insight from unstructured user-generated web content requires new methodologies that are likely to be rooted in natural language processing and machine-learning techniques. Furthermore, the unprecedented scale of data begging to be analyzed necessitates the implementation of these methodologies on modern distributed computing platforms. In this paper, we describe a flexible new family of low-rank matrix approximation algorithms for modeling topics in a given corpus of documents (e.g., blog posts and tweets). We benchmark distributed optimization algorithms for running these models in a Hadoopi-enabled cluster environment. We describe online learning strategies for tracking the evolution of ongoing topics and rapidly detecting the emergence of new themes in a streaming setting.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Extracting insights from social media with large-scale matrix approximations

Abstract

Talk to us

Similar Papers

More From: IBM Journal of Research and Development

Lead the way for us

Journal: IBM Journal of Research and Development	Publication Date: Sep 1, 2011
Citations: 8

Similar Papers

The Comprehension of Figurative Language: What Is the Influence of Irony and Sarcasm on NLP Techniques?
Leila Weitzel ... Raul Freire Aguiar
-
Leila Weitzel, et. al.Leila Weitzel ... Raul Freire Aguiar
01 Jan 2015
01 Jan 2015

Twitter Archives and the Challenges of "Big Social Data" for Media and Communication Research
Jean Burgess ... Axel Bruns
M/C Journal | VOL. 15
Jean Burgess, et. al.Jean Burgess ... Axel Bruns
11 Oct 2012
M/C Journal | VOL. 15

Automated Identification of Aspirin-Exacerbated Respiratory Disease Using Natural Language Processing and Machine Learning: Algorithm Development and Evaluation Study.
Thanai Pongdee ... Sungrim Moon
JMIR AI | VOL. 2
Thanai Pongdee, et. al.Thanai Pongdee ... Sungrim Moon
12 Jun 2023
JMIR AI | VOL. 2

Development of Social Media Analytics System for Emergency Event Detection and Crisis Management
Shaheen Khatoon ... Majed A Alshamari
Computers, Materials & Continua | VOL. 68
Shaheen Khatoon, et. al.Shaheen Khatoon ... Majed A Alshamari
01 Jan 2020
Computers, Materials & Continua | VOL. 68

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Extracting insights from social media with large-scale matrix approximations

Abstract

Talk to us

Similar Papers

More From: IBM Journal of Research and Development