Robust machine learning algorithms for text analysis

Shikun Ke,José Luis Montiel Olea,James Nesbit

doi:10.3982/qe1825

Abstract

We study the Latent Dirichlet Allocation model, a popular Bayesian algorithm for text analysis. We show that the model's parameters are not identified, which suggests that the choice of prior matters. We characterize the range of values that the posterior mean of a given functional of the model's parameters can attain in response to a change in the prior, and we suggest two algorithms that report this range. Both of our algorithms rely on obtaining multiple Nonnegative Matrix Factorizations of either the posterior draws of the corpus' population term‐document frequency matrix or of its maximum likelihood estimator. The key idea is to maximize/minimize the functional of interest over all these nonnegative matrix factorizations. To illustrate the applicability of our results, we revisit recent work studying the effects of increased transparency on the communication structure of monetary policy discussions in the United States.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Robust machine learning algorithms for text analysis

Abstract

Talk to us

Similar Papers

More From: Quantitative Economics

Lead the way for us

Journal: Quantitative Economics	Publication Date: Jan 1, 2024
License type: CC BY-NC 4.0

Similar Papers

Confirmatory Factor Analyses in Psychological Test Adaptation and Development
Kay Brauer ... Matthias Ziegler
Psychological Test Adaptation and Development | VOL. 4
Kay Brauer, et. al.Kay Brauer ... Matthias Ziegler
01 Feb 2023
Psychological Test Adaptation and Development | VOL. 4

Learning Topic Models -- Going beyond SVD
Sanjeev Arora ... Ankur Moitra
-
Sanjeev Arora, et. al.Sanjeev Arora ... Ankur Moitra
01 Oct 2012
01 Oct 2012

Application of Weibull Distribution to Hidden Markov Model for Non-Negative Factorization Matrix
Edesiri Bridget Nkemnole ... Joshua Olaniyi Bamigbode
European Journal of Theoretical and Applied Sciences | VOL. 2
Edesiri Bridget Nkemnole, et. al.Edesiri Bridget Nkemnole ... Joshua Olaniyi Bamigbode
01 Jan 2024
European Journal of Theoretical and Applied Sciences | VOL. 2

Non-negative multiple matrix factorization with Euclidean and kullback-leibler mixed divergences
Masahiro Kohjima ... Tatsushi Matsubayashi
-
Masahiro Kohjima, et. al.Masahiro Kohjima ... Tatsushi Matsubayashi
01 Dec 2016
01 Dec 2016

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Robust machine learning algorithms for text analysis

Abstract

Talk to us

Similar Papers

More From: Quantitative Economics