Fetching the paper…
Reading the bibliography…
Topic models are widely used unsupervised models capable of learning topics - weighted lists of words and documents - from large collections of text documents.
1901
Earlier work this paper cites.
H. W. Kuhn, “The hungarian method for the assignment problem,” Naval research logistics quarterly , vol. 2, no. 1-2, pp. 83–97, 1955
1955
Earlier work this paper cites.
M. E. McCombs and D. L. Shaw, “The agenda-setting function of mass media,” Public Opinion Quarterly , vol. 36, no. 2, pp. 176–187, 1972
1972
Earlier work this paper cites.
H. M. Wallach, D. M. Mimno, and A. McCallum, “Rethinking LDA: Why Priors Matter,” in Advances in neural information processing systems , 2009, pp. 1973–1981
1981
Earlier work this paper cites.
S. Deerwester, S. T. Dumais, G. W. Furnas, T. K. Landauer, and R. Harshman, “Indexing by latent semantic analysis,” Journal of the American society for information science , vol. 41, no. 6, p. 391, 1990
1990
Earlier work this paper cites.
W. R. Gilks and P. Wild, “Adaptive Rejection Sampling for Gibbs Sampling,” Journal of the Royal Statistical Society. Series C (Applied Statistics) , vol. 41, no. 2, pp. 337–348, 1992
1992
Earlier work this paper cites.
C. Cortes and V. Vapnik, “Support-vector networks,” Machine Learning , vol. 20, no. 3, pp. 273–297, Sep 1995
1995
Earlier work this paper cites.
J. Pitman and M. Yor, “The two-parameter Poisson-Dirichlet distribution derived from a stable subordinator,” Annals of Probability , vol. 25, no. 2, pp. 855–900, Apr 1997
1997
Earlier work this paper cites.
D. D. Lee and H. S. Seung, “Learning the parts of objects by non-negative matrix factorization,” Nature , vol. 401, no. 6755, p. 788, 1999
1999
Earlier work this paper cites.
L. Breiman, “Random Forests,” Machine Learning , vol. 45, no. 1, pp. 5–32, Oct 2001
2001
Earlier work this paper cites.
D. M. Blei, A. Y. Ng, and M. I. Jordan, “Latent Dirichlet allocation,” Journal of Machine Learning Research , vol. 3, no. Jan, pp. 993–1022, 2003
2003
Earlier work this paper cites.
T. Jebara and R. Kondor, “Bhattacharyya and Expected Likelihood Kernels,” in Learning Theory and Kernel Machines . Springer Berlin Heidelberg, 2003, pp. 57–71
2003
Earlier work this paper cites.
C. X. Ling, J. Huang, and H. Zhang, “Auc: a statistically consistent and more discriminating measure than accuracy,” in Proceedings of the 18th International Joint Conference on Artificial intelligence , vol. 3, 2003, pp. 519–524
2003
Earlier work this paper cites.
B. Fitelson, “A probabilistic theory of coherence,” Analysis , vol. 63, no. 279, pp. 194–199, 2003
2003
Earlier work this paper cites.
T. L. Griffiths and M. Steyvers, “Finding scientific topics,” Proceedings of the National academy of Sciences , vol. 101, pp. 5228–5235, 2004
2004
Earlier work this paper cites.
F. Å. Nielsen, D. Balslev, and L. K. Hansen, “Mining the posterior cingulate: segregation between memory and pain components,” Neuroimage , vol. 27, no. 3, pp. 520–532, 2005
2005
Earlier work this paper cites.
P.-N. Tan, M. Steinbach, and V. Kumar, “Introduction to Data Mining,” Pearson , May 2005
2005
Earlier work this paper cites.
X. Wei and W. B. Croft, “Lda-based document models for ad-hoc retrieval,” in Proceedings of the 29th annual international ACM SIGIR conference on Research and development in information retrieval . ACM, 2006, pp. 178–185
2006
Earlier work this paper cites.
Y. W. Teh, M. I. Jordan, M. J. Beal, and D. M. Blei, “Hierarchical Dirichlet Processes,” Journal of the American Statistical Association , vol. 101, no. 476, pp. 1566–1581, 2006
2006
Earlier work this paper cites.
J. L. Boyd-Graber, D. M. Blei, and X. Zhu, “A topic model for word sense disambiguation,” in Proceedings of the 2007 Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning , 2007, pp. 1024–1033
2007
Earlier work this paper cites.
M. Steyvers and T. Griffiths, “Probabilistic topic models,” Handbook of latent semantic analysis , vol. 427, no. 7, pp. 424–440, 2007
2007
Earlier work this paper cites.
J. L. Boyd-Graber, D. M. Blei, and J. Zhu, “Probabalistic walks in semantic hierarchies as a topic model for WSD,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing , 2007
2007
Earlier work this paper cites.
C.-J. Lin, “Projected gradient methods for nonnegative matrix factorization,” Neural computation , vol. 19, no. 10, pp. 2756–2779, 2007
2007
Earlier work this paper cites.
A. De Waal and E. Barnard, “Evaluating topic models with stability,” in 19th Annual Symposium of the Pattern Recognition Association of South Africa , 2008
2008
Earlier work this paper cites.
I. Titov and R. T. McDonald, “A joint model of text and aspect ratings for sentiment summarization.” in ACL , 2008, pp. 308–316
2008
Earlier work this paper cites.
C. Boutsidis and E. Gallopoulos, “SVD based initialization: A head start for nonnegative matrix factorization,” Pattern Recognition , vol. 41, no. 4, pp. 1350–1362, Apr 2008
2008
Earlier work this paper cites.
H. M. Wallach, I. Murray, R. Salakhutdinov, and D. Mimno, “Evaluation methods for topic models,” in Proceedings of the 26th annual international conference on machine learning . ACM, 2009, pp. 1105–1112
2009
Earlier work this paper cites.
J. Chang, J. L. Boyd-Graber, S. Gerrish, C. Wang, and D. M. Blei, “Reading tea leaves: How humans interpret topic models.” in Proceedings of the 22nd International Conference on Neural Information Processing Systems , 2009, pp. 288–296
2009
Earlier work this paper cites.
C. Lin and Y. He, “Joint sentiment/topic model for sentiment analysis,” in Proceedings of the 18th ACM conference on Information and knowledge management . ACM, 2009, pp. 375–384
2009
Earlier work this paper cites.
L. AlSumait, D. Barbará, J. Gentle, and C. Domeniconi, “Topic significance ranking of lda generative models,” Machine Learning and Knowledge Discovery in Databases , pp. 67–82, 2009
2009
Earlier work this paper cites.
B. Settles, “Active learning literature survey,” University of Wisconsin-Madison Department of Computer Sciences, Tech. Rep., 2009
2009
Earlier work this paper cites.
D. Ramage, D. Hall, R. Nallapati, and C. D. Manning, “Labeled lda: A supervised topic model for credit attribution in multi-labeled corpora,” in Proceedings of the 2009 conference on empirical methods in natural language processing , 2009, pp. 248–256
2009
Earlier work this paper cites.
D. Newman, J. H. Lau, K. Grieser, and T. Baldwin, “Automatic evaluation of topic coherence,” in Human Language Technologies: The 2010 Annual Conference of the North American Chapter of the Association for Computational Linguistics . Association for Computational Linguistics, 2010, pp. 100–108
2010
Earlier work this paper cites.
K. M. Quinn, B. L. Monroe, M. Colaresi, M. H. Crespin, and D. R. Radev, “How to analyze political attention with minimal assumptions and costs,” American Journal of Political Science , vol. 54, no. 1, pp. 209–228, 2010
2010
Cited alongside, same era.
J. Grimmer, “A Bayesian Hierarchical Topic Model for Political Texts: Measuring Expressed Agendas in Senate Press Releases,” Political Analysis , vol. 18, no. 1, pp. 1–35, 2010
2010
Cited alongside, same era.
G. C. Cawley and N. L. C. Talbot, “On Over-fitting in Model Selection and Subsequent Selection Bias in Performance Evaluation,” Journal of Machine Learning Research , vol. 11, no. Jul, pp. 2079–2107, 2010
2010
Cited alongside, same era.
W. X. Zhao, J. Jiang, J. Weng, J. He, E.-P. Lim, H. Yan, and X. Li, “Comparing twitter and traditional media using topic models,” in European conference on information retrieval . Springer, 2011, pp. 338–349
2011
Cited alongside, same era.
D. Korenčić, S. Ristov, and J. Šnajder, “Getting the agenda right: measuring media agenda using topic models,” in Proceedings of the 2015 Workshop on Topic Models: Post-Processing and Applications . ACM, 2015, pp. 61–66
2015
Later among the works it cites.
Y. Wang, X. Zhao, Z. Sun, H. Yan, L. Wang, Z. Jin, L. Wang, Y. Gao, C. Law, and J. Zeng, “Peacock: Learning Long-Tail Topic Features for Industrial Applications,” ACM Trans. Intell. Syst. Technol. , vol. 6, no. 4, pp. 1–23, Aug 2015
2015
Later among the works it cites.
D. Greene and J. P. Cross, “Unveiling the political agenda of the european parliament plenary: A topical analysis,” in Proceedings of the ACM Web Science Conference . ACM, 2015, p. 2
2015
Later among the works it cites.
D. O’Callaghan, D. Greene, J. Carthy, and P. Cunningham, “An analysis of the coherence of descriptors in topic modeling,” Expert Systems with Applications , vol. 42, no. 13, pp. 5645–5657, 2015
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T.-I. Yang, A. J. Torget, and R. Mihalcea, “Topic modeling on historical newspapers,” in LaTeCH ’11 Proceedings of the 5th ACL-HLT Workshop on Language Technology for Cultural Heritage, Social Sciences, and Humanities . Association for Computational Linguistics, Jun 2011
2011
Cited alongside, same era.
M. Chen, X. Jin, and D. Shen, “Short text classification improved by learning multi-granularity topics,” in Proceedings of the Twenty-Second International Joint Conference on Artificial Intelligence , 2011, pp. 1776–1781
2011
Cited alongside, same era.
D. Mimno and D. Blei, “Bayesian checking for topic models,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 2011, pp. 227–237
2011
Cited alongside, same era.
D. Mimno, H. M. Wallach, E. Talley, M. Leenders, and A. McCallum, “Optimizing semantic coherence in topic models,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 2011, pp. 262–272
2011
Cited alongside, same era.
C. Musat, J. Velcin, S. Trausan-Matu, and M.-A. Rizoiu, “Improving topic evaluation using conceptual knowledge.” in 22nd International Joint Conference on Artificial Intelligence , vol. 3, 2011, pp. 1866–1871
2011
Cited alongside, same era.
C. Chen, L. Du, and W. Buntine, “Sampling Table Configurations for the Hierarchical Poisson-Dirichlet Process,” in Machine Learning and Knowledge Discovery in Databases . Springer Berlin Heidelberg, Sep 2011, pp. 296–311
2011
Cited alongside, same era.
J. Chuang, D. Ramage, C. Manning, and J. Heer, “Interpretation and trust: Designing model-driven visualizations for text analysis,” in Proceedings of the SIGCHI Conference on Human Factors in Computing Systems . ACM, 2012, pp. 443–452
2012
Cited alongside, same era.
S. Arora, R. Ge, and A. Moitra, “Learning topic models–going beyond svd,” in Foundations of Computer Science (FOCS), 2012 IEEE 53rd Annual Symposium on . IEEE, 2012, pp. 1–10
2012
Cited alongside, same era.
P. Xie, Y. Deng, and E. Xing, Diversifying Restricted Boltzmann Machine for Document Modeling . ACM, Aug 2015
2015
Later among the works it cites.
S. I. Nikolenko, S. Koltcov, and O. Koltsova, “Topic modelling for qualitative studies,” Journal of Information Science , vol. 43, no. 1, pp. 88–102, 2015
2015
Later among the works it cites.
C. May, F. Ferraro, A. McCree, J. Wintrode, D. Garcia-Romero, and B. Van Durme, “Topic identification and discovery on text and speech,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Lisbon, Portugal: Association for Computational Linguistics, Sep. 2015, pp. 2377–2387. [Online]. Available: https://www.aclweb.org/anthology/D15-1285
2015
Later among the works it cites.
C. Jacobi, W. van Atteveldt, and K. Welbers, “Quantitative analysis of large amounts of journalistic texts using topic modelling,” Digital Journalism , vol. 4, no. 1, pp. 89–106, 2016
2016
Later among the works it cites.
M. Brbić, M. Piškorec, V. Vidulin, A. Kriško, T. Šmuc, and F. Supek, “The landscape of microbial phenotypic traits and associated genes,” Nucleic acids research , 2016
2016
Later among the works it cites.
C. Puschmann and T. Scheffler, “Topic Modeling for Media and Communication Research: A Short Primer,” Aug 2016. [Online]. Available: https://papers.ssrn.com/sol3/papers.cfm?abstract_id=2836478
2016
Later among the works it cites.
S. J. Blair, Y. Bi, and M. D. Mulvenna, “Increasing topic coherence by aggregating topic models,” in International Conference on Knowledge Science, Engineering and Management . Springer, 2016, pp. 69–81
2016
Later among the works it cites.
P. Branco, L. Torgo, and R. P. Ribeiro, “A Survey of Predictive Modeling on Imbalanced Domains,” ACM Computing Surveys , vol. 49, no. 2, pp. 1–50, Nov 2016
2016
Later among the works it cites.
B. Krawczyk, “Learning from imbalanced data: open challenges and future directions,” Progress in Artificial Intelligence , vol. 5, no. 4, pp. 221–232, Nov 2016
2016
Later among the works it cites.
M. Roberts, B. Stewart, and D. Tingley, “Navigating the local modes of big data: The case of topic models,” in Computational Social Science . Cambridge University Press, New York, 2016, pp. 51–97
2016
Later among the works it cites.
M. E. Papadouka, N. Evangelopoulos, and G. Ignatow, “Agenda setting and active audiences in online coverage of human trafficking,” Information, Communication & Society , vol. 19, no. 5, pp. 655–672, 2016
2016
Later among the works it cites.
S. I. Nikolenko, “Topic quality metrics based on distributed word representations,” in Proceedings of the 39th International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 2016, pp. 1029–1032
2016
Later among the works it cites.
M. Yurochkin, A. Guha, and X. L. Nguyen, “Conic scan-and-cover algorithms for nonparametric topic modeling,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , ser. NIPS’17. Curran Associates Inc., 2017, p. 3881–3890
2017
Later among the works it cites.
N. Ramrakhiyani, S. Pawar, S. Hingmire, and G. K. Palshikar, “Measuring topic coherence through optimal word buckets,” in Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics . ACL, 2017, pp. 437–442
2017
Later among the works it cites.
M. Belford, B. M. Namee, and D. Greene, “Stability of topic modeling via matrix factorization,” Expert Systems with Applications , vol. 91, pp. 159 – 169, 2018
2018
Later among the works it cites.
D. Maier, A. Waldherr, P. Miltner, G. Wiedemann, A. Niekler, A. Keinert, B. Pfetsch, G. Heyer, U. Reber, T. Häussler et al. , “Applying lda topic modeling in communication research: Toward a valid and reliable methodology,” Communication Methods and Measures , vol. 12, no. 2-3, pp. 93–118, 2018
2018
Later among the works it cites.
D. Korenčić, S. Ristov, and J. Šnajder, “Document-based topic coherence measures for news media text,” Expert Systems with Applications , vol. 114, pp. 357–373, 2018
2018
Later among the works it cites.
F. Gilardi, C. R. Shipan, and B. Wueest, “Policy diffusion: The issue-definition stage,” 2018. [Online]. Available: https://fabriziogilardi.org/resources/papers/diffusion-policy-perceptions.pdf
2018
Later among the works it cites.
Q. Bai, K. Wei, M. Chen, Q. Hu, and L. He, “Mining Temporal Discriminant Frames via Joint Matrix Factorization: A Case Study of Illegal Immigration in the U.S. News Media,” in Knowledge Science, Engineering and Management . Springer International Publishing, 2018, pp. 260–267
2018
Later among the works it cites.
S. Suh, S. Shin, J. Lee, C. K. Reddy, and J. Choo, “Localized user-driven topic discovery via boosted ensemble of nonnegative matrix factorization,” Knowledge and Information Systems , vol. 56, no. 3, pp. 503–531, 2018
2018
Later among the works it cites.
L. Ying, J. M. Montgomery, and B. M. Stewart, “Inferring concepts from topics: Towards procedures for validating topics as measures,” PolMeth XXXVI, Cambridge, MA. Society for Political Methodology , vol. 5, 2019
2019
Later among the works it cites.
J. Rashid, S. M. Adnan Shah, A. Irtaza, T. Mahmood, M. W. Nisar, M. Shafiq, and A. Gardezi, “Topic modeling technique for text mining over biomedical text corpora through hybrid inverse documents frequency and fuzzy k-means clustering,” IEEE Access , vol. 7, pp. 146 070–146 080, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
Y. Xu, H. Nguyen, and Y. Li, “A semantic based approach for topic evaluation in information filtering,” IEEE Access , vol. 8, pp. 66 977–66 988, 2020
2020
Closest in time.
M. B. Pandur, J. Dobša, and L. Kronegger, “Topic modelling in social sciences: Case study of web of science,” 2020
2020
Closest in time.
M. Belford and D. Greene, “Ensemble topic modeling using weighted term co-associations,” Expert Systems with Applications , vol. 161, p. 113709, 2020
2020
Closest in time.
C. Doogan and W. Buntine, “Topic model or topic twaddle? re-evaluating semantic interpretability measures,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Online: Association for Computational Linguistics, Jun. 2021, pp. 3824–3848. [Online]. Available: https://www.aclweb.org/anthology/2021.naacl-main.300
2021
Closest in time.