Fetching the paper…
Reading the bibliography…
Data analysts have long sought to turn unstructured text data into meaningful concepts.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Nils Reimers and Iryna Gurevych. 2019 · 1908
Earlier work this paper cites.
Latent Dirichlet Allocation
David M Blei, Andrew Y Ng, and Michael I Jordan. 2003 · 2003
Earlier work this paper cites.
Finding Scientific Topics
Thomas L Griffiths and Mark Steyvers. 2004 · 2004
Earlier work this paper cites.
Constructing Grounded Theory: A Practical Guide through Qualitative Analysis
Kathy Charmaz. 2006 · 2006
Earlier work this paper cites.
Topic Significance Ranking of LDA Generative Models. In Proceedings of the 2009th European Conference on Machine Learning and Knowledge Discovery in Databases - Volume Part I (Bled, Slovenia) (ECMLPKDD’09) . Springer-Verlag, Berlin, Heidelberg, 67–82
Loulwah AlSumait, Daniel Barbará, James Gentle, and Carlotta Domeniconi. 2009 · 2009
Earlier work this paper cites.
Reading Tea Leaves: How Humans Interpret Topic Models. In Advances in Neural Information Processing Systems , Y. Bengio, D. Schuurmans, J. Lafferty, C. Williams, and A. Culotta (Eds.), Vol. 22. Curran Associates, Inc
Jonathan Chang, Sean Gerrish, Chong Wang, Jordan Boyd-Graber, and David Blei. 2009 · 2009
Earlier work this paper cites.
Topic modeling for the social sciences. In NIPS 2009 workshop on applications for topic models: text and beyond , Vol. 5. 1–4
Daniel Ramage, Evan Rosen, Jason Chuang, Christopher D Manning, and Daniel A McFarland. 2009 · 2009
Earlier work this paper cites.
Characterizing Microblogs with Topic Models
Daniel Ramage, Susan T. Dumais, and Daniel J. Liebling. 2010 · 2010
Earlier work this paper cites.
You Are What You Tweet: Analyzing Twitter for Public Health. In Proceedings of the international AAAI conference on web and social media , Vol. 5. 265–272
Michael Paul and Mark Dredze. 2011 · 2011
Earlier work this paper cites.
Topic Model Diagnostics: Assessing Domain Relevance via Topical Alignment. In Proceedings of the 30th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 28) , Sanjoy Dasgupta and David McAllester (Eds.). PMLR, Atlanta, Georgia, USA, 612–620
Jason Chuang, Sonal Gupta, Christopher Manning, and Jeffrey Heer. 2013 · 2013
Earlier work this paper cites.
Exploiting Affinities between Topic Modeling and the Sociological Perspective on Culture: Application to Newspaper Coverage of US Government Arts Funding
Paul DiMaggio, Manish Nag, and David Blei. 2013 · 2013
Earlier work this paper cites.
Computer-Assisted Content Analysis: Topic Models for Exploring Multiple Subjective Interpretations. In Advances in Neural Information Processing Systems workshop on human-propelled machine learning . 1–9
Jason Chuang, John D. Wilkerson, Rebecca Weiss, Dustin Tingley, and Brandon M Stewart. 2014 · 2014
Earlier work this paper cites.
Curiosity, Creativity, and Surprise as Analytic Tools: Grounded Theory Method
Michael Muller. 2014 · 2014
Earlier work this paper cites.
LDAvis: A method for visualizing and interpreting topics. In Proceedings of the Workshop on Interactive Language Learning, Visualization, and Interfaces . Association for Computational Linguistics, Baltimore, Maryland, USA, 63–70
Carson Sievert and Kenneth Shirley. 2014 · 2014
Earlier work this paper cites.
FeatureInsight: Visual support for error-driven feature ideation in text classification. In 2015 IEEE Conference on Visual Analytics Science and Technology (VAST) . IEEE, 105–112
Michael Brooks, Saleema Amershi, Bongshin Lee, Steven M Drucker, Ashish Kapoor, and Patrice Simard. 2015 · 2015
Earlier work this paper cites.
TopicCheck: Interactive Alignment for Assessing Topic Model Stability. In Proceedings of the 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Denver, Colorado, 175–184
Jason Chuang, Margaret E. Roberts, Brandon M. Stewart, Rebecca Weiss, Dustin Tingley, Justin Grimmer, and Jeffrey Heer. 2015 · 2015
Earlier work this paper cites.
A Frame of Mind: Using Statistical Models for Detection of Framing and Agenda Setting Campaigns. In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Beijing, China, 1629–1638
Oren Tsur, Dan Calacci, and David Lazer. 2015 · 2015
Earlier work this paper cites.
Bad Company—Neighborhoods in Neural Embedding Spaces Considered Harmful. In Proceedings of COLING 2016, the 26th International Conference on Computational Linguistics: Technical Papers . The COLING 2016 Organizing Committee, Osaka, Japan, 2785–2796
Johannes Hellrich and Udo Hahn. 2016 · 2016
Earlier work this paper cites.
Machine Learning and Grounded Theory Method: Convergence, Divergence, and Combination. In Proceedings of the 2016 ACM International Conference on Supporting Group Work (Sanibel Island, Florida, USA) (GROUP ’16) . Association for Computing Machinery, New York, NY, USA, 3–8
Michael Muller, Shion Guha, Eric P.S. Baumer, David Mimno, and N. Sadat Shami. 2016 · 2016
Earlier work this paper cites.
Comparing grounded theory and topic modeling: Extreme divergence or unlikely convergence?
Eric P. S. Baumer, David Mimno, Shion Guha, Emily Quan, and Geri K. Gay. 2017 · 2017
Earlier work this paper cites.
Aeonium: Visual analytics to support collaborative qualitative coding. In 2017 IEEE Pacific Visualization Symposium (PacificVis) . 220–229
Margaret Drouhard, Nan-Chen Chen, Jina Suh, Rafal Kocielnik, Vanessa Peña-Araya, Keting Cen, Xiangyi Zheng, and Cecilia R. Aragon. 2017 · 2017
Cited alongside, same era.
Accelerated Hierarchical Density Based Clustering. In Data Mining Workshops (ICDMW), 2017 IEEE International Conference on . IEEE, 33–42
Leland McInnes and John Healy. 2017 · 2017
Cited alongside, same era.
Using Machine Learning to Support Qualitative Coding in Social Science: Shifting the Focus to Ambiguity
Nan-Chen Chen, Margaret Drouhard, Rafal Kocielnik, Jina Suh, and Cecilia R. Aragon. 2018a · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Teaching Models to Express Their Uncertainty in Words
Stephanie Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
Later among the works it cites.
Leveraging large language models for multiple choice question answering
Joshua Robinson, Christopher Michael Rytting, and David Wingate. 2022 · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Later among the works it cites.
Problems with Cosine as a Measure of Embedding Similarity for High Frequency Words. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) . Association for Computational Linguistics, Dublin, Ireland, 401–423
Kaitlyn Zhou, Kawin Ethayarajh, Dallas Card, and Dan Jurafsky. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Regularizing and Optimizing LSTM Language Models. In 6th International Conference on Learning Representations, ICLR 2018, Vancouver, BC, Canada, April 30 - May 3, 2018, Conference Track Proceedings . OpenReview.net
Stephen Merity, Nitish Shirish Keskar, and Richard Socher. 2018 · 2018
Cited alongside, same era.
Analyzing Polarization in Social Media: Method and Application to Tweets on 21 Mass Shootings. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 2970–3005
Dorottya Demszky, Nikhil Garg, Rob Voigt, James Zou, Jesse Shapiro, Matthew Gentzkow, and Dan Jurafsky. 2019 · 2019
Cited alongside, same era.
Semantic Concept Spaces: Guided Topic Model Refinement using Word-Embedding Projections
Mennatallah El-Assady, Rebecca Kehlbeck, Christopher Collins, Daniel Keim, and Oliver Deussen. 2019 · 2019
Cited alongside, same era.
BERTopic: Leveraging BERT and c-TF-IDF to create easily interpretable topics
Maarten Grootendorst. 2020 · 2020
Cited alongside, same era.
On the Sentence Embeddings from Pre-trained Language Models. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Online, 9119–9130
Bohan Li, Hao Zhou, Junxian He, Mingxuan Wang, Yiming Yang, and Lei Li. 2020 · 2020
Cited alongside, same era.
The Disagreement Deconvolution: Bringing Machine Learning Performance Metrics In Line With Reality. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems (Yokohama, Japan) (CHI ’21) . Association for Computing Machinery, New York, NY, USA, Article 388, 14 pages
Mitchell L. Gordon, Kaitlyn Zhou, Kayur Patel, Tatsunori Hashimoto, and Michael S. Bernstein. 2021 · 2021
Cited alongside, same era.
Supporting Serendipity: Opportunities and Challenges for Human-AI Collaboration in Qualitative Analysis
Jialun Aaron Jiang, Kandrea Wade, Casey Fiesler, and Jed R. Brubaker. 2021 · 2021
Cited alongside, same era.
Designing Toxic Content Classification for a Diversity of Perspectives. In Seventeenth Symposium on Usable Privacy and Security (SOUPS 2021) . USENIX Association, 299–318
Deepak Kumar, Patrick Gage Kelley, Sunny Consolvo, Joshua Mason, Elie Bursztein, Zakir Durumeric, Kurt Thomas, and Michael Bailey. 2021 · 2021
Cited alongside, same era.
Breaking Out of the Ivory Tower: A Large-Scale Analysis of Patent Citations to HCI Research. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (Hamburg, Germany) (CHI ’23) . Association for Computing Machinery, New York, NY, USA, Article 760, 24 pages
Hancheng Cao, Yujie Lu, Yuting Deng, Daniel Mcfarland, and Michael S. Bernstein. 2023 · 2023
Later among the works it cites.
PaTAT: Human-AI Collaborative Qualitative Coding with Explainable Interactive Rule Synthesis. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (Hamburg, Germany) (CHI ’23) . Association for Computing Machinery, New York, NY, USA, Article 362, 19 pages
Simret Araya Gebreegziabher, Zheng Zhang, Xiaohang Tang, Yihao Meng, Elena L. Glassman, and Toby Jia-Jun Li. 2023 · 2023
Later among the works it cites.
Model Sketching: Centering Concepts in Early-Stage Machine Learning Model Design. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (Hamburg, Germany) (CHI ’23) . Association for Computing Machinery, New York, NY, USA, Article 741, 24 pages
Michelle S. Lam, Zixian Ma, Anne Li, Izequiel Freitas, Dakuo Wang, James A. Landay, and Michael S. Bernstein. 2023 · 2023
Later among the works it cites.
Lost in the Middle: How Language Models Use Long Contexts
Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. 2023 · 2023
Later among the works it cites.
Twitter’s Algorithm: Amplifying Anger, Animosity, and Affective Polarization
Smitha Milli, Micah Carroll, Sashrika Pandey, Yike Wang, and Anca D Dragan. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
TopicGPT: A Prompt-based Topic Modeling Framework
Chau Minh Pham, Alexander Hoyle, Simeng Sun, and Mohit Iyyer. 2023 · 2023
Later among the works it cites.
Whose Opinions Do Language Models Reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto. 2023 · 2023
Later among the works it cites.
Sensecape: Enabling Multilevel Exploration and Sensemaking with Large Language Models. In Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (San Francisco, CA, USA) (UIST ’23) . Association for Computing Machinery, New York, NY, USA, Article 1, 18 pages
Sangho Suh, Bryan Min, Srishti Palani, and Haijun Xia. 2023 · 2023
Later among the works it cites.
Large Language Models Enable Few-Shot Clustering
Vijay Viswanathan, Kiril Gashteovski, Carolin Lawrence, Tongshuang Wu, and Graham Neubig. 2023 · 2023
Later among the works it cites.
Megastudy identifying effective interventions to strengthen Americans’ democratic attitudes
Jan G Voelkel, Michael Stagnaro, James Chu, Sophia Pink, Joseph Mernyk, Chrystal Redekopp, Isaias Ghezae, Matthew Cashman, Dhaval Adjodah, Levi Allen, et al · 2023
Later among the works it cites.
Goal-Driven Explainable Clustering via Language Descriptions
Zihan Wang, Jingbo Shang, and Ruiqi Zhong. 2023 · 2023
Later among the works it cites.
Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding. In Companion Proceedings of the 28th International Conference on Intelligent User Interfaces (Sydney, NSW, Australia) (IUI ’23 Companion) . Association for Computing Machinery, New York, NY, USA, 75–78
Ziang Xiao, Xingdi Yuan, Q. Vera Liao, Rania Abdelghani, and Pierre-Yves Oudeyer. 2023 · 2023
Later among the works it cites.
Embedding Democratic Values into Social Media AIs via Societal Objective Functions
Chenyan Jia, Michelle S. Lam, Minh Chau Mai, Jeffrey T. Hancock, and Michael S. Bernstein. 2024 · 2024
Closest in time.
Can Large Language Models Transform Computational Social Science?
Caleb Ziems, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2024 · 2024
Closest in time.
Is Automated Topic Model Evaluation Broken?: The Incoherence of Coherence
Alexander Hoyle, Pranav Goel, Andrew Hian-Cheong, Denis Peskov, Jordan Boyd-Graber, and Philip Resnik. 2021 · 2033
Closest in time.