Baumann, Joachim, Paul Röttger, Aleksandra Urman, et al. 2025.
“Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation.” arXiv Preprint arXiv:2509.08825, ahead of print.
https://doi.org/10.48550/arXiv.2509.08825.
Bender, Emily M., Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021.
“On the Dangers of Stochastic Parrots: Can Language Models Be Too Big? 🦜.” (New York, NY, USA), FAccT ’21, March 1, 610–23.
https://doi.org/10.1145/3442188.3445922.
Bengio, Yoshua, Réjean Ducharme, and Pascal Vincent. 2000. “A neural probabilistic language model.” Advances in Neural Information Processing Systems 13.
Benoit, Kenneth, Drew Conway, Benjamin E. Lauderdale, Michael Laver, and Slava Mikhaylov. 2016.
“Crowd-Sourced Text Analysis: Reproducible and Agile Production of Political Data.” American Political Science Review 110 (2): 278–95.
https://doi.org/10.1017/S0003055416000058.
Biber, Douglas. 1993.
“Representativeness in Corpus Design.” Literary and Linguistic Computing 8 (4): 243–57.
https://doi.org/10.1093/llc/8.4.243.
Black, Ryan C., Sarah A. Treul, Timothy R. Johnson, and Jerry Goldman. 2011.
“Emotions, Oral Arguments, and Supreme Court Decision Making.” The Journal of Politics 73 (2): 572–81.
https://doi.org/10.1017/s002238161100003x.
Bucher, Martin Juan José, and Marco Martini. 2024.
“Fine-Tuned ‘Small’ LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification.” arXiv Preprint arXiv:2406.08660, ahead of print.
https://doi.org/10.48550/arXiv.2406.08660.
Coase, R. H. 1960.
“The Problem of Social Cost.” The Journal of Law & Economics 56 (4): 837–77.
https://doi.org/10.1086/674872.
Dietrich, Bryce J., Matthew Hayes, and Diana Z. O’Brien. 2019. “Pitch Perfect: Vocal Pitch and the Emotional Intensity of Congressional Speech.” American Political Science Review, Forthcoming.
Egami, Naoki, Musashi Hinck, Brandon M. Stewart, and Hanying Wei. 2024.
“Using Imperfect Surrogates for Downstream Inference: Design-Based Supervised Learning for Social Science Applications of Large Language Models.” Pre-published January 14.
https://doi.org/10.48550/arXiv.2306.04746.
Friedman, Milton, and Anna Jacobson Schwartz. 2008.
A Monetary History of the United States, 1867-1960. Princeton University Press.
https://books.google.com?id=Q7J_EUM3RfoC.
Gilardi, Fabrizio, Meysam Alizadeh, and Maël Kubli. 2023.
“ChatGPT Outperforms Crowd Workers for Text-Annotation Tasks.” Proceedings of the National Academy of Sciences 120 (30): e2305016120.
https://doi.org/10.1073/pnas.2305016120.
Grimmer, Justin. 2010.
“A Bayesian Hierarchical Topic Model for Political Texts: Measuring Expressed Agendas in Senate Press Releases.” Political Analysis 18 (1): 1–35.
https://doi.org/10.1093/pan/mpp034.
Grimmer, Justin, Margaret E. Roberts, and Brandon M. Stewart. 2022.
Text as Data: A New Framework for Machine Learning and the Social Sciences. Princeton University Press.
https://books.google.com?id=dL40EAAAQBAJ.
Grimmer, Justin, and Brandon M. Stewart. 2013. “Text as Data: The Promise and Pitfalls of Automatic Content Analysis Methods for Political Texts.” Political Analysis 21 (3): 267–97.
Lasswell, Harold D. 1927. “The Theory of Political Propaganda.” American Political Science Review 21 (03): 627–31.
Liu, Amy H. 2022.
“Pronoun Usage as a Measure of Power Personalization: A General Theory with Evidence from the Chinese-Speaking World.” British Journal of Political Science 52 (3, 3): 1258–75.
https://doi.org/10.1017/S0007123421000181.
Ollion, Étienne, Rubing Shen, Ana Macanovic, and Arnault Chatelain. 2024.
“The Dangers of Using Proprietary LLMs for Research.” Nature Machine Intelligence 6 (1): 4–5.
https://doi.org/10.1038/s42256-023-00783-6.
Pangakis, Nicholas, Samuel Wolken, and Neil Fasching. 2023.
“Automated Annotation with Generative AI Requires Validation.” arXiv Preprint arXiv:2306.00176, ahead of print.
https://doi.org/10.48550/arXiv.2306.00176.
Pennebaker, James W. 2017.
“Mind Mapping: Using Everyday Language to Explore Social & Psychological Processes.” Procedia Computer Science, Data analytics summit II; structuring the UNSTRUCTURED: The missing element of analytics, 14-16 december 2015, harrisburg, USA, vol. 118 (January): 100–107.
https://doi.org/10.1016/j.procs.2017.11.150.
Pennington, Jeffrey, Richard Socher, and Christopher Manning. 2014.
“GloVe: Global Vectors for Word Representation.” Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) (Doha, Qatar), October, 1532–43.
https://doi.org/10.3115/v1/D14-1162.
Stoltz, Dustin S., Marshall A. Taylor, and Sanuj Kumar. 2026.
“Selecting Language Models for Social Science: Start Small, Start Open, and Validate.” arXiv Preprint arXiv:2601.10926, ahead of print.
https://doi.org/10.48550/arXiv.2601.10926.
Stone, Philip J., Robert F. Bales, J. Zvi Namenwirth, and Daniel M. Ogilvie. 1962.
“The General Inquirer: A Computer System for Content Analysis and Retrieval Based on the Sentence as a Unit of Information.” Behavioral Science 7 (4): 484–98.
https://doi.org/10.1002/bs.3830070412.
Su, Yu-Sung, Yanqin Ruan, Siyu Sun, and Yu-Tzung Chang. 2020.
“A Pattern Recognition Framework for Detecting Changes in Chinese Internet Management System.” Journal of Social Computing 1 (1): 28–39.
https://doi.org/10.23919/JSC.2020.0004.
Törnberg, Petter. 2025.
“Large Language Models Outperform Expert Coders and Supervised Classifiers at Annotating Political Social Media Messages.” Social Science Computer Review 43 (6): 1181–95.
https://doi.org/10.1177/08944393241286471.
Vaswani, Ashish, Noam Shazeer, Niki Parmar, et al. 2017. “Attention Is All You Need.” Advances in Neural Information Processing Systems 30.
Zhang, Han, and Jennifer Pan. 2019.
“CASM: A Deep-Learning Approach for Identifying Collective Action Events with Text and Image Data from Social Media.” Sociological Methodology 49 (1): 1–57.
https://doi.org/10.1177/0081175019860244.
Zipf, George Kingsley. (1949) 2016. Human Behavior and the Principle of Least Effort: An Introduction to Human Ecology. Addison-Wesley Press. Reprint, Ravenio Books.
Social Media