References
The bibliography below includes only works assigned or cited in at least one of the four syllabi.
Course markers:
🟥 Data Literacy
🟨 Working Efficiently with Data and Code
🟩 Computational Social Science
🟪 Generative AI for Applied Research
Bibliography
Abdurahman, Suhaib, Alireza Salkhordeh Ziabari, Alexander K. Moore, Daniel M. Bartels, and Morteza Dehghani. 2025. “A Primer for Evaluating Large Language Models in Social-Science Research.” Advances in Methods and Practices in Psychological Science 8 (2): 1–25. https://doi.org/10.1177/25152459251325174.
Abraham, Louis, Charles Arnal, and Antoine Marie. 2025. “Prompt Selection Matters: Enhancing Text Annotations for Social Sciences with Large Language Models.” Journal of Computational Social Science 8: 73. https://doi.org/10.1007/s42001-025-00388-6.
Abramson, Corey M. et al. 2026. “Qualitative Research in an Era of Artificial Intelligence: A Pragmatic Approach to Data Analysis, Workflow, and Computation.” Annual Review of Sociology 52. https://doi.org/10.1146/annurev-soc-011824-104836.
Adcock, Robert, and David Collier. 2001. “Measurement Validity: A Shared Standard for Qualitative and Quantitative Research.” American Political Science Review 95 (3): 529–46. https://doi.org/10.1017/S0003055401003100.
Alammar, Jay, and Maarten Grootendorst. 2024. Hands-on Large Language Models. O’Reilly Media.
Alvarado-Mena, Edwin. 2026a. “How to Organize Your Computer for Data Work.” July 28. https://alvaradocss.com/posts/organize-data-projects/.
Alvarado-Mena, Edwin. 2026b. “Rendering Professional Documents with Quarto.” July 28. https://alvaradocss.com/posts/professional-documents-quarto/.
Amershi, Saleema, Dan Weld, Mihaela Vorvoreanu, et al. 2019. “Guidelines for Human-AI Interaction.” Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems, 1–13. https://doi.org/10.1145/3290605.3300233.
Angrist, Joshua D, and Jörn-Steffen Pischke. 2009. Mostly Harmless Econometrics: An Empiricist’s Companion. Princeton university press.
Angrist, Joshua D, and Jörn-Steffen Pischke. 2014. Mastering’metrics: The Path from Cause to Effect. Princeton University Press.
Ashwin, Julian, Aditya Chhabra, and Vijayendra Rao. 2026. “Using Large Language Models for Qualitative Analysis Can Introduce Serious Bias.” Sociological Methods & Research 55 (3): 795–839. https://doi.org/10.1177/00491241251338246.
Athey, Susan, and Guido W Imbens. 2019. “Machine Learning Methods That Economists Should Know About.” Annual Review of Economics 11 (1): 685–725. https://doi.org/10.1146/annurev-economics-080217-053433.
Atreja, Shubham, Joshua Ashkinaze, Lingyao Li, Julia Mendelsohn, and Libby Hemphill. 2025. “What’s in a Prompt?: A Large-Scale Experiment to Assess the Impact of Prompt Design on the Compliance and Accuracy of LLM-Generated Text Annotations.” Proceedings of the International AAAI Conference on Web and Social Media 19 (1). https://doi.org/10.1609/icwsm.v19i1.35807.
Atteveldt, Wouter van, Damian Trilling, and Carlos Arcíla Calderón. 2022. Computational Analysis of Communication. Wiley Blackwell.
Bail, Christopher A. 2012. “The Fringe Effect: Civil Society Organizations and the Evolution of Media Discourse about Islam Since the September 11th Attacks.” American Sociological Review 77 (6): 855–79.
Bail, Christopher A. 2024. “Can Generative AI Improve Social Science?” Proceedings of the National Academy of Sciences 121 (21): e2314021121. https://doi.org/10.1073/pnas.2314021121.
Barrie, Christopher, Lisa P. Argyle, James Bisbee, et al. 2026. AI and Research Methods. https://doi.org/10.33774/apsa-2026-h59kk.
Beaulieu, Alan. 2020. Learning SQL: Generate, Manipulate, and Retrieve Data. O’Reilly Media.
Bender, Emily M., Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021. “On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?” Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency, 610–23. https://doi.org/10.1145/3442188.3445922.
Bergstrom, Carl T., and Jevin D. West. 2020. Calling Bullshit: The Art of Skepticism in a Data-Driven World. Random House.
Bisbee, James, Joshua D Clinton, Cassy Dorff, Brenton Kenkel, and Jennifer M Larson. 2024. “Synthetic Replacements for Human Survey Data? The Perils of Large Language Models.” Political Analysis 32 (4): 401–16. https://doi.org/10.1017/pan.2024.5.
Bisbee, James, and Arthur Spirling. 2025. “What to Do When Humans Are No Longer the Gold Standard: Large Language Models, State of the Art and Robustness.” Unpublished manuscript.
Bommasani, Rishi et al. 2021. On the Opportunities and Risks of Foundation Models. https://arxiv.org/abs/2108.07258.
Bond, Robert M., Christopher J. Fariss, Jason J. Jones, et al. 2012. “A 61-Million-Person Experiment in Social Influence and Political Mobilization.” Nature 489: 295–98. https://doi.org/10.1038/nature11421.
Borgatti, Stephen P., Martin G. Everett, and Jeffrey C. Johnson. 2018. Analyzing Social Networks. 2nd ed. SAGE.
Borgatti, Stephen P., Ajay Mehra, Daniel J. Brass, and Giuseppe Labianca. 2009. “Network Analysis in the Social Sciences.” Science 323 (5916): 892–95. https://doi.org/10.1126/science.1165821.
Borgman, Christine L. 2015. Big Data, Little Data, No Data: Scholarship in the Networked World. MIT Press.
boyd, danah, and Kate Crawford. 2012. “Critical Questions for Big Data.” Information, Communication & Society 15 (5): 662–79. https://doi.org/10.1080/1369118X.2012.678878.
Brach, William, Kristián Košťál, and Michal Ries. 2025. “The Effectiveness of Large Language Models in Transforming Unstructured Text to Standardized Formats.” IEEE Access 13: 91808–25. https://doi.org/10.1109/ACCESS.2025.3573030.
Brandt, Patrick T, Sultan Alsarra, Vito D’Orazio, et al. 2026. “Extractive Versus Generative Language Models for Political Conflict Text Classification.” Political Analysis 34 (3): 344–72.
Broman, Karl W, and Kara H Woo. 2018. “Data Organization in Spreadsheets.” The American Statistician 72 (1): 2–10. https://doi.org/10.1080/00031305.2017.1375989.
Brown, Tom B., Benjamin Mann, Nick Ryder, et al. 2020. “Language Models Are Few-Shot Learners.” Advances in Neural Information Processing Systems 33: 1877–901. https://proceedings.neurips.cc/paper/2020/hash/1457c0d6bfcb4967418bfb8ac142f64a-Abstract.html.
Bryan, Jennifer. 2018. “Excuse Me, Do You Have a Moment to Talk about Version Control?” The American Statistician 72 (1): 20–27. https://doi.org/10.1080/00031305.2017.1399928.
Buolamwini, Joy, and Timnit Gebru. 2018. “Gender Shades: Intersectional Accuracy Disparities in Commercial Gender Classification.” Proceedings of the 1st Conference on Fairness, Accountability and Transparency, Proceedings of machine learning research, vol. 81: 77–91. https://proceedings.mlr.press/v81/buolamwini18a.html.
Cairo, Alberto. 2016. The Truthful Art: Data, Charts, and Maps for Communication. New Riders.
Cairo, Alberto. 2019. How Charts Lie: Getting Smarter about Visual Information. W. W. Norton & Company.
Card, Dallas, Serina Chang, Chris Becker, et al. 2022. “Computational Analysis of 140 Years of US Political Speeches Reveals More Positive but Increasingly Polarized Framing of Immigration.” Proceedings of the National Academy of Sciences 119 (31): e2120510119. https://doi.org/10.1073/pnas.2120510119.
Carlini, Nicholas, Florian Tramèr, Eric Wallace, et al. 2021. “Extracting Training Data from Large Language Models.” 30th USENIX Security Symposium (USENIX Security 21), 2633–50. https://www.usenix.org/conference/usenixsecurity21/presentation/carlini-extracting.
Chacon, Scott, and Ben Straub. 2014. Pro Git. 2nd ed. Apress.
Chae, Youngjin, and Thomas Davidson. 2026. “Large Language Models for Text Classification: From Zero-Shot Learning to Instruction-Tuning.” Sociological Methods & Research 55 (2): 501–67.
Chang, Yupeng, Xu Wang, Jindong Wang, et al. 2024. “A Survey on Evaluation of Large Language Models.” ACM Transactions on Intelligent Systems and Technology 15 (3): 1–45. https://doi.org/10.1145/3641289.
Cirone, Alexandra, and Arthur Spirling. 2021. “Turning History into Data: Data Collection, Measurement, and Inference in HPE.” Journal of Historical Political Economy 1 (1): 127–54.
Cleveland, William S., and Robert McGill. 1984. “Graphical Perception: Theory, Experimentation, and Application to the Development of Graphical Methods.” Journal of the American Statistical Association 79 (387): 531–54. https://doi.org/10.1080/01621459.1984.10478080.
Collier, David, Jody LaPorte, and Jason Seawright. 2012. “Putting Typologies to Work: Concept Formation, Measurement, and Analytic Rigor.” Political Research Quarterly 65 (1): 217–32. https://doi.org/10.1177/1065912912437162.
Cunningham, Scott. 2021. Causal Inference: The Mixtape. Yale University Press.
D’Ignazio, Catherine, and Lauren F. Klein. 2020. Data Feminism. MIT Press.
De Mesquita, Ethan Bueno, and Anthony Fowler. 2021. Thinking Clearly with Data: A Guide to Quantitative Reasoning and Analysis. Princeton University Press.
Denny, Matthew J, and Arthur Spirling. 2018. “Text Preprocessing for Unsupervised Learning: Why It Matters, When It Misleads, and What to Do about It.” Political Analysis 26 (2): 168–89. https://doi.org/10.1017/pan.2017.44.
Diez, David M., Christopher D. Barr, and Mine Çetinkaya-Rundel. 2019. OpenIntro Statistics. 4th ed. OpenIntro, Inc. https://www.openintro.org/book/os/.
Dong, Qingxiu, Lei Li, Damai Dai, et al. 2024. “A Survey on in-Context Learning.” Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, 1107–28.
Donoho, David. 2017. “50 Years of Data Science.” Journal of Computational and Graphical Statistics 26 (4): 745–66. https://doi.org/10.1080/10618600.2017.1384734.
Duflo, Esther, Rachel Glennerster, and Michael Kremer. 2007. “Using Randomization in Development Economics Research: A Toolkit.” Handbook of Development Economics 4: 3895–962. https://doi.org/10.1016/S1573-4471(07)04061-2.
Dunivin, Zackary Okun, Mobina Noori, Seth Frey, and Curtis Atkinson. 2026. Self-Reflection in Automated Qualitative Coding: Improving Text Annotation Through Secondary LLM Critique. https://arxiv.org/abs/2601.09905.
Edelmann, Achim, Tom Wolff, Danielle Montagne, and Christopher A. Bail. 2020. “Computational Social Science and Sociology.” Annual Review of Sociology 46: 61–81. https://doi.org/10.1146/annurev-soc-121919-054621.
Egami, Naoki, Christian J. Fong, Justin Grimmer, Margaret E. Roberts, and Brandon M. Stewart. 2022. “How to Make Causal Inferences Using Texts.” Science Advances 8 (42): eabg2652. https://doi.org/10.1126/sciadv.abg2652.
Eisenstein, Jacob. 2019. Introduction to Natural Language Processing. MIT Press.
Enamorado, Ted, Benjamin Fifield, and Kosuke Imai. 2019. “Using a Probabilistic Model to Assist Merging of Large-Scale Administrative Records.” American Political Science Review 113 (2): 353–71. https://doi.org/10.1017/S0003055418000783.
Epstein, Joshua M. 2008. “Why Model?” Journal of Artificial Societies and Social Simulation 11 (4): 12. https://www.jasss.org/11/4/12.html.
Es, Shahul, Jithin James, Luis Espinosa-Anke, and Steven Schockaert. 2024. “RAGAs: Automated Evaluation of Retrieval Augmented Generation.” Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations, 150–58. https://doi.org/10.18653/v1/2024.eacl-demo.16.
Feuerriegel, Stefan et al. 2026. “A Reporting Checklist for Large Language Models in Behavioural Science.” Nature Human Behaviour 10 (7): 1182–86. https://doi.org/10.1038/s41562-026-02492-7.
Fotheringham, A. Stewart, and David W. S. Wong. 1991. “The Modifiable Areal Unit Problem in Multivariate Statistical Analysis.” Environment and Planning A 23 (7): 1025–44. https://doi.org/10.1068/a231025.
Freedman, David A. 1991. “Statistical Models and Shoe Leather.” Sociological Methodology 21: 291–313. https://doi.org/10.2307/270939.
Futschek, Gerald. 2006. “Algorithmic Thinking: The Key for Understanding Computer Science.” International Conference on Informatics in Secondary Schools-Evolution and Perspectives, 159–68.
Gallegos, Isabel O., Ryan A. Rossi, Joe Barrow, et al. 2024. “Bias and Fairness in Large Language Models: A Survey.” Computational Linguistics 50 (3): 1097–179. https://doi.org/10.1162/coli_a_00524.
Gallifant, Jack, Majid Afshar, Saleem Ameen, et al. 2025. “The TRIPOD-LLM Reporting Guideline for Studies Using Large Language Models.” Nature Medicine 31 (1): 60–69. https://doi.org/10.1038/s41591-024-03425-5.
Gandrud, Christopher. 2020. Reproducible Research with r and RStudio. 3rd ed. Chapman; Hall/CRC.
Gebru, Timnit, Jamie Morgenstern, Briana Vecchione, et al. 2021. “Datasheets for Datasets.” Communications of the ACM 64 (12): 86–92. https://doi.org/10.1145/3458723.
Gelman, Andrew, Jennifer Hill, and Aki Vehtari. 2020. Regression and Other Stories. Cambridge University Press.
Gelman, Andrew, and Guido Imbens. 2013. Why Ask Why? Forward Causal Inference and Reverse Causal Questions. National Bureau of Economic Research.
Gentzkow, Matthew, Bryan Kelly, and Matt Taddy. 2019. “Text as Data.” Journal of Economic Literature 57 (3): 535–74. https://doi.org/10.1257/jel.20181020.
Gerring, John. 2012. “Mere Description.” British Journal of Political Science 42 (4): 721–46. https://doi.org/10.1017/S0007123412000130.
Gilardi, Fabrizio, Meysam Alizadeh, and Maël Kubli. 2023. “ChatGPT Outperforms Crowd Workers for Text-Annotation Tasks.” Proceedings of the National Academy of Sciences 120 (30): e2305016120. https://doi.org/10.1073/pnas.2305016120.
Glaeser, Edward L, Scott Duke Kominers, Michael Luca, and Nikhil Naik. 2018. “Big Data and Big Cities: The Promises and Limitations of Improved Measures of Urban Life.” Economic Inquiry 56 (1): 114–37.
Goldberg, Amir, and Sarah K Stein. 2018. “Beyond Social Contagion: Associative Diffusion and the Emergence of Cultural Variation.” American Sociological Review 83 (5): 897–932. https://doi.org/10.1177/0003122418797576.
Google. n.d.-a. “Google Python Style Guide.” Accessed September 2, 2026. https://google.github.io/styleguide/pyguide.html.
Google. n.d.-b. “Google’s r Style Guide.” Accessed September 2, 2026. https://google.github.io/styleguide/Rguide.html.
Granovetter, Mark. 1985. “Economic Action and Social Structure: The Problem of Embeddedness.” American Journal of Sociology 91 (3): 481–510. https://doi.org/10.1086/228311.
Greenland, Sander, Stephen J. Senn, Kenneth J. Rothman, et al. 2016. “Statistical Tests, p Values, Confidence Intervals, and Power: A Guide to Misinterpretations.” European Journal of Epidemiology 31: 337–50. https://doi.org/10.1007/s10654-016-0149-3.
Grimmer, Justin. 2015. “We Are All Social Scientists Now: How Big Data, Machine Learning, and Causal Inference Work Together.” PS: Political Science & Politics 48 (1): 80–83. https://doi.org/10.1017/S1049096514001784.
Grimmer, Justin, Margaret E Roberts, and Brandon M Stewart. 2022. Text as Data: A New Framework for Machine Learning and the Social Sciences. Princeton University Press.
Grimmer, Justin, and Brandon M. Stewart. 2013. “Text as Data: The Promise and Pitfalls of Automatic Content Analysis Methods for Political Texts.” Political Analysis 21 (3): 267–97. https://doi.org/10.1093/pan/mps028.
Hand, D. J. 1996. “Statistics and the Theory of Measurement.” Journal of the Royal Statistical Society: Series A (Statistics in Society) 159 (3): 445–73. https://doi.org/10.2307/2983326.
Henry, Adam Douglas. 2011. “Ideology, Power, and the Structure of Policy Networks.” Policy Studies Journal 39 (3): 361–83.
Heseltine, Michael, and Bernhard Clemm von Hohenberg. 2024. “Large Language Models as a Substitute for Human Experts in Annotating Political Text.” Research & Politics 11 (1): 20531680241236239.
Hofman, Jake M, Duncan J Watts, Susan Athey, et al. 2021. “Integrating Explanation and Prediction in Computational Social Science.” Nature 595 (7866): 181–88. https://doi.org/10.1038/s41586-021-03659-0.
Hogan, Aidan, Eva Blomqvist, Michael Cochez, et al. 2021. “Knowledge Graphs.” ACM Computing Surveys (Csur) 54 (4): 1–37. https://doi.org/10.1145/3447772.
Huntington-Klein, Nick. 2022. The Effect: An Introduction to Research Design and Causality. CRC Press.
IBM. n.d. “What Is Data Modeling?” Accessed August 21, 2026. https://www.ibm.com/think/topics/data-modeling.
Imai, Kosuke, and Kentaro Nakamura. 2026. “Causal Inference with Generative Artificial Intelligence: Application to Texts as Treatments.” Journal of the American Statistical Association, nos. just-accepted: 1–27.
Imai, Kosuke, and Nora Webb Williams. 2022. Quantitative Social Science: An Introduction in Tidyverse. Princeton University Press.
Ioannidis, John P. A. 2005. “Why Most Published Research Findings Are False.” PLOS Medicine 2 (8): e124. https://doi.org/10.1371/journal.pmed.0020124.
Irving, Damien, Kate Hertweck, Luke Johnston, Joel Ostblom, Charlotte Wickham, and Greg Wilson. 2021. Research Software Engineering with Python. CRC Press.
James, Gareth, Daniela Witten, Trevor Hastie, Robert Tibshirani, and Jonathan Taylor. 2023. An Introduction to Statistical Learning: With Applications in Python. Springer.
Janssens, Jeroen. 2021. Data Science at the Command Line. 2nd ed. O’Reilly Media.
Jurafsky, Daniel, and James H. Martin. 2026. Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition with Language Models. 3rd ed. https://web.stanford.edu/~jurafsky/slp3/.
Kapoor, Sayash, and Arvind Narayanan. 2023. “Leakage and the Reproducibility Crisis in Machine-Learning-Based Science.” Patterns 4 (9): 100804. https://doi.org/10.1016/j.patter.2023.100804.
Keele, Luke J, and Rocio Titiunik. 2015. “Geographic Boundaries as Regression Discontinuities.” Political Analysis 23 (1): 127–55. https://doi.org/10.1093/pan/mpu014.
Kim, Hannah, Kushan Mitra, Rafael Li Chen, Sajjadur Rahman, and Dan Zhang. 2024. “Meganno+: A Human-Llm Collaborative Annotation System.” Proceedings of the 18th Conference of the European Chapter of the Association for Computational Linguistics: System Demonstrations, 168–76.
King, Gary. 2013. Big Data Is Not about the Data! Presentation at the Golden Seeds Innovation Summit, New York, NY. https://gking.harvard.edu/files/gking/files/evbase-gs.pdf.
King, Gary, Robert O Keohane, and Sidney Verba. 2021. Designing Social Inquiry: Scientific Inference in Qualitative Research. Princeton University Press.
Kitchin, Rob. 2014. The Data Revolution: Big Data, Open Data, Data Infrastructures and Their Consequences. SAGE Publications.
Kitzes, Justin, Daniel Turek, and Fatma Deniz, eds. 2018. The Practice of Reproducible Research: Case Studies and Lessons from the Data-Intensive Sciences. University of California Press.
Landers, Richard N., Robert C. Brusso, Katelyn J. Cavanaugh, and Andrew B. Collmus. 2016. “A Primer on Theory-Driven Web Scraping: Automatic Extraction of Big Data from the Internet for Use in Psychological Research.” Psychological Methods 21 (4): 475–92. https://doi.org/10.1037/met0000081.
Lazer, David MJ, Alex Pentland, Duncan J Watts, et al. 2020. “Computational Social Science: Obstacles and Opportunities.” Science 369 (6507): 1060–62. https://doi.org/10.1126/science.aaz8170.
Lazer, David, Eszter Hargittai, Deen Freelon, et al. 2021. “Meaningful Measures of Human Society in the Twenty-First Century.” Nature 595 (7866): 189–96. https://doi.org/10.1038/s41586-021-03660-7.
Lazer, David, Ryan Kennedy, Gary King, and Alessandro Vespignani. 2014. “The Parable of Google Flu: Traps in Big Data Analysis.” Science 343 (6176): 1203–5. https://doi.org/10.1126/science.1248506.
Lazer, David, Alex Pentland, Lada Adamic, et al. 2009. “Computational Social Science.” Science 323 (5915): 721–23. https://doi.org/10.1126/science.1167742.
Lewis, Patrick, Ethan Perez, Aleksandra Piktus, et al. 2020. “Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.” Advances in Neural Information Processing Systems 33: 9459–74.
Li, Junyi, Xiaoxue Cheng, Xin Zhao, Jian-Yun Nie, and Ji-Rong Wen. 2023. “HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models.” Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, 6449–64. https://doi.org/10.18653/v1/2023.emnlp-main.397.
Licht, Hauke, Rupak Sarkar, Patrick Y. Wu, et al. 2025. “Measuring Scalar Constructs in Social Science with LLMs.” Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, 32144–71.
Llaudet, Elena, and Kosuke Imai. 2023. Data Analysis for Social Science: A Friendly and Practical Introduction. Princeton University Press.
Lovelace, Robin, Jakub Nowosad, and Jannes Muenchow. 2019. Geocomputation with r. Chapman; Hall/CRC.
Lundberg, Ian, Rebecca Johnson, and Brandon M Stewart. 2021. “What Is Your Estimand? Defining the Target Quantity Connects Statistical Evidence to Theory.” American Sociological Review 86 (3): 532–65. https://doi.org/10.1177/00031224211004187.
Luscombe, Alex, Kevin Dick, and Kevin Walby. 2022. “Algorithmic Thinking in the Public Interest: Navigating Technical, Legal, and Ethical Hurdles to Web Scraping in the Social Sciences.” Quality & Quantity 56 (3): 1023–44. https://doi.org/10.1007/s11135-021-01164-0.
Ma, Jing. 2025. “Causal Inference with Large Language Model: A Survey.” Findings of the Association for Computational Linguistics: NAACL 2025, 5901–13.
Macy, Michael W., and Robert Willer. 2002. “From Factors to Actors: Computational Sociology and Agent-Based Modeling.” Annual Review of Sociology 28: 143–66. https://doi.org/10.1146/annurev.soc.28.110601.141117.
Manning, Christopher D., Prabhakar Raghavan, and Hinrich Schütze. 2008. Introduction to Information Retrieval. Cambridge University Press.
Marwick, Ben, Carl Boettiger, and Lincoln Mullen. 2018. “Packaging Data Analytical Work Reproducibly Using r (and Friends).” The American Statistician 72 (1): 80–88. https://doi.org/10.1080/00031305.2017.1375986.
McVeety, Sam, and Amir Hormati. 2026. “Introducing the Open Knowledge Format.” Google Cloud, June 12. https://cloud.google.com/blog/products/data-analytics/how-the-open-knowledge-format-can-improve-data-sharing.
Meng, Xiao-Li. 2018. “Statistical Paradises and Paradoxes in Big Data (i): Law of Large Populations, Big Data Paradox, and the 2016 US Presidential Election.” The Annals of Applied Statistics 12 (2): 685–726. https://doi.org/10.1214/18-AOAS1161SF.
Mitchell, Margaret, Simone Wu, Andrew Zaldivar, et al. 2019. “Model Cards for Model Reporting.” Proceedings of the Conference on Fairness, Accountability, and Transparency, 220–29. https://doi.org/10.1145/3287560.3287596.
Molina, Mario, and Filiz Garip. 2019. “Machine Learning for Sociology.” Annual Review of Sociology 45 (1): 27–45. https://doi.org/10.1146/annurev-soc-073117-041106.
Moreau, David, Kristina Wiebels, and Carl Boettiger. 2023. “Containers for Computational Reproducibility.” Nature Reviews Methods Primers 3: 50. https://doi.org/10.1038/s43586-023-00236-9.
Mullainathan, Sendhil, and Jann Spiess. 2017. “Machine Learning: An Applied Econometric Approach.” Journal of Economic Perspectives 31 (2): 87–106. https://doi.org/10.1257/jep.31.2.87.
Munzert, Simon, Christian Rubba, Peter Meißner, and Dominic Nyhuis. 2015. Automated Data Collection with r: A Practical Guide to Web Scraping and Text Mining. Wiley.
Narayanan, Pavan Kumar. 2024. “Getting Started with Data Validation Using Pydantic and Pandera.” In Data Engineering for Machine Learning Pipelines: From Python Libraries to ML Pipelines and Cloud Platforms. Springer.
National Academies of Sciences, Engineering, and Medicine. 2019. Reproducibility and Replicability in Science. The National Academies Press. https://doi.org/10.17226/25303.
Nelson, Laura K. 2020. “Computational Grounded Theory: A Methodological Framework.” Sociological Methods & Research 49 (1): 3–42. https://doi.org/10.1177/0049124117729703.
Niu, Cheng, Yuanhao Wu, Juno Zhu, et al. 2024. “RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models.” Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers).
Noy, Natalya F., and Deborah L. McGuinness. 2001. Ontology Development 101: A Guide to Creating Your First Ontology. Stanford Knowledge Systems Laboratory.
O’Sullivan, David, and David J. Unwin. 2010. Geographic Information Analysis. 2nd ed. Wiley.
OpenAI. 2026. Developer Quickstart. Https://developers.openai.com/api/docs/quickstart.
OpenAI. n.d.-a. “Function Calling and Tools.” Accessed September 2, 2026. https://developers.openai.com/api/docs/guides/function-calling.
OpenAI. n.d.-b. “Structured Outputs.” Accessed September 2, 2026. https://developers.openai.com/api/docs/guides/structured-outputs.
Paci, Simone. 2026. “Is Research Safe in the AI Revolution? LLMs Fall Short of Expert Benchmarks in Scientific Evidence Evaluation.” Chinese Political Science Review, 1–23.
Palmer, Alexis, and Arthur Spirling. 2023. “Large Language Models Can Argue in Convincing Ways about Politics, but Humans Dislike AI Authors: Implications for Governance.” Political Science 75 (3): 281–91.
Pangakis, Nick, and Sam Wolken. 2025. “Keeping Humans in the Loop: Human-Centered Automated Annotation with Generative AI.” Proceedings of the International AAAI Conference on Web and Social Media 19: 1471–92. https://doi.org/10.1609/icwsm.v19i1.35883.
Park, Joon Sung, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2023. “Generative Agents: Interactive Simulacra of Human Behavior.” Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology. https://doi.org/10.1145/3586183.3606763.
Pearl, Judea, and Dana Mackenzie. 2018. The Book of Why: The New Science of Cause and Effect. Basic Books.
Peng, Ji-Lun, Sijia Cheng, Egil Diau, et al. 2024. A Survey of Useful LLM Evaluation. https://doi.org/10.48550/arXiv.2406.00936.
Phoenix, James, and Mike Taylor. 2024. Prompt Engineering for Generative AI. " O’Reilly Media, Inc.".
Polak, Maciej P., and Dane Morgan. 2024. “Extracting Accurate Materials Data from Research Papers with Conversational Language Models and Prompt Engineering.” Nature Communications 15: 1569. https://doi.org/10.1038/s41467-024-45914-8.
Porsdam Mann, Sebastian, Anuraag A. Vazirani, Mateo Aboy, et al. 2024. “Guidelines for Ethical Use and Acknowledgement of Large Language Models in Academic Writing.” Nature Machine Intelligence 6: 1272–74. https://doi.org/10.1038/s42256-024-00922-7.
Posit. n.d. “Get Started with Quarto.” Accessed September 4, 2026. https://quarto.org/docs/get-started/.
Railsback, Steven F., and Volker Grimm. 2019. Agent-Based and Individual-Based Modeling: A Practical Introduction. 2nd ed. Princeton University Press.
Reimers, Nils, and Iryna Gurevych. 2019. “Sentence-BERT: Sentence Embeddings Using Siamese BERT-Networks.” Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing, 3982–92. https://doi.org/10.18653/v1/D19-1410.
Reinhart, Alex. 2015. Statistics Done Wrong: The Woefully Complete Guide. No Starch Press.
Roberts, Margaret E., Brandon M. Stewart, Dustin Tingley, et al. 2014. “Structural Topic Models for Open-Ended Survey Responses.” American Journal of Political Science 58 (4): 1064–82. https://doi.org/10.1111/ajps.12103.
Rodriguez, Pedro L, and Arthur Spirling. 2022. “Word Embeddings: What Works, What Doesn’t, and How to Tell the Difference for Applied Research.” The Journal of Politics 84 (1): 101–15.
Rogers, Simon, and Mark Girolami. 2016. A First Course in Machine Learning. CRC press.
Rule, Adam, Amanda Birmingham, Cristal Zuniga, et al. 2019. “Ten Simple Rules for Writing and Sharing Computational Analyses in Jupyter Notebooks.” PLOS Computational Biology 15 (7): e1007007. https://doi.org/10.1371/journal.pcbi.1007007.
Russell, Stuart J., and Peter Norvig. 2020. Artificial Intelligence: A Modern Approach. 4th ed. Pearson.
Salganik, Matthew J. 2019. Bit by Bit: Social Research in the Digital Age. Princeton University Press.
Samii, Cyrus. 2016. “Causal Empiricism in Quantitative Research.” The Journal of Politics 78 (3): 941–55. https://doi.org/10.1086/686690.
Sandve, Geir Kjetil, Anton Nekrutenko, James Taylor, and Eivind Hovig. 2013. “Ten Simple Rules for Reproducible Computational Research.” PLOS Computational Biology 9 (10): e1003285. https://doi.org/10.1371/journal.pcbi.1003285.
Schoon, Eric W, David Melamed, and Ronald L Breiger. 2024. Regression Inside Out. Cambridge University Press.
Schroeder, Hope, Marianne Aubin Le Quéré, Casey Randazzo, David Mimno, and Sarita Schoenebeck. 2025. “Large Language Models in Qualitative Research: Uses, Tensions, and Intentions.” Proceedings of the 2025 CHI Conference on Human Factors in Computing Systems, 1–17. https://doi.org/10.1145/3706598.3713120.
Schroeder, Hope, Deb Roy, and Jad Kabbara. 2025. “Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks.” Findings of the Association for Computational Linguistics: ACL 2025, 25771–95. https://doi.org/10.18653/v1/2025.findings-acl.1323.
Schulhoff, Sander, Michael Ilie, Nishant Balepur, et al. 2024. The Prompt Report: A Systematic Survey of Prompting Techniques. https://doi.org/10.48550/arXiv.2406.06608.
Shimizu, Cogan, and Pascal Hitzler. 2025. “Accelerating Knowledge Graph and Ontology Engineering with Large Language Models.” Journal of Web Semantics 85: 100862. https://doi.org/10.1016/j.websem.2025.100862.
Smaldino, Paul. 2023. Modeling Social Behavior: Mathematical and Agent-Based Models of Social Dynamics and Cultural Evolution. Princeton University Press.
Spiegelhalter, David. 2019. The Art of Statistics: How to Learn from Data. Basic Books.
Stevens, Stanley Smith. 1946. “On the Theory of Scales of Measurement.” Science 103 (2684): 677–80. https://doi.org/10.1126/science.103.2684.677.
The Carpentries. n.d.-a. “Instructor Training.” The Carpentries. Accessed September 4, 2026. https://preview.carpentries.org/instructor-training/.
The Carpentries. n.d.-b. “Programming Lessons.” Accessed September 5, 2026. https://software-carpentry.org/lessons/.
The Carpentries. n.d.-c. “The Unix Shell.” Accessed September 2, 2026. https://swcarpentry.github.io/shell-novice/.
The Carpentries. n.d.-d. “Version Control with Git.” Accessed September 2, 2026. https://swcarpentry.github.io/git-novice/.
The tidyverse team. n.d. “The Tidyverse Style Guide.” Accessed September 2, 2026. https://style.tidyverse.org/.
Törnberg, Petter. 2025. “Large Language Models Outperform Expert Coders and Supervised Classifiers at Annotating Political Social Media Messages.” Social Science Computer Review 43 (6). https://doi.org/10.1177/08944393241286471.
Tsaneva, Stefani, Danilo Dessì, Francesco Osborne, and Marta Sabou. 2025. “Knowledge Graph Validation by Integrating LLMs and Human-in-the-Loop.” Information Processing & Management 62 (5): 104145. https://doi.org/10.1016/j.ipm.2025.104145.
Tudorache, Tania. 2020. “Ontology Engineering: Current State, Challenges, and Future Directions.” Semantic Web 11 (1): 125–38. https://doi.org/10.3233/SW-190382.
Vaswani, Ashish, Noam Shazeer, Niki Parmar, et al. 2017. “Attention Is All You Need.” Advances in Neural Information Processing Systems 30: 5998–6008. https://proceedings.neurips.cc/paper/7181-attention-is-all-you-need.
Veselovsky, Veniamin, Manoel Horta Ribeiro, Akhil Arora, Martin Josifoski, Ashton Anderson, and Robert West. 2023. “Generating Faithful Synthetic Data with Large Language Models: A Case Study in Computational Social Science.” arXiv Preprint arXiv:2305.15041.
Wallach, Hanna, Meera Desai, A. Feder Cooper, et al. 2025. “Position: Evaluating Generative AI Systems Is a Social Science Measurement Challenge.” Proceedings of the 42nd International Conference on Machine Learning, Proceedings of machine learning research, vol. 267: 82232–51.
Wang, Angelina, Jamie Morgenstern, and John P. Dickerson. 2025. “Large Language Models That Replace Human Participants Can Harmfully Misportray and Flatten Identity Groups.” Nature Machine Intelligence 7: 400–411. https://doi.org/10.1038/s42256-025-00986-z.
Wang, Fahui, and Lingbo Liu. 2023. Computational Methods and GIS Applications in Social Science. CRC Press.
Wang, Xinru, Hannah Kim, Sajjadur Rahman, Kushan Mitra, and Zhengjie Miao. 2024. “Human-LLM Collaborative Annotation Through Effective Verification of LLM Labels.” Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems, 1–21.
Ward, Michael D, Brian D Greenhill, and Kristin M Bakke. 2010. “The Perils of Policy by p-Value: Predicting Civil Conflicts.” Journal of Peace Research 47 (4): 363–75. https://doi.org/10.1177/0022343309356491.
Wasserstein, Ronald L., and Nicole A. Lazar. 2016. “The ASA Statement on p-Values: Context, Process, and Purpose.” The American Statistician 70 (2): 129–33. https://doi.org/10.1080/00031305.2016.1154108.
Watts, Duncan J. 2017. “Should Social Science Be More Solution-Oriented?” Nature Human Behaviour 1: 0015. https://doi.org/10.1038/s41562-016-0015.
Wickham, Hadley. 2014. “Tidy Data.” Journal of Statistical Software 59: 1–23. https://doi.org/10.18637/jss.v059.i10.
Wickham, Hadley. 2019. Advanced r. 2nd ed. Chapman; Hall/CRC.
Wickham, Hadley, Mine Çetinkaya-Rundel, and Garrett Grolemund. 2023. R for Data Science. 2nd ed. O’Reilly Media. https://r4ds.hadley.nz/.
Wilke, Claus O. 2019. Fundamentals of Data Visualization. O’Reilly Media, Incorporated.
Wilson, Greg, D. A. Aruliah, C. Titus Brown, et al. 2014. “Best Practices for Scientific Computing.” PLOS Biology 12 (1): e1001745. https://doi.org/10.1371/journal.pbio.1001745.
Wilson, Greg, Jennifer Bryan, Karen Cranston, Justin Kitzes, Lex Nederbragt, and Tracy K. Teal. 2017. “Good Enough Practices in Scientific Computing.” PLOS Computational Biology 13 (6): e1005510. https://doi.org/10.1371/journal.pcbi.1005510.
Wing, Jeannette M. 2006. “Computational Thinking.” Communications of the ACM 49 (3): 33–35. https://doi.org/10.1145/1118178.1118215.
Wooldridge, Jeffrey M. 2019. Introductory Econometrics: A Modern Approach. 7th ed. Cengage Learning.
Wulff, Dirk U., and Rui Mata. 2025. “Semantic Embeddings Reveal and Address Taxonomic Incommensurability in Psychological Measurement.” Nature Human Behaviour 9: 944–54. https://doi.org/10.1038/s41562-024-02089-y.
Wynter, Adrian de. 2025. “Awes, Laws, and Flaws from Today’s LLM Research.” Findings of the Association for Computational Linguistics: ACL 2025 (Vienna, Austria), 12834–54. https://doi.org/10.18653/v1/2025.findings-acl.664.
Xie, Yihui, J. J. Allaire, and Garrett Grolemund. 2018. R Markdown: The Definitive Guide. Chapman; Hall/CRC.
Xie, Yueqi, Lemeng Liang, Shuzhen Li, et al. 2026. “Evaluating the Statistical Realism of LLM-Generated Social Science Data.” Proceedings of the National Academy of Sciences 123 (19): e2538145123. https://doi.org/10.1073/pnas.2538145123.
Xie, Zhengnan, Alice Saebom Kwak, Enfa Fane, et al. 2022. “Extracting Space Situational Awareness Events from News Text.” Proceedings of the Thirteenth Language Resources and Evaluation Conference, 6077–82.
Yarkoni, Tal, and Jacob Westfall. 2017. “Choosing Prediction over Explanation in Psychology: Lessons from Machine Learning.” Perspectives on Psychological Science 12 (6): 1100–1122. https://doi.org/10.1177/1745691617693393.
Yu, Wenhao et al. 2024. “ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks.” Findings of the Association for Computational Linguistics: NAACL 2024 (Mexico City, Mexico), 1333–51. https://doi.org/10.18653/v1/2024.findings-naacl.85.
Zelle, John M. 2017. Python Programming: An Introduction to Computer Science. 3rd ed. Franklin, Beedle & Associates.
Zeller, Andreas. 2009. Why Programs Fail: A Guide to Systematic Debugging. 2nd ed. Morgan Kaufmann.
Zeng, Qiuhai, Claire Jin, Xinyue Wang, Yuhan Zheng, and Qunhua Li. 2025. “AIRepr: An Analyst-Inspector Framework for Evaluating Reproducibility of LLMs in Data Science.” Findings of the Association for Computational Linguistics: EMNLP 2025, 10170–201. https://doi.org/10.18653/v1/2025.findings-emnlp.539.
Zhao, Chengshuai, Zhen Tan, Chau-Wai Wong, Xinyan Zhao, Tianlong Chen, and Huan Liu. 2025. “SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human Intervention.” Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 8473–503. https://doi.org/10.18653/v1/2025.acl-long.416.
Ziems, Caleb, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2024. “Can Large Language Models Transform Computational Social Science?” Computational Linguistics 50 (1): 237–91. https://doi.org/10.1162/coli_a_00502.
Zook, Matthew, Solon Barocas, danah boyd, et al. 2017. “Ten Simple Rules for Responsible Big Data Research.” PLOS Computational Biology 13 (3): e1005399. https://doi.org/10.1371/journal.pcbi.1005399.
Zufall, Elise, and Tyler A Scott. 2024. “Syntactic Measurement of Governance Networks from Textual Data, with Application to Water Management Plans.” Policy Studies Journal 52 (4): 941–54. https://doi.org/10.1111/psj.12556.
Zupic, Ivan. 2026. “A Human-Centered Workflow for Using Large Language Models in Content Analysis.” arXiv Preprint arXiv:2603.19271.