Can Large Language Models Transform Computational Social Science
Bibliographic Data
Large language models (LLMs) are capable of successfully performing many language processing tasks zero-shot (without training data). If zero-shot LLMs can also reliably classify and explain social phenomena like persuasiveness and political ideology, then LLMs could augment the computational social science (CSS) pipeline in important ways. This work provides a road map for using LLMs as CSS tools. Towards this end, we contribute a set of prompting best practices and an extensive evaluation pipeline to measure the zero-shot performance of 13 language models on 25 representative English CSS benchmarks. On taxonomic labeling tasks (classification), LLMs fail to outperform the best fine-tuned models but still achieve fair levels of agreement with humans. On free-form coding tasks (generation), LLMs produce explanations that often exceed the quality of crowdworkers' gold references. We conclude that the performance of today's LLMs can augment the CSS research pipeline in two ways: (1) serving as zero-shot data annotators on human annotation teams, and (2) bootstrapping challenging creative generation tasks (e.g., explaining the underlying attributes of a text). In summary, LLMs are posed to meaningfully participate in social science analysis in partnership with humans
Bootstrapping (finance · Computational and Text Analysis Methods · Natural Language Processing Techniques · Topic Modeling
Inside the meat locker
Bridging the silos in affective AI
Designing social surveys for understanding farming and natural resource management
Mining for Meaning
Durably reducing conspiracy beliefs through dialogues with AI
Vibe researching
What kind of field is computational social science? Fragmentation, pluralism, and unification
The fidelity of large language models in social science
Large Language Models as Psychological Simulators
Mapping climate change coverage
Reconfiguring Responsibility
Advancing qualitative analysis in professional disaster and risk communication
TikTok engagement traces over time and health risky behaviors
How does social support detected automatically in discussion forums relate to online learning burnout? The moderating role of students’ self-regulated learning
Evolving linguistic divergence on polarizing social media
Methods for aggregating investor sentiment from social media
Machine-assisted quantitizing designs
Best practices for responsibly using AI tools in social sciences research
Understanding Large Language Model Driven Social Bots
MMInfluencer
DataPoll
Examining media’s coverage of Covid-19 vaccines and social media sentiments on vaccine manufacturers’ stock prices
Towards the future of pedestrian–AV interaction
Empathy in action
Automating Content Analysis With Multiple LLM Agents
Partisan Knowledge Claims in Congressional Oversight
Leveraging prompt-based LLMs for automated scoring and feedback generation in higher education
Exploring the potential of generative AI to complement multi-stakeholder landscape preference assessment
Beyond sentiment
Digital grievances
Using language models to label clusters of scientific documents
Evolving Landscapes, Shifting Narratives
Toward open-source foundation model ecosystem
Comparative analysis of GPT-4, Gemini, and Ernie as gloss sign language translators in special education
Who believes in science? A computational tool for identifying language invoking or disputing scientific knowledge
Goodbye human annotators? Content analysis of social policy debates using ChatGPT
Going with the Mainstream
Co-Opetition in Crowdsourcing Challenges
Variation across Regions and Demographics in African American Language Morphosyntax
Stepping stones or party placements? Post-ministerial cabinet careers of Belgian ministerial advisers (1999–2020)
Beyond single snapshots
Natural language processing for social science research
Collaborative Growth
The AI-Reflexivity Checklist (ARC)
Using Digital Tools to Understand Global Development Continuums
Misogyny as a Structural Hub
Classification bias of LLMs in detecting incivility towards female and male politicians in German social media discourse
Multi-Resolution Design
Tailoring generative AI chatbots for multiethnic communities in disaster preparedness communication
Finding love in algorithms
Applications of GPT in Political Science Research
Measuring Politicians’ Public Personality Traits Using Computational Text Analysis
Political Debate
Prompting the Machine
The Efficacy of Large Language Models and Crowd Annotation for Accurate Content Analysis of Political Social Media Messages
Embracing Dialectic Intersubjectivity
Navigating the Risks of Using Large Language Models for Text Annotation in Social Science Research
Survey of Cultural Awareness in Language Models
Socially Aware Language Technologies
Evaluating Synthetic Data Generation from User Generated Text
Administrative Decision-Making with Generative AI
Computational Basis of Large Language Models’ Decision Making in Social Simulation
Simulating social perception with large language models
Fine-tuned large language models can replicate expert coding better than trained coders
Mapping (A)Ideology
When Does Surveillance Trigger Resistance? Public Response to Escalating Digital Control in China
Analyzing narrative contagion through digital storytelling in social media conversations
Studying economic black holes
Open-source LLMs for text annotation
The ethics of generative AI in social science research
The role of generative AI in navigating trade-offs in policy research design
Codebook LLMs
Synthetically generated text for supervised text analysis
Positioning Political Texts with Large Language Models by Asking and Averaging
What happened to Putin’s friends? The radical right’s reaction to the Russian invasion on social media
Intelligent Computing Social Modeling and Methodological Innovations in Political Science in the Era of Large Language Models
AI as super-controversy
Political expression of academics on Twitter
Large language models (LLM) in computational social science
Large language models' varying accuracy in recognizing risk-promoting and health-supporting sentiments in public health discourse
Using large language models for preprocessing and information extraction from unstructured text
Start Generating
The Russia-Ukraine War in Chinese Social Media
Generative Multimodal Models for Social Science
Updating 'The Future of Coding
Integrating Generative Artificial Intelligence into Social Science Research
Using Large Language Models for Qualitative Analysis can Introduce Serious Bias
Correcting the Measurement Errors of AI-Assisted Labeling in Image Analysis Using Design-Based Supervised Learning
Machine Bias. How Do Generative Language Models Answer Opinion Polls
The Mixed Subjects Design
Large Language Models for Text Classification
Seeded Topic Models in Digital Archives
Meaning in Hyperspace
Content analysis in communication research
The general inquirer
Affective Computing
The Grammar of Society
Networks, Crowds, and Markets
Explanation in artificial intelligence
Personalized Persuasion
Choosing Prediction Over Explanation in Psychology
A General Psychoevolutionary Theory of Emotion
Persistent Anti-Muslim Bias in Large Language Models
Private traits and attributes are predictable from digital records of human behavior
Word embeddings quantify 100 years of gender and ethnic stereotypes
Twitter mood predicts the stock market
To Explain or to Predict?
Generative Agents
Evaluation methods for topic models
The Risk of Racial Bias in Hate Speech Detection
Prediction and explanation in social systems
Beyond Western, Educated, Industrial, Rich, and Democratic (Weird) Psychology
Experimental evidence for tipping points in social convention
ChatGPT outperforms crowd workers for text-annotation tasks
Construal-level theory of psychological distance.
Computational Social Science
Risk, Ambiguity, and the Savage Axioms
What makes a metaphor literary? Answers from two computational studies
Social Media and the Decision to Participate in Political Protest
Terrified
Event History Modeling
An Introduction to Sociolinguistics
The online disinhibition effect
Empathic communities
Humor and social distance in elementary school children
Dialogue Act Modeling for Automatic Tagging and Recognition of Conversational Speech
Novel Event Detection and Classification for Historical Texts
Cultural Performances
Emotional Feedback and the Viral Spread of Social Media Messages About Autism Spectrum Disorders
Effects of Humor on Persuasion
Dynamic models of segregation
Exploiting affinities between topic modeling and the sociological perspective on culture
Gender identity and lexical variation in social media
Analytical sociology and computational social science
Conceptions of Time and Events in Social Science Methods
Emotions, Oral Arguments, and Supreme Court Decision Making
Out of One, Many
A Bayesian Hierarchical Topic Model for Political Texts
Solidarity or Schism
An Automated Information Extraction Tool for International Conflict Data with Performance as Good as Human Coders
Politeness
Framing
Framing responsibility for political issues
Computational Sociolinguistics
Identity and interaction
Adapting computational text analysis to social science (and vice versa)
Identification of leaders, lurkers, associates and spammers in a social network
The Geometry of Culture
Significant themes in 19th-century literature
Choice Under Uncertainty
Personality and Political Attitudes
Machine Translation
Digital Footprints
Cycles of Conflict, a Century of Continuity
Do Anti-Immigrant Laws Shape Public Sentiment? A Study of Arizona's SB 1070 Using Twitter Data
| Unique citing works | 93 |
|---|---|
| Citations per year | 46,5 |
| Citation span | 2024 - 2026 (3) |
| Citation velocity | current |
| Highly cited | No |
| Citation types | Neutral: 90 |