Out of One, Many: Using Language Models to Simulate Human Samples-Reference-Cited by-同舟云学术

Out of One, Many: Using Language Models to Simulate Human Samples

Published:2023-02-21 Issue:3 Volume:31 Page:337-351
ISSN:1047-1987
Container-title:Political Analysis
language:en
Short-container-title:Polit. Anal.

Author:

Argyle Lisa P.^ORCID,Busby Ethan C.,Fulda Nancy,Gubler Joshua R.^ORCID,Rytting Christopher,Wingate David

Abstract

AbstractWe propose and explore the possibility that language models can be studied as effective proxies for specific human subpopulations in social science research. Practical and research applications of artificial intelligence tools have sometimes been limited by problematic biases (such as racism or sexism), which are often treated as uniform properties of the models. We show that the “algorithmic bias” within one such tool—the GPT-3 language model—is instead both fine-grained and demographically correlated, meaning that proper conditioning will cause it to accurately emulate response distributions from a wide variety of human subgroups. We term this propertyalgorithmic fidelityand explore its extent in GPT-3. We create “silicon samples” by conditioning the model on thousands of sociodemographic backstories from real human participants in multiple large surveys conducted in the United States. We then compare the silicon and human samples to demonstrate that the information contained in GPT-3 goes far beyond surface similarity. It is nuanced, multifaceted, and reflects the complex interplay between ideas, attitudes, and sociocultural context that characterize human attitudes. We suggest that language models with sufficient algorithmic fidelity thus constitute a novel and powerful tool to advance understanding of humans and society across a variety of disciplines.

Publisher

Cambridge University Press (CUP)

Subject

Political Science and International Relations,Sociology and Political Science

Reference41 articles.

1. Affect, Not Ideology

2. Understanding the Role of Racism in Contemporary US Public Opinion

3. Word Embeddings for the Analysis of Ideological Placement in Parliamentary Corpora

Cited by 95 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The use of ChatGPT for personality research: Administering questionnaires using generated personas;Personality and Individual Differences;2024-10

2. Negative social tipping dynamics resulting from and reinforcing Earth system destabilization;Earth System Dynamics;2024-09-10

3. Understanding AI-Generated Experiments in Tourism: Replications Using GPT Simulations;Journal of Travel Research;2024-09-05

4. The use of synthetic data in tourism;Annals of Tourism Research;2024-09

5. Performance and biases of Large Language Models in public opinion simulation;Humanities and Social Sciences Communications;2024-08-28