Questioning the Survey Responses of Large Language Models (2306.07951v4)

Published 13 Jun 2023 in cs.CL

Abstract: Surveys have recently gained popularity as a tool to study LLMs. By comparing survey responses of models to those of human reference populations, researchers aim to infer the demographics, political opinions, or values best represented by current LLMs. In this work, we critically examine this methodology on the basis of the well-established American Community Survey by the U.S. Census Bureau. Evaluating 43 different LLMs using de-facto standard prompting methodologies, we establish two dominant patterns. First, models' responses are governed by ordering and labeling biases, for example, towards survey responses labeled with the letter "A". Second, when adjusting for these systematic biases through randomized answer ordering, models across the board trend towards uniformly random survey responses, irrespective of model size or pre-training data. As a result, in contrast to conjectures from prior work, survey-derived alignment measures often permit a simple explanation: models consistently appear to better represent subgroups whose aggregate statistics are closest to uniform for any survey under consideration.

References (47)

Citations (17)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

GitHub

GitHub - socialfoundations/surveying-language-models: Code to reproduce the paper "Questioning the Survey Responses of Large Language Models" (9 stars)

Tweets

https://twitter.com/StefanResearch/status/1772203384478499152

https://twitter.com/ShashwatGoel7/status/1776858168334700886

Questioning the Survey Responses of Large Language Models (2306.07951v4)

Summary

Related Papers

GitHub

Tweets