Skip to main navigation Skip to search Skip to main content

Whose Values Prevail? Bias in Large Language Model Value Alignment

  • Ruoxi Qi (Co-first Author)
  • , Gleb Papyshev (Co-first Author)
  • , Kellee Tsai
  • , Antoni B. Chan
  • , Janet Hsiao

Research output: Chapters, Conference Papers, Creative and Literary WorksRGC 32 - Refereed conference paper (with host publication)peer-review

Abstract

As large language models (LLMs) are increasingly integrated into our lives, concerns have been raised about whether they are biased towards the values of particular cultures. We show that while LLMs were biased toward the values of WEIRD populations, some non-Western populations, including East Asia and Russia, were also represented relatively well. Notably, the Rich dimension was the strongest predictor of LLM's alignment instead of the most discussed Western dimension. This suggests the need to attend to less prosperous populations instead of focusing only on easily accessible populations. We also found that one source of this bias could be unbalanced training data as approximated by an Internet Freedom measure, and that prompting the model to act as individuals from different populations reduced the bias but could not eliminate it. These findings raise the importance of training process disclosure and the consideration of culture-specific models to ensure ethical usage of LLMs. © 2025 by the author(s).
Original languageEnglish
Title of host publicationProceedings of the 47th Annual Conference of the Cognitive Science Society
PublisherUniversity of California
Pages665-672
Publication statusPublished - Jul 2025
Event47th Annual Meeting of the Cognitive Science Society (CogSci 2025) - San Francisco, United States
Duration: 30 Jul 20252 Aug 2025

Publication series

NameProceedings of the Annual Meeting of the Cognitive Science Society
Volume47
ISSN (Electronic)1069-7977

Conference

Conference47th Annual Meeting of the Cognitive Science Society (CogSci 2025)
PlaceUnited States
CitySan Francisco
Period30/07/252/08/25

Research Keywords

  • Large Language Model (LLM)
  • Value Alignment
  • WEIRD Population

Fingerprint

Dive into the research topics of 'Whose Values Prevail? Bias in Large Language Model Value Alignment'. Together they form a unique fingerprint.

Cite this