arXiv:2512.09772cs.CL2025-12

测试大模型文化倾向,发现DeepSeek和GPT-5偏美,GPT-4在英语下偏中。

DeepSeek's WEIRD Behavior: The cultural alignment of Large Language Models and the effects of prompt language and cultural prompting

  • 用文化提示+语言切换调整模型文化倾向
  • DeepSeek-V3/V3.1与GPT-5始终偏向美国数据
  • GPT-4o/4.1能随提示语言灵活适配中美文化

文化是人际互动的核心,对大语言模型(LLMs)的人类化交互至关重要。本文基于霍夫斯泰德的国际调查数据,分析DeepSeek-V3、V3.1、GPT-4、GPT-4.1、GPT-4o及GPT-5的文化对齐情况。通过系统提示结合语言与文化提示策略,尝试使模型对齐美国与中国。结果显示:DeepSeek-V3、V3.1及GPT-5即使使用文化提示或切换提示语言,仍显著偏向美国数据,未实现对中国的强或弱对齐;而GPT-4在英文提示下更接近中国数据,但文化提示可使其转向美国;低资源模型GPT-4o与GPT-4.1则能根据提示语言和文化提示,有效适配中美文化。

原文摘要 · Abstract (English)

Culture is a core component of human-to-human interaction and plays a vital role in how we perceive and interact with others. Advancements in the effectiveness of Large Language Models (LLMs) in generating human-sounding text have greatly increased the amount of human-to-computer interaction. As this field grows, the cultural alignment of these human-like agents becomes an important field of study. Our work uses Hofstede's VSM13 international surveys to understand the cultural alignment of the following models: DeepSeek-V3, V3.1, GPT-4, GPT-4.1, GPT-4o, and GPT-5. We use a combination of prompt language and cultural prompting, a strategy that uses a system prompt to shift a model's alignment to reflect a specific country, to align these LLMs with the United States and China. Our results show that DeepSeek-V3, V3.1, and OpenAI's GPT-5 exhibit a close alignment with the survey responses of the United States and do not achieve a strong or soft alignment with China, even when using cultural prompts or changing the prompt language. We also find that GPT-4 exhibits an alignment closer to China when prompted in English, but cultural prompting is effective in shifting this alignment closer to the United States. Other low-cost models, GPT-4o and GPT-4.1, respond to the prompt language used (i.e., English or Simplified Chinese) and cultural prompting strategies to create acceptable alignments with both the United States and China.

文化对齐大模型提示工程

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。