arXiv:2411.06032cs.CL2024-11被引 32

用文化心理学框架评估大模型的文化价值观差异

LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output

  • 基于GLOBE框架设计评测基准,自动分析大模型输出的文化倾向
  • 对比中美大模型发现东西方文化价值存在系统性差异
  • 为大模型文化对齐与人机协作提供新评估思路

大量研究致力于减少生成内容中的有害或偏见信息,并使AI输出更符合人类意图;然而,关于大模型文化价值观的研究仍处于初期阶段。文化价值观深刻影响社会运行机制,体现成员的规范、优先级和决策模式。为此,我们借鉴文化心理学理论和经实证验证的GLOBE框架,提出LLM-GLOBE基准,用于评估大模型的文化价值体系,并利用该基准比较中国与美国大模型的文化价值观。方法上引入创新的“大模型作为陪审团”流水线,自动化评估开放式生成内容,实现概念层面的大规模分析。结果揭示了东西方文化价值体系间的异同,表明开放生成任务是评估文化价值观更具前景的方向。研究还探讨了其对后续模型开发、评估与部署的影响,尤其在大模型文化对齐及人工智能文化价值对人机协作结果的影响方面。

原文摘要 · Abstract (English)

Immense effort has been dedicated to minimizing the presence of harmful or biased generative content and better aligning AI output to human intention; however, research investigating the cultural values of LLMs is still in very early stages. Cultural values underpin how societies operate, providing profound insights into the norms, priorities, and decision making of their members. In recognition of this need for further research, we draw upon cultural psychology theory and the empirically-validated GLOBE framework to propose the LLM-GLOBE benchmark for evaluating the cultural value systems of LLMs, and we then leverage the benchmark to compare the values of Chinese and US LLMs. Our methodology includes a novel "LLMs-as-a-Jury" pipeline which automates the evaluation of open-ended content to enable large-scale analysis at a conceptual level. Results clarify similarities and differences that exist between Eastern and Western cultural value systems and suggest that open-generation tasks represent a more promising direction for evaluation of cultural values. We interpret the implications of this research for subsequent model development, evaluation, and deployment efforts as they relate to LLMs, AI cultural alignment more broadly, and the influence of AI cultural value systems on human-AI collaboration outcomes.

文化对齐大模型评估GLOBE框架

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。