用基督教视角评估AI对人类成长的影响,发现当前模型在信仰维度表现差31分。
Evaluating Artificial Intelligence Through a Christian Understanding of Human Flourishing
- 构建基督教视角的评估框架,从七个维度衡量AI对人类成长的塑造作用。
- 20个前沿模型在信仰与灵性维度平均低31分,整体表现下降约17分。
- 揭示AI默认倾向世俗程序化思维,缺乏神学连贯性,适合伦理与神学研究者参考。
人工智能对齐本质上是品格塑造问题,而非仅是安全问题。随着大语言模型越来越多地介入道德反思与精神探索,它们不仅提供信息,更成为数字教义传播工具,主动塑造人类的理解、决策与道德思考。为使这种形塑影响可观察、可测量,我们提出「繁荣型人工智能基准:基督教单轮评估」(FAI-C-ST),用于评估前沿模型在七个维度上是否符合基督教的人类繁荣观。通过对20个前沿模型进行比较,我们发现当前系统并非世界观中立,而是默认采用一种缺乏神学根基的程序性世俗主义,导致所有维度表现系统性下降约17分,尤其在信仰与灵性维度下降达31分。这表明价值观对齐的差距并非技术局限,而是源于训练目标优先追求广泛可接受性与安全性,而非深层且内在一致的道德或神学推理。
原文摘要 · Abstract (English)
Artificial intelligence (AI) alignment is fundamentally a formation problem, not only a safety problem. As Large Language Models (LLMs) increasingly mediate moral deliberation and spiritual inquiry, they do more than provide information; they function as instruments of digital catechesis, actively shaping and ordering human understanding, decision-making, and moral reflection. To make this formative influence visible and measurable, we introduce the Flourishing AI Benchmark: Christian Single-Turn (FAI-C-ST), a framework designed to evaluate Frontier Model responses against a Christian understanding of human flourishing across seven dimensions. By comparing 20 Frontier Models against both pluralistic and Christian-specific criteria, we show that current AI systems are not worldview-neutral. Instead, they default to a Procedural Secularism that lacks the grounding necessary to sustain theological coherence, resulting in a systematic performance decline of approximately 17 points across all dimensions of flourishing. Most critically, there is a 31-point decline in the Faith and Spirituality dimension. These findings suggest that the performance gap in values alignment is not a technical limitation, but arises from training objectives that prioritize broad acceptability and safety over deep, internally coherent moral or theological reasoning.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。