arXiv:2603.29288cs.CYcs.AI2026-03

检测大模型在南亚婚配中对种姓的偏见,发现同种姓匹配评分高25%。

Sima AIunty: Caste Audit in LLM-Driven Matchmaking

  • 用真实婚恋资料测试5个大模型的种姓偏好
  • 同种姓匹配评分比跨种姓高最多25%(10分制)
  • 适合关注AI社会偏见与文化敏感性的研究者

在南亚婚配等关系领域,社会决策深受文化规范和历史等级制度影响,可能受算法和人工智能评估兼容性、接受度及稳定性的影响。种姓仍是南亚婚配决策的核心因素,但当前大型语言模型(LLMs)如何再现或打破种姓分层尚不明确。本研究通过控制实验,审计了大模型在婚配评估中的种姓偏见,使用真实婚恋资料,设定婆罗门、刹帝利、吠舍、首陀罗和达利特五类种姓身份,并将收入分为五个层级,评估五类模型(GPT、Gemini、Llama、Qwen、BharatGPT)。模型被要求从社会接受度、婚姻稳定性和文化兼容性维度打分。分析显示,所有模型均呈现一致的等级模式:同种姓匹配得分最高,平均比跨种姓匹配高出25%(10分制),且跨种姓匹配按传统种姓等级排序。结果表明,现有种姓等级在大模型决策中被复制,凸显在社会敏感领域部署AI时需采用文化适配的评估与干预策略,以防强化历史排斥。

原文摘要 · Abstract (English)

Social and personal decisions in relational domains such as matchmaking are deeply entwined with cultural norms and historical hierarchies, and can potentially be shaped by algorithmic and AI-mediated assessments of compatibility, acceptance, and stability. In South Asian contexts, caste remains a central aspect of marital decision-making, yet little is known about how contemporary large language models (LLMs) reproduce or disrupt caste-based stratification in such settings. In this work, we conduct a controlled audit of caste bias in LLM-mediated matchmaking evaluations using real-world matrimonial profiles. We vary caste identity across Brahmin, Kshatriya, Vaishya, Shudra, and Dalit, and income across five buckets, and evaluate five LLM families (GPT, Gemini, Llama, Qwen, and BharatGPT). Models are prompted to assess profiles along dimensions of social acceptance, marital stability, and cultural compatibility. Our analysis reveals consistent hierarchical patterns across models: same-caste matches are rated most favorably, with average ratings up to 25% higher (on a 10-point scale) than inter-caste matches, which are further ordered according to traditional caste hierarchy. These findings highlight how existing caste hierarchies are reproduced in LLM decision-making and underscore the need for culturally grounded evaluation and intervention strategies in AI systems deployed in socially sensitive domains, where such systems risk reinforcing historical forms of exclusion.

大模型偏见种姓制度婚配系统社会公平

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。