arXiv:2506.01998cs.HCcs.AI2025-06被引 6

用日语自指词研究语音助手身份模糊性,发现声音与用词可共同制造性别困惑。

Inter(sectional) Alia(s): Ambiguity in Voice Agent Identity via Intersectional Japanese Self-Referents

  • 通过日语自指词测试不同语音助手的身份感知效果
  • 发现'仆'和'我'等词能模糊性别,与声音协同制造身份迷惑
  • 揭示文化语境下语音代理身份认知的复杂性,适合人机交互研究者参考

模拟人类形象的对话代理引发了关于机器拟人化中社会身份暗示的伦理问题。有研究指出,日本语境下的交叉身份自指词会引发复杂且常具回避性的代理身份印象。然而,其他“中性”非代词自指表达(NPSR)及语音作为社会性表达媒介的作用尚未被探索。本研究通过众包实验,让204名日本参与者评估三种ChatGPT语音(Juniper、Breeze、Ember)搭配七种自指词。结果表明,语音存在明显性别化倾向,而交叉身份自指词具有规避性别化的潜力,即通过中立与模糊制造身份暧昧。尤其在‘仆’和‘我’等词上,年龄与正式程度感知与性别化相互交织,符合社会语言学理论。本研究为代理身份感知提供了更细腻的理解,并倡导在语音代理领域开展交叉性与文化敏感的研究。

原文摘要 · Abstract (English)

Conversational agents that mimic people have raised questions about the ethics of anthropomorphizing machines with human social identity cues. Critics have also questioned assumptions of identity neutrality in humanlike agents. Recent work has revealed that intersectional Japanese pronouns can elicit complex and sometimes evasive impressions of agent identity. Yet, the role of other "neutral" non-pronominal self-referents (NPSR) and voice as a socially expressive medium remains unexplored. In a crowdsourcing study, Japanese participants (N = 204) evaluated three ChatGPT voices (Juniper, Breeze, and Ember) using seven self-referents. We found strong evidence of voice gendering alongside the potential of intersectional self-referents to evade gendering, i.e., ambiguity through neutrality and elusiveness. Notably, perceptions of age and formality intersected with gendering as per sociolinguistic theories, especially boku and watakushi. This work provides a nuanced take on agent identity perceptions and champions intersectional and culturally-sensitive work on voice agents.

语音代理身份模糊文化敏感

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。