arXiv:2502.15720cs.CYcs.AI2025-02被引 1

提出可开源且由社区所有、对齐、控制的忠诚AI构建路径

Training AI to be Loyal

  • 通过开源+社区治理实现模型所有权与控制权归属
  • 定义社区所有模型需经集体批准使用并共享经济收益
  • 适合关注去中心化AI治理与可持续激励的开发者

忠诚AI是其所属社区的忠实代表。当社区拥有模型的所有权、价值观对齐及控制权时,该AI即为忠诚。社区所有模型须经社区批准方可使用,并共同分享经济收益;社区对齐模型需符合社区共识的价值观;社区控制模型应执行社区设计的功能。为保障对忠诚AI社区的无许可访问,模型必须开源。本文核心科学问题在于:如何构建既开源又由社区所有与治理的模型?本文提出开放、可盈利且忠诚的模型(OML)具体实现路径,基于前期工作arXiv:2411.03887(1)及一个基于密码学-机器学习的库(http://github.com/sentient-agi/oml-1.0-fingerprinting)。

原文摘要 · Abstract (English)

Loyal AI is loyal to the community that builds it. An AI is loyal to a community if the community has ownership, alignment, and control. Community owned models can only be used with the approval of the community and share the economic rewards communally. Community aligned models have values that are aligned with the consensus of the community. Community controlled models perform functions designed by the community. Since we would like permissionless access to the loyal AI's community, we need the AI to be open source. The key scientific question then is: how can we build models that are openly accessible (open source) and yet are owned and governed by the community. This seeming impossibility is the focus of this paper where we outline a concrete pathway to Open, Monetizable and Loyal models (OML), building on our earlier work on OML, arXiv:2411.03887(1) , and a representation via a cryptographic-ML library http://github.com/sentient-agi/oml-1.0-fingerprinting .

AI治理开源模型社区控制

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。