AI决策需随时间动态对齐多方利益,避免僵化评估。
Pluralistic Alignment Over Time
- 引入时间维度评估多利益相关方价值对齐
- 支持不同时间点体现不同群体的价值偏好
- 适合长期交互型AI系统的设计与评估
若人工智能系统在时间上持续做决策,我们应如何评估其与一组利益相关方(可能价值观冲突)的对齐程度?本文主张考虑时间因素,包括利益相关方满意度的变化及可能的时间延展性偏好。我们提出将近期公平性评估方法拓展至一种新型多元对齐形式:时间多元主义,即在不同时段反映不同利益相关方的价值取向。
原文摘要 · Abstract (English)
If an AI system makes decisions over time, how should we evaluate how aligned it is with a group of stakeholders (who may have conflicting values and preferences)? In this position paper, we advocate for consideration of temporal aspects including stakeholders' changing levels of satisfaction and their possibly temporally extended preferences. We suggest how a recent approach to evaluating fairness over time could be applied to a new form of pluralistic alignment: temporal pluralism, where the AI system reflects different stakeholders' values at different times.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。