arXiv:2606.03312cs.ROcs.AI2026-06

构建家庭机器人价值冲突评估基准,揭示模型偏好与人类价值观的偏差。

RobotValues: Evaluating Household Robots When Human Values Conflict

论文配图:RobotValues: Evaluating Household Robots When Human Values Conflict
图 1 · 摘自论文原文
  • 通过大模型生成1万例家庭场景,设计多价值冲突任务
  • 80%情况下模型无法纠正默认偏好,选错优先隐私的动作
  • 适合关注机器人伦理决策与价值观对齐的研究者

尽管家用机器人常以任务完成度评估,但日常家庭环境中常存在价值冲突情境,要求机器人在人类自主、效率或社交适宜性等价值间做出权衡。然而,目前缺乏评估机器人价值取向的基准。我们提出RobotValues,一个包含10,000个价值冲突场景的基准,每个场景基于真实家庭图像,提供多个体现不同人类价值的合理行动选项。该基准通过大语言模型辅助生成、利益相关者驱动的价值提取、图像生成及自动质量控制构建。使用该基准评估机器人中的视觉-语言模型,发现模型存在默认偏好(如安全与顺从),而忽视隐私优先动作;当被指令优先考虑与其偏好冲突的价值时,模型在80%情况下仍选择错误动作。结果表明,家用机器人评估应不仅关注任务完成或安全合规,还需考察其在价值冲突中选择合适行为的能力。

原文摘要 · Abstract (English)

While household robots are often evaluated based on task completion, everyday domestic environments involve value-conflicting situations in which robots are expected to choose actions that prioritize other values than task success, such as human autonomy, efficiency, or social appropriateness. Yet, there are no benchmarks for evaluating robots' value preferences in such scenarios. We introduce RobotValues, a benchmark to evaluate household robot planners in 10K value-conflict scenarios. Each instance consists of a realistic household image with multiple plausible robot actions that prioritize different human values. We construct RobotValues through LLM-assisted scenario generation, stakeholder-grounded value extraction, image generation and automatic quality control. Using RobotValues we evaluate VLMs used in robotics and find that models exhibit default value preferences, including safety and accommodation, while underselecting privacy-prioritizing actions. When the models are instructed to prioritize specific values that conflict with their own preferences, they often fail to override their default actions, choosing incorrect actions for 80% of the time. These findings suggest that household robot evaluation should measure not only task completion or safety compliance, but also whether robots can choose among plausible actions when human values conflict.

机器人评估价值对齐家庭机器人伦理决策

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。