arXiv:2605.18423cs.ROcs.CY2026-05

构建可量化评估自主系统伦理能力的基准测试框架。

REBAR: Reference Ethical Benchmark for Autonomy Readiness

论文配图:REBAR: Reference Ethical Benchmark for Autonomy Readiness
图 1 · 摘自论文原文
  • 用神经符号大模型分析场景伦理难度并生成测试用例。
  • 通过仿真环境评估,输出可重复的自主性成熟度评分。
  • 适合需验证系统伦理合规性的研发与监管方使用。

随着自主系统日益复杂,客观衡量其伦理与法律合规性的指标对揭示系统局限性、确保滥用者问责至关重要。现有伦理具身AI框架多为定性,依赖安全护栏或定向红队测试,但护栏常直接禁止危险行为,缺乏用户干预或可解释原因。为此,本文提出参考伦理自主性准备度基准(REBAR),一个用于自主系统的量化测试与评估框架。REBAR将运行指标映射为可计算的自主性成熟度等级(ARL)评分体系,以量化伦理表现。其核心创新包括:基于神经符号大模型的场景伦理难度计算与解释方法、由大模型驱动的大规模测试实例生成,以及高度逼真的仿真环境。通过此严谨测试流程评估白盒自主解决方案,REBAR提供客观、可复现的基准得分,弥合抽象原则与可验证、可问责自主性之间的差距。

原文摘要 · Abstract (English)

As autonomous systems grow more advanced, objective metrics to evaluate their ethical and legal compliance are critical for informing end users of their limitations and ensuring accountability of those who misuse them. Current ethical embodied AI frameworks remain mostly qualitative, focusing on system design (through safety guardrails or targeted red teaming), and the realized guardrails often directly disallow unsafe behavior without providing the user with an override or interpretable reason. Instead, there is a need for computable metrics through rigorous testing that allow a user to determine the applicability of the system to the task. To address this gap, we introduce the Reference Ethical Benchmark for Autonomy Readiness (REBAR), a quantitative test and evaluation framework for autonomous systems. REBAR maps operating metrics into a computable Autonomy Readiness Level (ARL) rubric that can quantify ethical performance. Key innovations of the framework include a neuro-symbolic Large Language Model (LLM) approach to calculate and explain the ethical difficulty of scenarios, LLM-driven at-scale generation of test instances, and a versatile, photorealistic simulation environment. By evaluating white-box autonomy solutions through this rigorous testing pipeline, REBAR delivers an objective and repeatable benchmark score, bridging the gap between abstract principles and verifiable, accountable autonomy.

自主系统伦理评估基准测试

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。