通过提前预测意图分歧,在自动驾驶决策前设置安全闸门,防止交互后规划失败。
Gating Before Commitment: Anticipating Intent Divergence to Prevent Post-Interaction Decision Failures in Autonomous Driving

- 在决策层引入语言引导的意图模块,用平滑的意图-几何差异分数触发安全闸门。
- 在10次重放中均成功拦截车辆偏离路径,比走廊出口早161毫秒响应。
- 可减少误触发,适合用于高风险自动驾驶场景中的实时安全防护。
车辆交互中的意图误判导致反复规划失败。本文研究一种决策层机制:语言引导的意图模块读取结构化描述,计算平滑的意图-几何差异得分,并在提交计划前进行门控,位于走廊包络线之前。在一次离路偏离和四段碰撞片段的回放测试中,仅此门控层能修复规划错误;主案例中,门控在漂移开始后72毫秒触发,但比走廊退出早161毫秒,确保所有十次重放轨迹均在走廊内。首次校准在5.9分钟内产生九次误触发,源于冲突评分不确定性(半冲突);预注册重构将不确定性视为回避,使误触发率降至每分钟0.341次。两个消融实验界定模型作用:完整评分在部署有效条件下对五次失败中的四次检测最快,无否决规则下三例中三次检测成功(000871快一周期;000228在不确定路段提前触发,五段视频无法分类为信号或偶然);去掉置信度项则损失两次检测。几何规则在同误报率下,检测能力提升超三倍。证据支持该门控机制,模型核心作用为最快检测及对几何规则的不确定性否决。
原文摘要 · Abstract (English)
Intent misinterpretation during vehicle interactions causes recurring planning failures. We study a decision layer in which a language-guided intent module reads structured descriptors, computes a smoothed intent-geometry divergence score, and gates the planned maneuver before commitment, upstream of a corridor envelope. On a replayed off-road departure and four crash clips under a frozen, disclosed implementation, gating is the only layer that repairs the plan: on the main case it fires 72 ms after the drift onset but 161 ms before the corridor exit, keeping the trajectory in the corridor in all ten replays. The first calibration draws nine false triggers in 5.9 minutes, each from scoring uncertainty as half a conflict; a preregistered redesign treating uncertainty as abstention cuts this to 0.341 per minute. Two ablations bound the model's contribution: the full score detects fastest on four of five failures under the deployed eligibility, three of five against the unvetoed rule (000871 by one cycle; 000228 by a pre-onset fire on an uncertain stretch that five clips cannot classify as signal or coincidence; dropping the confidence term costs two detections), while on in-domain tracks at equal false positives the geometric rule more than triples its detection. The evidence supports the gating mechanism; the model's demonstrated roles are the fastest detection on these failures and an uncertainty veto on the geometric rule.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。