arXiv:2608.23979cs.AIcs.CR2026-08

用可审计规则替代黑箱推荐,让投票者自主决定看哪些论点。

Rules Before Oracles: Auditable, User-Configurable Argument Selection for Deliberative Polling

  • 设计可公开验证的推荐规则,参数由用户掌控,确保透明性。
  • 在论点覆盖和支持度上优于黑箱模型,差距随对抗压力增大而拉大。
  • 适合关注公平与可问责性的公共决策系统开发者或研究者。

在协商式民意调查中,当论点数量超过人阅读能力时,需有机制筛选每个投票者可见内容,现有做法依赖不透明的模型,使选民无法复现或质疑影响其决策的信息暴露。本文提出基于公开规则的推荐机制,以可公开计算的证据为基础,参数由用户持有,将可读性作为可用机制的准入条件而非与准确性的权衡。我们形式化了对二元理由集合的投票机制,评估标准包括理由覆盖率、到达顺序和被支持的质量;给出七个可验证的标准,并提出一种一跳反向支持流规则,由关系权重函数参数化。通过约17,000次种子配对运行的智能体模拟器记录每轮投票的所有提案组合。所生成提案比标签阅读上限低0.035,表明任何无约束排序器的优势有限且微小。仅考虑覆盖率时,在非退化的创作条件下,该规则与随机提案无异(因无视顺序与善意);但在另外两项指标上,其表现始终领先,优势随对抗压力加剧而扩大,支持质量得分达随机方案的3.3倍。一旦真实比例的提案缺乏理由,覆盖率优势回归并持续扩大。在同质化泛滥下,平坦权重策略使完整性从0.81降至0.34,而作者计数归一化仅降至0.44,表明权重函数是关键安全控制,能挽回10%的完整性。排序选择本质是在覆盖率与支持质量之间的权衡,这种抉择只能由可读规则赋予受影响者,且已映射至开源点对点平台。

原文摘要 · Abstract (English)

In a deliberative poll, once submissions outnumber what anyone will read, some mechanism chooses which arguments each voter sees, acquiring much of the decision; practice delegates it to opaque learned rankers, so a voter cannot recompute or contest the exposure that shaped their vote. We ask whether it can be a published rule over publicly recomputable evidence with parameters held by the voter, treating legibility as an admissibility condition on usable mechanisms, not an objective traded against accuracy. We formalise a poll over bipolar justification sets, judging a slate by reason coverage, the order it arrives in, and captured endorsement mass; we give seven checkable criteria for a civic recommender and a rule meeting them: a one-hop reversed endorsement flow parameterised by a relation-weight function. An agentic simulator records every slate at every vote, over about 17,000 seed-paired runs. Served slates fall 0.035 short of a label-reading ceiling upper-bounding every selection procedure, opaque ones included: any unconstrained ranker's advantage is bounded and small. On coverage alone, with non-degenerate authoring, the rule is indistinguishable from a random slate, a null due to an order-blind, charity-blind instrument; on the other two it leads at every prefix by a margin widening with adversarial pressure and dominates on mass by a factor of 3.3. Once a realistic fraction of submissions carries no reasons, the coverage margin returns and grows. Label-homogeneous flooding collapses completeness from 0.81 to 0.34 under a flat weight policy, only to 0.44 under author-count normalisation, making the weight function a security control worth 10% of completeness. The choice between ranking arms is a position on a coverage-versus-mass frontier, not a fact, the kind of choice only a legible rule can hand to the person it affects. It maps onto an open-source peer-to-peer platform.

协商民主可审计推荐公共决策

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。