arXiv:2608.22223stat.APcs.AI2026-08

根据验证目标优化稀有脑部扫描分配,提升研究效率。

Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN

  • 按目标影响度分配稀有PET扫描,而非仅看预测不确定性
  • 在200次扫描预算下,目标对齐验证使置信区间缩小至随机验证的92.3%
  • 适用于阿尔茨海默病临床/科研验证,尤其关注基因与年龄效应

抗淀粉样蛋白治疗和血液生物标志物正将阿尔茨海默病的诊断流程转变为两阶段模式:先用低成本方法广泛筛查,再在关键节点使用稀缺的确认性淀粉样蛋白正电子发射断层扫描(PET)进行验证。尽管PET仍为淀粉样蛋白负荷的金标准测量手段,但其扫描资源、试验预算及支付方证据包均有限。本文提出一个操作性问题:何时简单透明的验证已足够,何时需引入拟合残差不确定性评分以增加复杂性?对于加权协议目标,个体i的验证价值等于目标影响力与协议残差不确定性的乘积。通用不确定性采样仅依赖后者,可能导致资源浪费于难以预测但对科学、临床或商业主张贡献弱的受试者。本文将该原则应用于A4/LEARN PET数据集,将实际观测的PET数据视为稀缺确认研究的设计实验室。在主要对比——携带APOE4基因型与非携带者的中心值24及以上阳性(以Centiloid为单位)时,简单的APOE4平衡验证几乎捕获了全部目标特异性增益:在200次扫描预算下,目标对齐验证的置信区间宽度比为0.923,而目标特定评分得0.914,通用不确定性采样仅为0.980。其他目标表现不同:年龄斜率分析和基于截断点的阳性判定中,目标特定评分带来更大收益。实践启示清晰:稀缺协议测量应依据待验证的科学主张进行分配,而不仅依据预测不确定性。

原文摘要 · Abstract (English)

Anti-amyloid therapies and blood-based biomarkers are changing Alzheimer disease workups into a two-stage measurement workflow: screen broadly with cheaper information, then spend scarce confirmatory amyloid measurements where they support the decision that will be reported. Amyloid positron-emission tomography (PET) remains one such protocol measurement for amyloid burden, but PET slots, trial budgets, and payer-facing evidence packages are finite. This paper asks a deliberately operational question: when is simple transparent PET validation enough, and when is a fitted residual-uncertainty score worth the added complexity? For a weighted protocol target, the first-order value of validating subject i is the product of target influence and residual protocol uncertainty. Generic uncertainty sampling uses only the second factor and can spend PET measurements on subjects that are hard to predict but weak for the scientific, clinical, or commercial claim. We apply this rule to the A4/LEARN PET archive, treating observed PET as a design laboratory for scarce-confirmation studies. For the primary APOE4 carrier versus non-carrier contrast in Centiloid 24-or-higher PET positivity, simple APOE4-balanced validation recovers nearly all of the target-specific gain: at PET budget 200, the confidence-interval width ratio relative to random validation is 0.923 for APOE4 balancing and 0.914 for target-specific scoring, while generic uncertainty sampling is 0.980. Other targets behave differently: target-specific scoring gives larger gains for an age-slope analysis and for cutoff-indexed PET positivity. The practical message is simple: spend scarce protocol measurements according to the claim being validated, not only according to prediction uncertainty.

阿尔茨海默病PET扫描目标对齐资源优化

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。