人标记重点时,选择偏好藏个性,但判断什么重要靠的是大家共识。
Personal Salience: Highlighting Is Social, but Individuality Lives in Selection
- 区分出通用重要性、群体共识和个体选择三类信号,用共读控制实验分离变量。
- 个人历史能精准预测自己选中的重点(+0.14差距),但对判断何为重要影响微弱(+0.017)。
- 个体差异主要体现在已知重要段落中如何挑选,而非判断哪些值得标亮。
社交性高亮允许用户标记对自己有意义的文本。我们探究这些自然痕迹中能否还原个体特征,采用共读控制(同一文档被多人标记)来固定文本与主题,检验个人历史是否比他人更能预测其标记行为。研究分离出三类信号:通用重要性(结构)、群体重要性(他人标记)和个人重要性(个体残差)。结果发现,标记行为高度社会性:个人历史对预测自身标记的贡献极小(嵌入评分差距仅+0.017,虽小但显著),远低于群体共识或信息占优基线(能看到同文档他人标记),甚至不如基于其他文档历史训练的前沿大模型;而当已知某些段落已被广泛认为重要时,个人历史却能强有力且无泄漏地预测其选择(差距+0.14)。主题分解显示这种偏好稳定且集中于主题层面,相比同主题读者,个体差异缩小6-8倍,细粒度主题无法进一步分离该残差。关键发现是不对称性:相同评分器下,个体信号在选择环节的作用是重要性判断环节的6-8倍。方法上,传统历史条件评估存在泄漏(约42%样本中目标自身标记进入画像,使得分虚高+0.15 AP),小群体会夸大个性化;本研究采用去泄漏设计、密集群体和模型匹配对照,结论更可靠。强调亮点:真正的个体印记藏于‘从众多重要项中挑出谁’,而非‘认为什么重要’。
原文摘要 · Abstract (English)
Social highlighters let people mark passages that matter to them. We ask how much of an individual is recoverable from these naturalistic traces, using a co-readership identity control (the same document highlighted by many users) that holds document and topic fixed and asks whether a person's own history predicts their marks better than another reader's does. We separate generic salience (structure), crowd salience (what others marked), and personal salience (the individual residual). First, highlighting is social: which sentences you mark is predicted far better by the crowd than by structure or by a personal model, and even a well-estimated crowd, an information-privileged baseline that sees others' marks on the same document, beats a frontier LLM twin built from your other-document history; the within-document personal signal is at most a whisper (own-vs-other gap +0.017 by an embedding scorer, small but significant). Second, in sharp contrast, individuality lives in selection: asked which of the already-salient passages are yours, your own history is a strong, leakage-free predictor (gap +0.14). A topic decomposition shows this is largely stable thematic preference: it shrinks ~6-8x against a topically-matched peer, and a thin residual cannot be separated from finer topic. The non-obvious part is an asymmetry: under the same scorer the individual signal is ~6-8x weaker in salience than in selection. Methodologically, naive history-conditioning evaluations leak (the target's own marks enter the profile in ~42% of pairs, inflating personal scores by up to +0.15 AP) and small crowds overstate personalization; our results are leakage-free, use a dense crowd, and a model-matched control. Highlights carry a genuine individual signature, but a thin layer over a strong shared one, surfacing far more in which salient things a person selects than in what is salient.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。