研究中文生成式搜索如何选源、引用和呈现信息,揭示界面差异与内容可信度机制。
What Do Chinese-Language Generative Search Engines Cite and Surface? A Large-Scale Empirical Study
- 对比4大平台8个接口,系统分析16万条引用数据的选源逻辑。
- 品牌引用率仅8.3%,联系方式仅12.4%来自有联系方式的页面。
- 发现网页时效性影响内容寿命,且App与网页端引用差异显著。
生成式AI问答系统正重塑信息获取方式,将内容可见性从排名结果转向生成答案中的检索、引用与呈现。我们对四大主流平台的中文生成式搜索进行了大规模实证研究,覆盖网页与应用端共8个界面、614个查询及每组合3次重复。从21.4万条原始记录中构建出16.08万条清洗后的引用级数据集,分析引用行为、来源归属、实体曝光与跨界面一致性。五项发现:第一,品牌在引用池中被选择的比例为8.3%,含联系方式的页面中12.4%将信息引入答案;第二,内容匹配度、跨源出现频次与语义角色是重要预测因素,而5118-Baidu综合质量评分未成为任一结果的主要预测因子;第三,带发布时间的页面半衰期约为高时效性查询39天、低时效性查询68天;第四,约13%的品牌曝光无法匹配同期引用池,约71%的联系方式曝光无法匹配爬取正文;第五,同一平台的App与网页端引用集合存在系统性差异。结果揭示了中文生成式搜索的信息筛选与呈现机制,表明界面类型是关键分析维度。
原文摘要 · Abstract (English)
Generative AI question-answering systems increasingly mediate information access, shifting content visibility from ranked search results to retrieval, citation, and presentation in generated answers. We conduct a large-scale empirical study of Chinese-language generative search across the Web and App interfaces of four mainstream platforms. The controlled design covers eight platform interfaces, 614 queries, and three replications per query-platform-interface combination. From 214,119 raw records, we construct a cleaned citation-level dataset of 160,860 records and analyze citation behavior, source attribution, entity exposure, and cross-interface consistency. Five findings emerge. First, brands in the citation pool were selectively surfaced in answers: the overall brand-selection rate was 8.3%, and 12.4% of retrieved sources containing contact information contributed contact information to answers. Second, content fit, cross-source occurrence count, and semantic role were relatively important in predictive models, whereas the 5118-Baidu Composite Quality Score was not the leading predictor for any examined outcome. Third, among cited pages with publication dates, fitted half-lives were approximately 39 days for high-timeliness queries and 68 days for low-timeliness queries. Fourth, approximately 13% of brand exposures could not be matched to the contemporaneous citation pool, and approximately 71% of contact-information exposures could not be matched to the crawled body text. Fifth, source sets differed systematically between the App and Web interfaces of the same platform. These results characterize how Chinese-language generative search systems select, attribute, and surface information and show that interface type is an important dimension of analysis.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。