arXiv:2606.16874cs.CLcs.CE2026-06

分析Reddit用户自述揭露诈骗链路,揭示多年趋势与路径演化。

Understanding Scam Trends and Rail Paths from Reddit Self-Disclosure Narratives

论文配图:Understanding Scam Trends and Rail Paths from Reddit Self-Disclosure Narratives
图 1 · 摘自论文原文
  • 基于2023–2025年Reddit自述帖,构建多阶段诈骗链数据集。
  • 发现诈骗路径以多链路为主,不同年份主导类型和环节各异。
  • 适合安全研究者与平台风控团队参考,助力生成式诈骗模拟。

在线诈骗行为具有多阶段特征,其生命周期包含时序排列的链路与事件,而非孤立信号。现有研究虽分析了诈骗类型与链路特征,但缺乏跨年度趋势追踪。同时,因缺少开源标注数据集,链路间关系研究受限。为此,本文利用2023至2025年Reddit诈骗相关子版块的自述帖子,收集21,304篇含身份、通信、平台或支付至少一环的帖子,通过启发式标注进行趋势分析;再采用大模型辅助方法标注1,800篇含明确或可恢复诈骗链的帖子,经人工验证后用于路径分析;最后对帖子评论运行主题模型,探究社区支持行为。结果表明:诈骗过程以多链路为主,各年份主导类型与链路组件变化明显,不同诈骗类型在路径复杂度上系统性差异显著。社区支持行为随时间趋于细化。本研究支持合成诈骗链数据生成与人工智能风险评估,但结论可能不适用于其他平台。

原文摘要 · Abstract (English)

Online scam behavior is inherently multi-stage, and the lifecycle includes temporally ordered rails and events rather than isolated signals. Existing works analyze characteristics of scam types and rails, but they do not track scam trends across years. Moreover, the work on the relations between rails is hampered due to the lack of open-source datasets with annotations and coverage of different scam types. To address these gaps, we build a dataset to analyze the yearly trend of scam characteristics and rail paths using Reddit self-disclosure narratives from 2023 to 2025. We collect 21,304 posts from scam-related subreddits with at least one rail among identity, communication, platform, and payment for trend analysis by heuristic annotation. Then, we label 1,800 posts containing explicit or recoverable scam chains by an LLM-assisted method for scam path analysis. The method is evaluated with human annotation. Lastly, we run a topic model on the comments of the posts to analyze the community support behavior. The results reveal that scam processes are predominantly multi-rail. Across years, different scam types and rail components dominate. Different scam types vary systematically in path complexity. Reddit support behaviors have become more detailed over time. This work supports synthetic scam chain data simulation and AI-related scam risk assessment, though findings may not generalise to other platforms.

诈骗分析社交平台链路建模

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。