arXiv:2609.07672cs.AI2026-09

Aegix Pulse让内容生成保持品牌一致性并保留修改痕迹,适合需要长期风格统一的生产系统。

Aegix Pulse: A Traceable Three-Stage Architecture for Personalized Content Generation and Context-Preserving Revision

论文配图:Aegix Pulse: A Traceable Three-Stage Architecture for Personalized Content Generation and Context-Preserving Revision
图 1 · 摘自论文原文
  • 分三阶段处理任务澄清、品牌档案构建与可控生成,全程追踪内容来源。
  • 修复时保留上下文可提升任务一致性0.29点(5分制),但统计不显著。
  • 适合对品牌连贯性要求高的企业级内容生成场景。

生产级内容生成系统需融合用户当前任务、长期品牌身份、历史证据及修改反馈。我们提出Aegix Pulse——一种面向生产的三阶段架构,将当前任务澄清与任务人格确定、长期账户档案(品牌基因)构建、受控生成与修订过程分离,并在各版本间保持溯源记录。通过96个合成社交媒体生成任务评估四个预注册假设。四种初始生成条件逐步引入任务人格、账户档案和成功历史风格证据;两种修订条件对比普通修订与上下文保持修订。实验生成480条完整生成记录与1,440次盲评大模型评分,辅以人工评审。相比仅使用任务人格,加入账户档案使平均品牌一致性得分提高0.1562分(5分制),但校正多重比较后p=0.1224,未达显著。上下文保持修订相比普通修订使平均任务保留得分提升0.2917分(校正后p=0.2432),亦未显著。任务人格本身有微弱效应,成功历史证据在当前设置下未带来额外提升。人工验证未能一致复现大模型评分趋势,且评审者间一致性低。结果为持续品牌上下文与上下文保持修订提供初步证据,同时指明更强证据处理与评估方法的改进方向。

原文摘要 · Abstract (English)

Production content-generation systems must integrate a user's immediate task, long-term brand identity, historical evidence, and revision feedback. We present Aegix Pulse, a production-oriented three-stage architecture that separates current-task clarification and Task Persona finalization, long-term Account Profile (Brand DNA) assembly, and controlled generation and revision while preserving provenance across content versions. We evaluate four preregistered claims using 96 synthetic social-media generation tasks. Four initial-generation conditions progressively introduced a Task Persona, Account Profile, and successful-history style evidence, while two revision conditions compared plain and context-preserving revision. The experiment produced 480 completed generation records and 1,440 blinded LLM-Judge evaluations, supplemented by human review. Adding the Account Profile increased mean brand-consistency scores by 0.1562 points on a five-point scale compared with Task Persona alone (Holm-adjusted p=.1224). Preserving task and brand context during revision increased mean task-preservation scores by 0.2917 points compared with plain revision (Holm-adjusted p=.2432). Neither improvement was statistically conclusive after multiple-comparison correction. Task Persona alone showed a small observed effect, while successful-history evidence provided no additional improvement in brand consistency under the current setting. Human validation did not consistently reproduce the LLM-Judge effect directions and showed low inter-reviewer agreement. These findings provide preliminary evidence for persistent brand context and context-preserving revision while identifying priorities for stronger evidence processing and evaluation.

内容生成品牌一致性上下文保留可追溯性

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。