法律AI中检索增强生成的三大结构性缺陷及其解决方案
Beyond Probabilistic Similarity: Structural, Temporal, and Causal Limitations of Retrieval-Augmented Generation in the Legal Domain
- 从法律知识的层级、时间与因果结构出发,揭示检索机制的三重盲区
- 发现现有RAG系统在法律内容时效性、结构完整性和来源可追溯性上存在根本缺陷
- 提出面向法律推理的确定性架构设计原则,适合法律AI系统开发者与政策制定者
检索增强生成(RAG)已成为应对法律AI不可靠性的标准架构,但高调失败事件仍频发,包括提交给法院的虚构引用和将过时法律内容误标为现行有效。我们指出,这些并非仅因模型规模不足导致的幻觉,而是概率性检索与法律知识的层级性、时间动态性及制度溯源结构之间存在根本性不匹配所致。本文从三个层面展开:首先,基于古典法律理论,阐明法律知识的三项本体属性——层级与部分整体结构、受操作封闭性约束的时间动态演进、以及基于论证义务的制度溯源因果性;其次,识别出检索系统的三种病理:部分整体盲视、时间盲视与因果不透明,每种均提供操作定义、失效机制、典型例证及诊断标准;最后,以该框架审视当前技术进展,发现现有方法对三方面要求处理不均,尚未形成协同一致的范式。由此提炼出四项法律检索的确定性设计原则:本体优先、事件实体化、双时间正确性、确定性交互协议。该框架聚焦于‘适用何法及处于何种状态’的法律问题,主要面向立法与宪法检索,解释时间亦作为明确延伸。
原文摘要 · Abstract (English)
Retrieval-Augmented Generation (RAG) has become a standard architectural response to unreliability in legal AI, yet high-profile failures, including fabricated citations submitted to courts and anachronistic legal content presented as current, continue to appear across jurisdictions. We argue that these failures are not residual confabulations to be eliminated by scaling language models, but symptoms of an architectural mismatch between probabilistic retrieval and the hierarchical, temporal, and institutional structure of legal knowledge. We develop the argument in three moves. First, we articulate the ontological commitment of legal knowledge as a triad of properties derivable from classical legal theory: hierarchical and mereological structure, diachronic dynamism under operational closure, and causal traceability of institutional provenance grounded in the duty of justification. Second, we identify three corresponding pathologies of retrieval (mereological blindness, diachronic blindness, and causal opacity), each developed with an operational definition, a failure mechanism, a canonical example, and detection criteria for diagnostic use. Third, we review the state of the art through this lens, showing that existing approaches address these requirements unevenly and do not yet compose into a paradigm that treats them as co-constitutive. From this analysis we derive four architectural commitments that characterize the deterministic-by-design direction for legal retrieval: ontological primacy, event reification, bitemporal correctness, and deterministic interaction protocols. The framework concerns quaestio juris (which norms apply and in what state) rather than the downstream tasks that act on identified norms, and addresses legislative and constitutional retrieval primarily, with interpretive time as an explicit extension.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。