针对法律文档长篇问答难题,提出可解析复杂结构的智能问答系统。
Long-Context Long-Form Question Answering for Legal Domain
- 拆解法律专有词汇,提升文档检索精度
- 解析嵌套章节与脚注,准确关联内容结构
- 生成精准覆盖全篇的长篇回答,适合法律从业者使用
法律文件具有复杂的文档布局,包含多层嵌套章节、冗长脚注,并运用复杂的语法和领域专用词汇以确保精确性与权威性。这些特性使问答任务极具挑战,尤其是当答案跨越多页(需长上下文)且要求全面(长篇回答)时。本文针对法律领域的长上下文长篇问答问题,提出一个问答系统:(a) 拆解领域专有词汇以提升源文档检索效果;(b) 解析复杂文档布局,准确分离章节与脚注并合理关联;(c) 使用精准的领域词汇生成综合性答案。此外,我们引入一种覆盖率度量,将性能按召回率分类,便于人工评估。基于法律与企业税务专业人士的协作,构建了一个QA数据集。通过全面实验与消融研究,验证了该系统的可用性与有效性。
原文摘要 · Abstract (English)
Legal documents have complex document layouts involving multiple nested sections, lengthy footnotes and further use specialized linguistic devices like intricate syntax and domain-specific vocabulary to ensure precision and authority. These inherent characteristics of legal documents make question answering challenging, and particularly so when the answer to the question spans several pages (i.e. requires long-context) and is required to be comprehensive (i.e. a long-form answer). In this paper, we address the challenges of long-context question answering in context of long-form answers given the idiosyncrasies of legal documents. We propose a question answering system that can (a) deconstruct domain-specific vocabulary for better retrieval from source documents, (b) parse complex document layouts while isolating sections and footnotes and linking them appropriately, (c) generate comprehensive answers using precise domain-specific vocabulary. We also introduce a coverage metric that classifies the performance into recall-based coverage categories allowing human users to evaluate the recall with ease. We curate a QA dataset by leveraging the expertise of professionals from fields such as law and corporate tax. Through comprehensive experiments and ablation studies, we demonstrate the usability and merit of the proposed system.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。