构建德语雇佣合同合法性公平性审查基准数据集
AI-assisted German Employment Contract Review: A Benchmark Dataset
- 构建匿名化标注的德语雇佣合同条款数据集
- 提供法律合规与公平性评估基线模型
- 助力法律AI辅助审查,适合法律科技研究者
雇佣合同在全球范围内用于约定雇主与雇员的工作条件。理解并审查合同中无效或不公平条款需要深厚的法律体系与术语知识。近年来自然语言处理(NLP)技术为辅助此类审查带来了希望。然而,由于缺乏专家标注的数据集,将NLP应用于法律文本尤其困难。为解决这一问题,并作为我们利用NLP辅助律师合同审查工作的起点,本文发布了一个匿名化且经过标注的德语雇佣合同条款合法性与公平性审查基准数据集,同时提供基线模型评估。
原文摘要 · Abstract (English)
Employment contracts are used to agree upon the working conditions between employers and employees all over the world. Understanding and reviewing contracts for void or unfair clauses requires extensive knowledge of the legal system and terminology. Recent advances in Natural Language Processing (NLP) hold promise for assisting in these reviews. However, applying NLP techniques on legal text is particularly difficult due to the scarcity of expert-annotated datasets. To address this issue and as a starting point for our effort in assisting lawyers with contract reviews using NLP, we release an anonymized and annotated benchmark dataset for legality and fairness review of German employment contract clauses, alongside with baseline model evaluations.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。