DocSpiral让人工标注文档更高效,边标边训练模型,逐步减少人工干预。
DocSpiral: A Platform for Integrated Assistive Document Annotation through Human-in-the-Spiral
- 采用螺旋迭代设计,人标注数据训练模型,模型反过来减轻后续标注负担。
- 实验显示标注时间减少至少41%,三轮迭代后模型性能持续提升。
- 适合需要处理扫描件文档的科研与医疗领域,支持快速构建AI模型。
从领域特定的图像文档(如扫描报告)中获取结构化数据对下游任务至关重要,但受文档多样性影响仍具挑战性。许多文档仅以图像形式存在,需人工标注来训练自动化提取系统。我们提出首个面向图像文档的Human-in-the-Spiral辅助标注平台DocSpiral,旨在解决从特定领域图像文档中提取结构化信息的问题。其螺旋式设计建立了一个迭代循环:人工标注用于训练模型,模型性能提升后逐步减少人工介入。DocSpiral集成文档格式标准化、全面的标注界面、评估指标仪表盘及模型开发API接口,形成统一工作流。实验表明,该框架在模型训练的三轮迭代中,标注时间减少至少41%,且性能持续提升。通过免费开放系统,我们希望降低文档处理领域人工智能/机器学习模型开发门槛,推动大语言模型在地质科学和医疗等图像密集型领域的应用。系统已公开:https://app.ai4wa.com。演示视频:https://app.ai4wa.com/docs/docspiral/demo。
原文摘要 · Abstract (English)
Acquiring structured data from domain-specific, image-based documents such as scanned reports is crucial for many downstream tasks but remains challenging due to document variability. Many of these documents exist as images rather than as machine-readable text, which requires human annotation to train automated extraction systems. We present DocSpiral, the first Human-in-the-Spiral assistive document annotation platform, designed to address the challenge of extracting structured information from domain-specific, image-based document collections. Our spiral design establishes an iterative cycle in which human annotations train models that progressively require less manual intervention. DocSpiral integrates document format normalization, comprehensive annotation interfaces, evaluation metrics dashboard, and API endpoints for the development of AI / ML models into a unified workflow. Experiments demonstrate that our framework reduces annotation time by at least 41\% while showing consistent performance gains across three iterations during model training. By making this annotation platform freely accessible, we aim to lower barriers to AI/ML models development in document processing, facilitating the adoption of large language models in image-based, document-intensive fields such as geoscience and healthcare. The system is freely available at: https://app.ai4wa.com. The demonstration video is available: https://app.ai4wa.com/docs/docspiral/demo.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。