统一GEC评估指标接口,让纠错系统对比更公平可靠。
gec-metrics: A Unified Library for Grammatical Error Correction Evaluation
- 提供统一接口,确保评估实现一致
- 支持元评估与可视化分析,便于指标改进
- 易扩展的API设计,适合研究者快速开发
我们提出gec-metrics,一个用于使用和开发语法错误修正(GEC)评估指标的统一库。该库通过统一接口实现评估标准的一致性,确保系统间比较的公平性。其设计强调API易用性,具备高度可扩展性,并内置元评估功能、分析脚本与可视化工具,助力评估指标的持续优化。代码以MIT许可证开源,也可作为可安装包使用,视频演示已发布于YouTube。
原文摘要 · Abstract (English)
We introduce gec-metrics, a library for using and developing grammatical error correction (GEC) evaluation metrics through a unified interface. Our library enables fair system comparisons by ensuring that everyone conducts evaluations using a consistent implementation. Moreover, it is designed with a strong focus on API usage, making it highly extensible. It also includes meta-evaluation functionalities and provides analysis and visualization scripts, contributing to developing GEC evaluation metrics. Our code is released under the MIT license and is also distributed as an installable package. The video is available on YouTube.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。