评测越中、越老双向机器翻译系统,融合自动与人工评估。
ViBidirectionMT-Eval: Machine Translation for Vietnamese-Chinese and Vietnamese-Lao language pair
- 构建越中、越老双向翻译模型,覆盖4种语言方向。
- 在1000对新闻与通用语料上测试,使用BLEU与SacreBLEU指标。
- 引入双语专家人工评估,提升模型性能排名可靠性。
本文报告了VLSP 2022-2023机器翻译共享任务的成果,聚焦越南语-中文和越南语-老挝语的机器翻译。该任务是第9届和第10届越南语言与语音处理研讨会(VLSP 2022, VLSP 2023)的一部分。目标是构建针对越南语-中文和越南语-老挝语翻译的机器翻译系统(对应4种翻译方向)。提交系统在1000对测试数据(新闻与通用领域)上进行评估,采用BLEU [11] 和 SacreBLEU [12] 等标准指标。此外,系统输出还由中文和老挝语专家进行人工判断,人工评估在模型性能排序中发挥关键作用,确保评估更全面。
原文摘要 · Abstract (English)
This paper presents an results of the VLSP 2022-2023 Machine Translation Shared Tasks, focusing on Vietnamese-Chinese and Vietnamese-Lao machine translation. The tasks were organized as part of the 9th, 10th annual workshop on Vietnamese Language and Speech Processing (VLSP 2022, VLSP 2023). The objective of the shared task was to build machine translation systems, specifically targeting Vietnamese-Chinese and Vietnamese-Lao translation (corresponding to 4 translation directions). The submission were evaluated on 1,000 pairs for testing (news and general domains) using established metrics like BLEU [11] and SacreBLEU [12]. Additionally, system outputs also were evaluated with human judgment provided by experts in Chinese and Lao languages. These human assessments played a crucial role in ranking the performance of the machine translation models, ensuring a more comprehensive evaluation.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。