为海关关税分类工具建立评估基准,推动贸易数字化
Benchmarking Harmonized Tariff Schedule Classification Models
- 借鉴语言模型评测思路,构建多维度评估框架
- 对比多家主流工具,揭示准确率与代码匹配度差异
- 适合跨境电商、报关系统开发者参考优化
海关关税分类行业对电子商务和国际贸易至关重要,但目前缺乏统一的分类解决方案评估标准。本研究受语言模型评测方法启发,为美国进口商品建立并测试了一个基准评估框架,用于系统比较主流的HTS分类工具。框架涵盖速度、准确率、合理性及HTS编码匹配度等关键指标,全面评估了Zonos、Tarifflo、Avalara和WCO BACUDA等领先方案的表现,识别出各工具的优势与局限。结果指明行业整体改进方向,为国际商贸与电商领域更高效、标准化的关税分类提供支持。
原文摘要 · Abstract (English)
The Harmonized Tariff System (HTS) classification industry, essential to e-commerce and international trade, currently lacks standardized benchmarks for evaluating the effectiveness of classification solutions. This study establishes and tests a benchmark framework for imports to the United States, inspired by the benchmarking approaches used in language model evaluation, to systematically compare prominent HTS classification tools. The framework assesses key metrics--such as speed, accuracy, rationality, and HTS code alignment--to provide a comprehensive performance comparison. The study evaluates several industry-leading solutions, including those provided by Zonos, Tarifflo, Avalara, and WCO BACUDA, identifying each tool's strengths and limitations. Results highlight areas for industry-wide improvement and innovation, paving the way for more effective and standardized HTS classification solutions across the international trade and e-commerce sectors.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。