arXiv:2507.07116cs.DCcs.AI2025-07

对比三类分布式账本存储语义数据,发现私有链最高效。

Analysing semantic data storage in Distributed Ledger Technologies for Data Spaces

  • 用真实知识图谱测试公有、私有、混合链的语义数据存储性能。
  • 私有链在存储效率和资源消耗上最优,混合链兼顾审计与效率。
  • 适合关注数据主权与系统性能平衡的研究者和开发者参考。

数据空间正成为多参与方之间实现自主、安全、可信数据交换的去中心化基础设施。为实现语义互操作性,语义网技术与知识图谱被提出应用。尽管分布式账本技术(DLT)适合作为数据空间的基础架构,但其在语义数据高效存储方面仍存在显著差距。本文基于真实世界知识图谱,对公有、私有及混合型DLT进行了系统的语义数据存储评估,比较了性能、存储效率、资源消耗以及更新与查询能力。结果表明,私有DLT在存储与管理语义内容方面最为高效,而混合DLT在公共可审计性与运行效率之间提供了良好平衡。研究进一步讨论了根据去中心化数据生态中数据主权需求选择合适DLT基础设施的策略。

原文摘要 · Abstract (English)

Data spaces are emerging as decentralised infrastructures that enable sovereign, secure, and trustworthy data exchange among multiple participants. To achieve semantic interoperability within these environments, the use of semantic web technologies and knowledge graphs has been proposed. Although distributed ledger technologies (DLT) fit as the underlying infrastructure for data spaces, there remains a significant gap in terms of the efficient storage of semantic data on these platforms. This paper presents a systematic evaluation of semantic data storage across different types of DLT (public, private, and hybrid), using a real-world knowledge graph as an experimental basis. The study compares performance, storage efficiency, resource consumption, and the capabilities to update and query semantic data. The results show that private DLTs are the most efficient for storing and managing semantic content, while hybrid DLTs offer a balanced trade-off between public auditability and operational efficiency. This research leads to a discussion on the selection of the most appropriate DLT infrastructure based on the data sovereignty requirements of decentralised data ecosystems.

数据空间分布式账本知识图谱语义存储

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。