The term Data Space, understood as the secure exchange of data in distributed systems, ensuring openness, transparency, decentralization, sovereignty, and interoperability of information, has gained importance during the last years. However, Data Spaces are in an initial phase of definition, and new research is necessary to address their requirements. The Open Data ecosystem can be understood as one of the precursors of Data Spaces as it provides mechanisms to ensure the interoperability of information through resource discovery, information exchange, and aggregation via metadata. However, Data Spaces require more advanced capabilities including the automatic and scalable generation and publication of high-quality metadata. In this work, we present a set of software tools that facilitate the automatic generation and publication of metadata, the modeling of datasets through standards, and the assessment of the quality of the generated metadata. We validate all these tools through the YODA Open Data Portal showing how they can be connected to integrate Open Data into Data Spaces.
翻译:术语“数据空间”被理解为分布式系统中确保信息开放性、透明性、去中心化、主权及互操作性的安全数据交换,近年来其重要性日益凸显。然而,数据空间尚处于定义初期阶段,需要开展新的研究以满足其需求。开放数据生态系统可被视为数据空间的先驱之一,因其通过资源发现、信息交换及基于元数据的聚合,提供了确保信息互操作性的机制。但数据空间需要更高级的能力,包括元数据的自动化、可扩展生成与发布。本文提出了一套软件工具,用于实现元数据的自动生成与发布、数据集的标准建模以及生成元数据质量的评估。我们通过YODA开放数据门户验证了所有工具,展示了如何将这些工具连接起来,将开放数据集成至数据空间中。