Charles Explorer logo
🇬🇧

Visualizer of Dataset Similarity Using Knowledge Graph

Publication at Faculty of Mathematics and Physics |
2020

Abstract

Many institutions choose to make their datasets available as Open Data. Open Data datasets are described by publisher-provided metadata and are registered in catalogs such as the European Data Portal.

In spite of that, findability still remain a major issue. One of the main reasons is that metadata is captured in different contexts and with different background knowledge, so that keyword-based search provided by the catalogs is insufficient.

A solution is to use an enriched querying that employs a dataset similarity model built on a shared context represented by a knowledge graph. However, the "black-box" dataset similarity may not fit well the user needs.

If an explainable similarity model is used, then the issue can be tackled by providing users with a visualisation of the dataset similarity. This paper introduces a web-based tool for dataset similarity visualisation called ODIN (Open Dataset INspector).

ODIN visualises knowledge graph-based dataset similarity, offering thus an explanation to the user. To understand the similarity, users can discover additional datasets that match their needs or reformulate the query to better reflect the knowledge graph.

Last but not least, the user can analyze and/or design the similarity model itself.