<?xml version="1.0" encoding="utf-8" standalone="yes" ?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Posts | Sven Lieber</title>
    <link>https://sven-lieber.org/en/post/</link>
      <atom:link href="https://sven-lieber.org/en/post/index.xml" rel="self" type="application/rss+xml" />
    <description>Posts</description>
    <generator>Wowchemy (https://wowchemy.com)</generator><language>en-us</language><lastBuildDate>Wed, 29 Oct 2025 09:00:00 +0100</lastBuildDate>
    <image>
      <url>https://sven-lieber.org/media/icon_hufdd866d90d76849587aac6fbf27da1ac_464_512x512_fill_lanczos_center_3.png</url>
      <title>Posts</title>
      <link>https://sven-lieber.org/en/post/</link>
    </image>
    
    <item>
      <title>ENDORSE 2025</title>
      <link>https://sven-lieber.org/en/2025/10/29/endorse-2025/</link>
      <pubDate>Wed, 29 Oct 2025 09:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2025/10/29/endorse-2025/</guid>
      <description>&lt;p&gt;The European Economic Area is one of the best use cases for (semantic) data integration:
the many different languages, legislations, rules and systems
must be or become interoperable!
On October 8 and 9, 2025, I attended
the &lt;strong&gt;3rd European Data Conference on Reference Data and Semantics (ENDORSE)&lt;/strong&gt;
in the European Commission&amp;rsquo;s building Charlemagne in Brussels, Belgium.
In this blog post I will reflect on the conference,
provide details from some of the presentations,
and talk about my own presentation in the Cultural Heritage session.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;Several hundred people participated in this year&amp;rsquo;s ENDORSE conference in person
and many watched the sessions via an online livestream.
A total of 85 speakers participated in 45 accepted presentations (out of 61 submissions),
8 posters, 2 panels and 3 keynotes.&lt;/p&gt;
&lt;p&gt;The &lt;a href=&#34;https://ror.org/016jmhr98&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Publications Office of the European Union&lt;/a&gt;,
which organized the conference
already highlighted several key messages
in a LinkedIn post.&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;(1) Data, semantics and AI must work hand in hand to build services that are truly citizen- and business-centric.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;blockquote&gt;
&lt;p&gt;(2) The power of AI depends on the quality, precision and multilingual nature of the data it learns from, a uniquely European strength.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;blockquote&gt;
&lt;p&gt;(3) Semantic interoperability used to be a niche field. Today it’s a strategic enabler of Europe’s digital transformation.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;blockquote&gt;
&lt;p&gt;(4) Collaboration, co-creation and trust remain the foundations of an interoperable and inclusive digital ecosystem.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;blockquote&gt;
&lt;p&gt;(5) Execution matters! Turning insights and knowledge into concrete actions is how we deliver meaningful impact.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;blockquote&gt;
&lt;p&gt;(6) A key takeaway? Reference data is the glue that binds all these areas, a message that resonated strongly throughout our vibrant networking sessions.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;I can only agree!
This conference was a great networking event for the &lt;em&gt;GovTech&lt;/em&gt; sector
where Semantic Web standards and technologies are put into practice.&lt;/p&gt;
&lt;p&gt;I keep this post short and only focus on my own takeaways, a few details from some of the presentations
and my own contribution related to the use of Wikibase in the MetaBelgica project.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;My main takeaways&lt;/strong&gt;&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;
&lt;p&gt;From vocabularies to application profiles: &lt;strong&gt;governance matters&lt;/strong&gt;&lt;/p&gt;
&lt;/li&gt;
&lt;li&gt;
&lt;p&gt;&lt;strong&gt;Linked Data meets Generative AI meets data spaces&lt;/strong&gt;: existing structured knowledge can be offered via the MCP protocol to generative AI models, but also well-documented specifications make it more human and machine-friendly&lt;/p&gt;
&lt;/li&gt;
&lt;li&gt;
&lt;p&gt;The professional conference setup really made a difference: professional moderators, time keepers and online streaming all made the conference visit a nice experience&lt;/p&gt;
&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;One downside of the conference: the speaker&amp;rsquo;s presentations, one of the main outcomes of the conference,
are only available via links on the ENDORSE website
and hence are not FAIR.
I would have loved to cite all the presentations that I mention with a DOI.
Providing all presentations via e.g. a Zenodo community (like done for the &lt;a href=&#34;https://zenodo.org/communities/cordi-2025/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;CoRDI conference&lt;/a&gt;)
would have been a great solution to make the outcome of ENDORSE FAIR.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;data-stewardship-in-the-age-of-ai&#34;&gt;Data stewardship in the Age of AI&lt;/h2&gt;
&lt;p&gt;In his keynote about the future of data stewardship,
&lt;a href=&#34;https://orcid.org/0000-0002-0107-2984&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Stefaan Verhulst&lt;/a&gt;
from &lt;a href=&#34;https://thegovlab.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;the Govlab&lt;/a&gt;
covered many interesting topics around what he called the
&lt;em&gt;emergence of open data winter while having an AI summer&lt;/em&gt;.&lt;/p&gt;
&lt;p&gt;According to Stefaan we witness the emergence of an open data winter:
access to data is backsliding, so is the excitement for open gov data
and even Creative Commons sees a decline in licensing.
A general GenAI anxiety is noticeable,
yet there is hope: we are able to ride the fourth wave of Open Data.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The fourth wave of open data: Generative AI&#34; srcset=&#34;
               /media/endorse-2025/2025-10-09_endorse-fourth-wave-open-data_hu04a040fe607b2e5f19abd35dfb300a86_228596_8e35185b30b13dc23f046db99d4ec12f.webp 400w,
               /media/endorse-2025/2025-10-09_endorse-fourth-wave-open-data_hu04a040fe607b2e5f19abd35dfb300a86_228596_0ad65e15797f1de9b7c443d55ae6a78b.webp 760w,
               /media/endorse-2025/2025-10-09_endorse-fourth-wave-open-data_hu04a040fe607b2e5f19abd35dfb300a86_228596_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/endorse-2025/2025-10-09_endorse-fourth-wave-open-data_hu04a040fe607b2e5f19abd35dfb300a86_228596_8e35185b30b13dc23f046db99d4ec12f.webp&#34;
               width=&#34;760&#34;
               height=&#34;426&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In a nutshell,
this inspiring keynote
introduced the following six shifts in data stewardship,
to become more strategic and prevent an open data winter.&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;&lt;strong&gt;supply =&amp;gt; demand&lt;/strong&gt; (become demand-centric, key questions regarding a problem and need for evidence for decision makers)&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;datasets =&amp;gt; Knowledge Graphs&lt;/strong&gt; for interoperability&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;structured data =&amp;gt; unstructured data&lt;/strong&gt; (Large Language Models like unstructured data, but it was never part of data management)&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;consent =&amp;gt; social license&lt;/strong&gt; (to engage with communities, current consent is flawed in multiple ways)&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;FAIR =&amp;gt; FAIR-R&lt;/strong&gt; (AI-ready, authoritative provenance, model context protocol (connection from AI to external systems))&lt;/li&gt;
&lt;li&gt;&lt;strong&gt;Open Data =&amp;gt; data commons&lt;/strong&gt; (GenAI anxiety, common-crawl + wikipedia fuels model training to a large extent, following the theories of Elenor Ostrom: govern resources in ways that are not extractive and align with principles and expectations of those that created the data or make it accessible, data commons prices)&lt;/li&gt;
&lt;/ol&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of Stefaan&amp;rsquo;s presentation is available &lt;a href=&#34;https://www.youtube.com/watch?v=CTwLPpXCvWo&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=33&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
  &lt;/div&gt;
&lt;/div&gt;
&lt;p&gt;Stefaan’s talk looked into problems at the horizon
and how we could deal with them.
Data governance and openness
are not static goals but evolving practices.
The next presentations showed how these principles are already
being applied in concrete public sector contexts.&lt;/p&gt;
&lt;h2 id=&#34;interoperable-public-services-and-data-spaces&#34;&gt;Interoperable Public Services and Data Spaces&lt;/h2&gt;
&lt;p&gt;In her keynote,
&lt;a href=&#34;https://no.linkedin.com/in/kjersti-steien-0019765b&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Kjersti Steien&lt;/a&gt;
from the &lt;a href=&#34;https://www.digdir.no/digdir/about-norwegian-digitalisation-agency/887&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Norwegian Digitalisation Agency (Digdir)&lt;/a&gt;
talked about the implementation of Public Services in Norway.&lt;/p&gt;
&lt;p&gt;Similar to other European countries,
Norway also have a catalog of public services.
Even though Norway is not a member of the EU,
they use the &lt;em&gt;Core Public Service Vocabulary Application Profile&lt;/em&gt; (CPSV-AP) of the EU.&lt;/p&gt;
&lt;p&gt;More specifically,
their service catalog is described in
the Norwegian extension CPSV-AP-NO
and can also be connected
to relevant concepts described with SKOS-AP-NO,
and datasets according to the DCAT-AP-NO model.
A nice detail: the Norwegian CPSV extension also includes services delivered by the private sector.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Current challenges for Digdir in Norway&#34; srcset=&#34;
               /media/endorse-2025/2025-10-09_endorse-norway_hub8d9901e3c544c3891d8e39e672c84ca_69932_fb8f2df5b81b9349601a5bf7cb3f0b2a.webp 400w,
               /media/endorse-2025/2025-10-09_endorse-norway_hub8d9901e3c544c3891d8e39e672c84ca_69932_b6570493651b54a599f139835ca8f0f7.webp 760w,
               /media/endorse-2025/2025-10-09_endorse-norway_hub8d9901e3c544c3891d8e39e672c84ca_69932_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/endorse-2025/2025-10-09_endorse-norway_hub8d9901e3c544c3891d8e39e672c84ca_69932_fb8f2df5b81b9349601a5bf7cb3f0b2a.webp&#34;
               width=&#34;760&#34;
               height=&#34;395&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;This example shows how EU application profiles
can provide a shared basis even beyond EU borders.
At the same time, national approaches still vary quite a lot.&lt;/p&gt;
&lt;p&gt;During the conference I have learned,
that also Germany has a catalog of public services:
the &lt;a href=&#34;https://fimportal.de&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FIM portal&lt;/a&gt; (presentation by &lt;a href=&#34;https://orcid.org/0000-0001-6423-7427&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Felicitats Löffler&lt;/a&gt; and &lt;a href=&#34;https://orcid.org/0000-0003-1478-1867&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marianne Mauch&lt;/a&gt; from the state of Thuringia, the &lt;em&gt;green heart of Germany&lt;/em&gt;).
However, it is not derived from CPSV-AP
and I don&amp;rsquo;t think it is mapped to it yet.
This brings me to an interesting point
mentioned by Kjersti:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;It would be great if the EU would recommend CPSV-AP as one of the specifications to use with the Single Digital Gateway&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;These different implementations underline that even with good standards,
alignment still depends on manual curation and coordination between actors.&lt;/p&gt;
&lt;p&gt;Overall the work on public service descriptions
makes me think of two things:
first, how important manual data curation is.
Secondly, on some work I did a few years ago in the FAST project,
on providing personalized workflows for life events
to improve the customer journey (&lt;a href=&#34;https://doi.org/10.5281/zenodo.8001639&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;presentation&lt;/a&gt;)&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of Kjersti&amp;rsquo;s presentation is available &lt;a href=&#34;https://www.youtube.com/watch?v=TKUBX7-Dbwo&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=40&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
and the recording of Felicitas and Marianne &lt;a href=&#34;https://www.youtube.com/watch?v=f7Jkiaa4i5E&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=35&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
  &lt;/div&gt;
&lt;/div&gt;
&lt;p&gt;After several talks on interoperability across administrations,
the next keynote (which was the first of the conference actually)
focused on Knowledge systems in general
and what it means to make data understandable for both humans and machines.&lt;/p&gt;
&lt;h2 id=&#34;provenance-and-agreement&#34;&gt;Provenance and Agreement&lt;/h2&gt;
&lt;p&gt;In his keynote,
&lt;a href=&#34;https://orcid.org/0000-0003-0183-6910&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Paul Groth&lt;/a&gt;
from &lt;a href=&#34;https://ror.org/008xxew50&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;VU Amsterdam&lt;/a&gt;
(co-author of the W3C PROV standard
and the FAIR data principles)
covered several topics related to Knowledge Engineering
and Large Language Models (LLMs).&lt;/p&gt;
&lt;p&gt;One interesting thing that I&amp;rsquo;ve learned about in Paul&amp;rsquo;s talk is C2PA,
a global standard for content authenticity,
that recently was integrated into Android.
Basically it adds &lt;em&gt;Content Credentials&lt;/em&gt;
via cryptographical methods to digital media
and thus the necessary &lt;strong&gt;provenance&lt;/strong&gt; to
verify that a certain image or video was created and not generated by AI.
Something similar I&amp;rsquo;ve seen at the poster of the &lt;a href=&#34;https://doc.piveau.io/hub/piveau-x/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Piveau-X catalog&lt;/a&gt;,
uploaded content needs to be cryptographically verified.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Paul Groth presenting a graph about Knowledge Engineering Requirements over time&#34; srcset=&#34;
               /media/endorse-2025/2025-10-08_endorse-knowledge-engineering_hu28d5f4fecb7a91abe4e42a078cfb00ca_4073677_37608207ec68b3c274214f2358f4a190.webp 400w,
               /media/endorse-2025/2025-10-08_endorse-knowledge-engineering_hu28d5f4fecb7a91abe4e42a078cfb00ca_4073677_c38d36be0d6e8327250522ef2198c6c2.webp 760w,
               /media/endorse-2025/2025-10-08_endorse-knowledge-engineering_hu28d5f4fecb7a91abe4e42a078cfb00ca_4073677_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/endorse-2025/2025-10-08_endorse-knowledge-engineering_hu28d5f4fecb7a91abe4e42a078cfb00ca_4073677_37608207ec68b3c274214f2358f4a190.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Paul also talked about how important,
but also how time consuming,
it is to work on standards and find agreement.
He gave the example of the &lt;a href=&#34;https://www.w3.org/TR/prov-overview/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;W3C PROV specification&lt;/a&gt;,
where he and others were involved in
in 8820 public E-mails,
666 issues,
600 Wiki pages,
6000 mercurial commits
and 152 teleconferences.&lt;/p&gt;
&lt;p&gt;What I take home from this talk:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;Make standards more human readable, so they are more machine (LLM) readable&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;Adding more documentation to your specifications and data models is never a bad idea :-)&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of Paul&amp;rsquo;s presentation is available &lt;a href=&#34;https://www.youtube.com/watch?v=HPKpVXAep1Q&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=4&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
  &lt;/div&gt;
&lt;/div&gt;
&lt;hr&gt;
&lt;p&gt;Paul’s talk highlighted the effort that goes into developing common standards
and the importance of documenting them well.
The following presentation by
&lt;a href=&#34;https://be.linkedin.com/in/bert-van-nuffelen-a349634&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Bert Van Nuffelen&lt;/a&gt;
from &lt;a href=&#34;https://www.vlaanderen.be/digitaal-vlaanderen/over-ons&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Digitaal Vlaanderen&lt;/a&gt;
gave a concrete example about Knowledge Engineering
of the &lt;em&gt;Open Standards for Linked Organizations&lt;/em&gt; (OSLO)
in Flanders, Belgium.&lt;/p&gt;
&lt;p&gt;He highlighted the fact that reuse of knowledge is important,
but also that it can lead to questions regarding data governance.
For example, certain constraints should not be defined
directly on the concepts in the namespace of a vocabulary,
but rather in the application profile
that reuses terms of this vocabulary.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Data governance issue when reusing specifications: an illustration from the Flemish OSLO standards&#34; srcset=&#34;
               /media/endorse-2025/2025-10-08_endorse-oslo_hua43e628f2c7e97d65a435c4da1d6e2f1_95455_0010c657696fe5edb4f835e55f15694c.webp 400w,
               /media/endorse-2025/2025-10-08_endorse-oslo_hua43e628f2c7e97d65a435c4da1d6e2f1_95455_342983769d8728971ac627502c8db21a.webp 760w,
               /media/endorse-2025/2025-10-08_endorse-oslo_hua43e628f2c7e97d65a435c4da1d6e2f1_95455_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/endorse-2025/2025-10-08_endorse-oslo_hua43e628f2c7e97d65a435c4da1d6e2f1_95455_0010c657696fe5edb4f835e55f15694c.webp&#34;
               width=&#34;760&#34;
               height=&#34;420&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;However, how does the governance of an application profile
actually works in practice?
After all, the reused terms are governed
in vocabularies of different namespaces.
The emergence of &lt;em&gt;data spaces&lt;/em&gt; aggravates the situation,
because several data spaces can make use
of a mix of different application profiles and vocabulary terms.&lt;/p&gt;
&lt;p&gt;The OSLO presentation did not stop at data governance.
It also touched on how AI systems can interact directly with these well-defined semantics.&lt;/p&gt;
&lt;p&gt;OSLO allows language models to query its structured descriptions
through a Model Context Protocol (MCP) server,
which is an interesting step towards connecting Linked Data with generative AI.
I first learned about MCP at this conference.
While Paul Groth advocated for more machine-readable standards,
the MCP approach demonstrates how machines can use them in practice.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of Bert&amp;rsquo;s presentation is available &lt;a href=&#34;https://www.youtube.com/watch?v=W1txkBe_UKo&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=28&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
  &lt;/div&gt;
&lt;/div&gt;
&lt;hr&gt;
&lt;p&gt;Talking about data spaces,
they were another recurring topic at this year’s ENDORSE.
In one of the panels it was said that, in theory,
you could even run a data space on paper.
This highlights how much the concept still depends on governance rather than technology.
So far there is not one single standard, but several coexisting initiatives and implementations.&lt;/p&gt;
&lt;p&gt;During the discussion, moderated by &lt;a href=&#34;https://orcid.org/0000-0001-6917-2167&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Pieter Colpaert&lt;/a&gt;,
one speaker mentioned the browser wars of the 1990s:
many competing implementations, not all interoperable yet.
The International Data Spaces Association (IDSA) was mentioned as one of the key players,
and the European Commission already recognizes more than 70 &lt;em&gt;common European data spaces&lt;/em&gt;
even if it is not always clear to me what qualifies as one.&lt;/p&gt;
&lt;p&gt;What I found particularly interesting was the perspective on procurement.
Instead of one large platform
where only big companies can compete,
data spaces could allow smaller providers to offer interoperable services.
The sovereignty of participants and the absence of monopolies
were highlighted as essential principles.
The main recommendation from the panel:
don’t make the definition of a data space too strict,
it’s better to have working examples than theoretical perfection.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of the data spaces panel is available &lt;a href=&#34;https://www.youtube.com/watch?v=f_DPml-yKW8&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=41&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;
  &lt;/div&gt;
&lt;/div&gt;
&lt;p&gt;These discussions about interoperability,
governance and shared infrastructures
also resonated strongly in the cultural heritage session,
where I presented our own approach to data management in the MetaBelgica project.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;metabelgica-central-vs-decentralized-data-management&#34;&gt;MetaBelgica: central vs decentralized data management&lt;/h2&gt;
&lt;p&gt;This year&amp;rsquo;s ENDORSE also featured a session on cultural heritage.
Besides presentations about DE-BIAS from
&lt;a href=&#34;https://orcid.org/0000-0001-5286-1388&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Orfeas Menis Mastromichalakis&lt;/a&gt;,
the Dutch data space for cultural heritage
from &lt;a href=&#34;https://orcid.org/0000-0002-2884-3523&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Enno Meijers&lt;/a&gt;
and a case study for historical encyclopedias
from &lt;a href=&#34;https://orcid.org/0000-0002-3731-6397&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Thora Hagen&lt;/a&gt;,
I presented work of our &lt;a href=&#34;https://www.kbr.be/en/projects/metabelgica/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;MetaBelgica&lt;/a&gt; project.&lt;/p&gt;
&lt;p&gt;In particular, my presentation focused about our choices on data management.
On the one hand,
due to existing legacy systems
and duplicate efforts in data curation,
we opted to &lt;strong&gt;not have a fully decentralized solution&lt;/strong&gt;.&lt;/p&gt;
&lt;p&gt;On the other hand,
we also did &lt;strong&gt;not choose a fully centralized solution&lt;/strong&gt;,
to avoid having a single-point of failure system
in which partners from several institutions have to maintain all their data.&lt;/p&gt;
&lt;p&gt;Instead we have chosen for a Wikibase system
that I situate in-between with a best-of-both worlds approach:
&lt;strong&gt;Loosely coupled FAIR data&lt;/strong&gt;.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The advantages of using a Wikibase for MetaBelgica&#34; srcset=&#34;
               /media/endorse-2025/2025-10-08_endorse-metabelgica_huc849e16af550dcbb3f52afb91b18a8c8_175071_771a4f4d45340fbd23c9504fbe8f8877.webp 400w,
               /media/endorse-2025/2025-10-08_endorse-metabelgica_huc849e16af550dcbb3f52afb91b18a8c8_175071_e318c4163c96fc3cb4c15a80a1e06e74.webp 760w,
               /media/endorse-2025/2025-10-08_endorse-metabelgica_huc849e16af550dcbb3f52afb91b18a8c8_175071_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/endorse-2025/2025-10-08_endorse-metabelgica_huc849e16af550dcbb3f52afb91b18a8c8_175071_771a4f4d45340fbd23c9504fbe8f8877.webp&#34;
               width=&#34;760&#34;
               height=&#34;430&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;We rely on the mature data management software Wikibase (with Technology Readiness Level 9),
while keeping a balance of control and collaboration.
For example, we keep local autonomy
and possible additional institution-specific data,
while providing de-duplicated high quality FAIR reference data
in a single trustworthy system with Persistent Identifiers for the public.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    The recording of my presentation is available &lt;a href=&#34;https://www.youtube.com/watch?v=rrhf9HNI9CI&amp;amp;list=PLT5rARDev_rmgK8ZddP7p3oKFaj5A2whJ&amp;amp;index=31&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt; and my slides via the following DOI: &lt;a href=&#34;https://doi.org/10.5281/zenodo.17344191&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.17344191&lt;/a&gt;.
  &lt;/div&gt;
&lt;/div&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;ENDORSE 2025 once again showed how Europe combines technical precision with collaboration.
From open data stewardship and interoperable public services
to AI-ready standards,
the conference made clear that the future of digital government
depends on connecting structured knowledge and the people maintaining it.
This post only covered a few talks that I found particularly interesting,
there are many more presentations (and recordings)
to be discovered on the &lt;a href=&#34;https://op.europa.eu/en/web/endorse-2025/programme#&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;ENDORSE website&lt;/a&gt;.&lt;/p&gt;
&lt;h2 id=&#34;references&#34;&gt;References&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;Lieber, S. (2019, June 25). The FAST project: Criterion and Evidence from the Public Service UX perspective. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/zenodo.8001639&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.8001639&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Lieber, S. (2025, October 8). Opening up Belgian Cultural Heritage Reference Data with Wikibase. The European Data Conference on Reference Data And Semantics (ENDORSE), Brussels, Belgium. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/zenodo.17344191&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.17344191&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;</description>
    </item>
    
    <item>
      <title>CoRDI 2025</title>
      <link>https://sven-lieber.org/en/2025/09/07/cordi-2025/</link>
      <pubDate>Sun, 07 Sep 2025 09:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2025/09/07/cordi-2025/</guid>
      <description>&lt;p&gt;In the old days, research was done by famous individuals.
Nowadays, research is a team effort!
Working together requires
making agreements about how to collaborate
and using common tools.
These kinds of agreements (or standards) and tools
are especially crucial given the amount of data nowadays.
Me and around 700 other people attended the &lt;strong&gt;2nd Conference on Research Data Infrastructure&lt;/strong&gt;,
in Aachen, Germany from August 26 to 28, 2025.
In this blog post I will reflect
on the various presentations and posters I saw
as well as the discussions I had.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;In August 2025,
the &lt;a href=&#34;https://ror.org/05qj6w324&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;&lt;em&gt;German National Research Data Infrastructure (NFDI)&lt;/em&gt;&lt;/a&gt;
organized the second edition of the
Conference on Research Data Infrastructure (CoRDI).
Around &lt;strong&gt;700 people&lt;/strong&gt; attended this international multi-track conference.
This year&amp;rsquo;s edition did not follow a standard peer-review,
but was curated based on the conference organizers
and a balance of the various disciplines and topics
in the field of research data infrastructure.&lt;/p&gt;
&lt;p&gt;Even though my contribution was not selected,
I anyway attended the conference as this is a great chance to network.
Next to the small things,
such as
live music from the RWTH Aachen big band,
the daily buffet with default vegetarian/vegan options,
and exciting robot demos at the social event,
the conference featured Research Data Management (RDM) in various disciplines.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;My main takeaways&lt;/strong&gt;&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;
&lt;p&gt;&lt;strong&gt;Standards and identifiers are the hidden champions&lt;/strong&gt;: Whether in workflows, Knowledge Graphs, or infrastructures, persistent identifiers and authority data are the backbone that allow collaboration across disciplines.&lt;/p&gt;
&lt;/li&gt;
&lt;li&gt;
&lt;p&gt;&lt;strong&gt;Infrastructure is both technical and human&lt;/strong&gt;: Behind cables, servers, and workflows, it’s people, governance, and communities that keep infrastructures alive.&lt;/p&gt;
&lt;/li&gt;
&lt;li&gt;
&lt;p&gt;&lt;strong&gt;Bridging scales is key&lt;/strong&gt;: The inspiring keynote reminded me how crucial large-scale infrastructures and global governance are. But in the sessions and posters, I also saw how much progress depends on practical tools, workflows, and community initiatives that researchers can actually use today. The real challenge is linking both scales together.&lt;/p&gt;
&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Even though in the rest of this post I will only focus on my personal &lt;em&gt;main highlights&lt;/em&gt;
it is still a quite lengthy post.
Feel free to immediately jump to the last section on my &lt;a href=&#34;#reflections--takeaways&#34;&gt;reflections&lt;/a&gt;
or other sections that interest you by using the table of contents below.&lt;/p&gt;


&lt;details class=&#34;toc-inpage d-print-none  &#34; open&gt;
  &lt;summary class=&#34;font-weight-bold&#34;&gt;Table of Contents&lt;/summary&gt;
  &lt;nav id=&#34;TableOfContents&#34;&gt;
  &lt;ul&gt;
    &lt;li&gt;&lt;a href=&#34;#context&#34;&gt;Context&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#keynotes&#34;&gt;Keynotes&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#rob-finn---life-science-data&#34;&gt;Rob Finn - Life Science data&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#cathrin-stöver---research-and-education-infrastructure&#34;&gt;Cathrin Stöver - Research and Education Infrastructure&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#persistent-identifiers-authority-data--knowledge-graphs&#34;&gt;Persistent identifiers, Authority data &amp;amp; Knowledge Graphs&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#pids-small-links-big-impact&#34;&gt;PIDs: Small links, big impact&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#authority-data-in-the-humanities&#34;&gt;Authority data in the humanities&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#knowledge-graphs-and-ontologies&#34;&gt;Knowledge Graphs and Ontologies&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#ai&#34;&gt;AI&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#infrastructure&#34;&gt;Infrastructure&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#governance&#34;&gt;Governance&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#global-governance&#34;&gt;Global governance&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#governance-in-practice&#34;&gt;Governance in practice&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#training--communities&#34;&gt;Training &amp;amp; Communities&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#reflections--takeaways&#34;&gt;Reflections &amp;amp; Takeaways&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#references&#34;&gt;References&lt;/a&gt;&lt;/li&gt;
  &lt;/ul&gt;
&lt;/nav&gt;
&lt;/details&gt;

&lt;h2 id=&#34;keynotes&#34;&gt;Keynotes&lt;/h2&gt;
&lt;p&gt;This year&amp;rsquo;s conference featured two keynotes:
&lt;strong&gt;Rob Finn&lt;/strong&gt; talked about the complex data ecosystem of life science data
and
&lt;strong&gt;Cathrin Stöver&lt;/strong&gt; provided insights into the physical networks
that make all of our work possible in the first place.&lt;/p&gt;
&lt;h3 id=&#34;rob-finn---life-science-data&#34;&gt;Rob Finn - Life Science data&lt;/h3&gt;
&lt;p&gt;On the first conference day,
&lt;a href=&#34;https://orcid.org/0000-0001-8626-2148&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Rob Finn&lt;/a&gt;
from the &lt;a href=&#34;https://ror.org/02catss52&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;European Bioinformatics Institute&lt;/a&gt;
talked about the kind of data that they process and publish.
As well as where he sees potential for AI.&lt;/p&gt;
&lt;p&gt;According to Rob,
and hinting towards work of &lt;a href=&#34;https://orcid.org/0000-0003-1219-2137&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Carole Goble&lt;/a&gt;,
there are ad-hoc workflows which can be made FAIR
by describing them with &lt;a href=&#34;https://www.wikidata.org/wiki/Q124366860&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Ro-Crate&lt;/a&gt;
and uploading those descriptions to the &lt;a href=&#34;https://workflowhub.eu/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;WorkflowHub&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;It&amp;rsquo;s not really my domain,
but what I take home from Rob&amp;rsquo;s talk
is the evolution of workflows.
Because I recognized my work for the &lt;a href=&#34;https://www.kbr.be/en/projects/beltrans/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELTRANS project&lt;/a&gt;,
where we create a corpus of book metadata of contemporary
translations between Dutch and French.
It&amp;rsquo;s a completely different area,
but in our case we also start with a bunch of
Python scripts, glued together with bash scripts.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;From Python scripts to the WorkflowHub&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_keynote-workflow_hu0ac40359d5d24d309b294d48798cf5b9_435819_5ec48c804b0fdce5332b6d02b7b04b65.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_keynote-workflow_hu0ac40359d5d24d309b294d48798cf5b9_435819_201b0f91242faf672c4057bc82934333.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_keynote-workflow_hu0ac40359d5d24d309b294d48798cf5b9_435819_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_keynote-workflow_hu0ac40359d5d24d309b294d48798cf5b9_435819_5ec48c804b0fdce5332b6d02b7b04b65.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;cathrin-stöver---research-and-education-infrastructure&#34;&gt;Cathrin Stöver - Research and Education Infrastructure&lt;/h3&gt;
&lt;p&gt;In her impressive keynote on the last conference day,
&lt;a href=&#34;https://www.wikidata.org/wiki/Q136008224&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Cathrin Stöver&lt;/a&gt;
the Chief Communication Officer at &lt;a href=&#34;https://ror.org/052hmv319&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;GÉANT&lt;/a&gt;,
discussed how GÉANT provides global connectivity for researchers.&lt;/p&gt;
&lt;p&gt;She wasn&amp;rsquo;t referring to a few online services or cloud solutions.
She discussed real-world 800 GBit traffic
via GÉANT&amp;rsquo;s own submarine cables,
as well as other success stories,
such as the well-known eduroam Wi-Fi internet access service.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Geant - 25 years of building global R&amp;amp;amp;D networks&#34; srcset=&#34;
               /media/cordi-2025/2025-08-28_cordi_geant-timeline_hu1c2e29d68346f9c2568ba08ead744de2_1304523_7f6a377eb3c02ad64e1cb9ba6d78ddf0.webp 400w,
               /media/cordi-2025/2025-08-28_cordi_geant-timeline_hu1c2e29d68346f9c2568ba08ead744de2_1304523_b81aa8a71c790b56044f079caa6194fc.webp 760w,
               /media/cordi-2025/2025-08-28_cordi_geant-timeline_hu1c2e29d68346f9c2568ba08ead744de2_1304523_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-28_cordi_geant-timeline_hu1c2e29d68346f9c2568ba08ead744de2_1304523_7f6a377eb3c02ad64e1cb9ba6d78ddf0.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;GÉANT provides research and education network services
by interconnecting 38 National Research and Education Networks (NRENs)
such as &lt;a href=&#34;https://ror.org/038t94116&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;&lt;em&gt;Deutsches Forschungsnetz&lt;/em&gt; (DFN)&lt;/a&gt; in Germany
or &lt;a href=&#34;https://ror.org/02d3h3h49&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELNET&lt;/a&gt; in Belgium
and serves around 50 million users across Europe.
In the last 15 years,
other parts of the world have been connected as well,
such as East Africa extensively and West Africa to a lesser extent.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;EU and national investments into Research &amp;amp;amp; Infrastructure&#34; srcset=&#34;
               /media/cordi-2025/2025-08-28_cordi_geant-rd-funding_hu3da12e04e44f7d782bf69d3c50272c2b_1720614_752d6c551f0ed26fe57cca03fa58310f.webp 400w,
               /media/cordi-2025/2025-08-28_cordi_geant-rd-funding_hu3da12e04e44f7d782bf69d3c50272c2b_1720614_424be7bda5b5d6610acc72ea4e7d5bb1.webp 760w,
               /media/cordi-2025/2025-08-28_cordi_geant-rd-funding_hu3da12e04e44f7d782bf69d3c50272c2b_1720614_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-28_cordi_geant-rd-funding_hu3da12e04e44f7d782bf69d3c50272c2b_1720614_752d6c551f0ed26fe57cca03fa58310f.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In addition to these interesting facts,
the keynote also focused on the conference organizing organization NFDI.
It must undergo evaluations and needs to become operational within a few years.
Based on GÉANT’s experiences,
Cathrin identified three important aspects for success:
&lt;strong&gt;validation, people, and funding&lt;/strong&gt;.
GÉANT receives top-down and bottom-up validation from both its users
and recognition on policy level, including an alignment with the
Sustainable Development Goals.
Evenly important is the people aspect with
a strong community and regular community conferences.
But also having a value conversation to keep people with passion.
And last but not least,
the aspect of secure and sufficient funding,
which due to the multiplier effect pays off.&lt;/p&gt;
&lt;p&gt;If you ask me, I think NFDI is on a good way!
There is a large user base and recognition from the policy level,
both from the German government
as well as via links to European research.
I have also met many passionate people at the conference.
And at least from an outsider perspective,
I think lots of money is invested in Germany for RDM at the moment.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;persistent-identifiers-authority-data--knowledge-graphs&#34;&gt;Persistent identifiers, Authority data &amp;amp; Knowledge Graphs&lt;/h2&gt;
&lt;p&gt;If there’s one area at CoRDI that resonated most with my own work,
it was the trio of persistent identifiers, authority data, and knowledge graphs.
They are deeply interconnected: identifiers give us stable references,
authority files give meaning to those references,
and knowledge graphs weave them together into broader ecosystems.&lt;/p&gt;
&lt;h3 id=&#34;pids-small-links-big-impact&#34;&gt;PIDs: Small links, big impact&lt;/h3&gt;
&lt;p&gt;Every Knowledge Graph or authority file starts with the simplest building block:
a persistent identifier.
Several posters and discussions at CoRDI showed just how central PIDs have become.&lt;/p&gt;
&lt;p&gt;Unfortunately I missed the PID panel,
because it was scheduled in parallel to the humanities and authority files session
But thanks to the poster sessions,
I could speak with experts directly.
For example at the poster of
&lt;a href=&#34;https://orcid.org/0009-0009-5034-6687&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Barbara Fischer&lt;/a&gt;
from the &lt;a href=&#34;https://ror.org/01n7gem85&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;German National Library (DNB)&lt;/a&gt;
and
&lt;a href=&#34;https://orcid.org/0000-0003-4448-3844&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Paul Vierkant&lt;/a&gt;
from &lt;a href=&#34;https://ror.org/04wxnsj81&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DataCite&lt;/a&gt; about the PID network Germany.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The poster of the PID network Germany&#34; srcset=&#34;
               /media/cordi-2025/2025-08-27_cordi_pid-network-de_hu805de766e05d652129a833d39fbce631_5813996_7f9737b8b42b27201882ddfbccd920bc.webp 400w,
               /media/cordi-2025/2025-08-27_cordi_pid-network-de_hu805de766e05d652129a833d39fbce631_5813996_fb16a7fb0ef915da4f94916d7bcf978e.webp 760w,
               /media/cordi-2025/2025-08-27_cordi_pid-network-de_hu805de766e05d652129a833d39fbce631_5813996_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-27_cordi_pid-network-de_hu805de766e05d652129a833d39fbce631_5813996_7f9737b8b42b27201882ddfbccd920bc.webp&#34;
               width=&#34;573&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;We don&amp;rsquo;t have something similar on national level in Belgium yet,
(probably also due to the special federalism in Belgium),
but I&amp;rsquo;m glad that my colleagues from the &lt;a href=&#34;https://www.belspo.be/belspo/coordination/resDev_openScience_en.stm&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FedOSC project&lt;/a&gt;
are also looking into the PID topic, at least for the Federal Scientific Institutions!
While talking about PIDs, we were also joined by &lt;a href=&#34;https://orcid.org/0000-0001-9488-1870&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Daniel Mietchen&lt;/a&gt;,
who argued that Wikibase already comes with PIDs out of the box.&lt;/p&gt;
&lt;p&gt;But identifiers on their own are just strings (or ideally Uniform Resource Identifiers (URIs)).
Their real value comes when we use them to describe entities
such as people, places, or works,
which is where authority data comes in.&lt;/p&gt;
&lt;hr&gt;
&lt;h3 id=&#34;authority-data-in-the-humanities&#34;&gt;Authority data in the humanities&lt;/h3&gt;
&lt;p&gt;Authority files give identity to PIDs,
turning abstract identifiers into meaningful references for people, places, and works.
The humanities session at CoRDI highlighted how this plays out in practice.&lt;/p&gt;
&lt;p&gt;The humanities session featured talks
about the &lt;a href=&#34;https://www.wikidata.org/wiki/Q36578&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;&lt;em&gt;Gemeinsame Normdatei (GND)&lt;/em&gt;&lt;/a&gt;, which is the authority file of German-speaking countries,
and the Knowledge Graph software Wikibase.&lt;/p&gt;
&lt;p&gt;In the presentation, &lt;em&gt;Authority Files and the Text+ Data Space&lt;/em&gt;
(&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736146&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736146&lt;/a&gt;),
&lt;a href=&#34;https://orcid.org/0009-0009-5034-6687&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Barbara Fischer&lt;/a&gt;
highlighted that&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;standards are not dropping from the sky&amp;rdquo;&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;Furthermore,
she argued that we should not lose diversity when using standards.
Barbara used English as an example of a language
used for understanding each other in science.
However, it does not replace one’s mother tongue,
with which one can express much more diverse concepts.
Barbara further argued that we have to give people the edit button,
or at least a button to suggest changes,
this will also enhance collaboration.&lt;/p&gt;
&lt;p&gt;Open Knowledge Graphs are one way to enable (large-scale) collaboration.
While most people are familiar with Wikidata,
there are also lesser-known, more domain-specific Knowledge Bases.
One such example is the &lt;a href=&#34;https://www.wikidata.org/wiki/Q90405608&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FactGrid&lt;/a&gt; Wikibase instance,
used by many historians.
Because the historians share a single Wikibase instance,
they are also forced to work together.&lt;/p&gt;
&lt;p&gt;FactGrid&amp;rsquo;s creator,
&lt;a href=&#34;https://orcid.org/0000-0001-9230-4666&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Olaf Simons&lt;/a&gt;
gave the presentation &lt;em&gt;Wikibase - the best software to communicate with the upcoming knowledge graphs?&lt;/em&gt;
(&lt;a href=&#34;https://doi.org/10.5281/zenodo.16892677&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16892677&lt;/a&gt;).
In it, he states a very concrete issue:
the 100 million item limit of Wikidata&amp;rsquo;s (and Wikibase&amp;rsquo;s) Blazegraph triple store!
Apparently the Wikidata graph was split as a workaround.
Yet, Olaf argues that these kind of workarounds do not solve the underlying issues.
He calls for the simple but effective solution:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;The NFDI should develop a joint interest with Wikimedia Germany to break the present 100 million item limit&amp;rdquo;&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Olaf Simons presents the 100 million item limit of Wikibase&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_100-mio-triples-challenge_huc5c78d0b689f2694310b8fd9dbbc7afb_5548299_570f6323885c58d18a35239ce606927c.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_100-mio-triples-challenge_huc5c78d0b689f2694310b8fd9dbbc7afb_5548299_b5d8db20f64eeeaa50eb87daa804cb01.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_100-mio-triples-challenge_huc5c78d0b689f2694310b8fd9dbbc7afb_5548299_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_100-mio-triples-challenge_huc5c78d0b689f2694310b8fd9dbbc7afb_5548299_570f6323885c58d18a35239ce606927c.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;I can very much relate to Olaf&amp;rsquo;s statement.
As a computer scientist I am used to developing custom software solutions
when I cannot find standard software that addresses the problem.
I often see software developers in different (engineering) fields
develop software even when standard software is already available.
One example is software to generate RDF from other data sources.
Standard software is important!
In the humanities, for example,
where tools with a user interface
and a high &lt;a href=&#34;https://www.wikidata.org/wiki/Q1478071&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Technology Readiness Level (TRL)&lt;/a&gt; are needed,
Wikibase plays an important role.
This fundamental software should be maintained
and further developed to overcome its current limitations.&lt;/p&gt;
&lt;p&gt;Authority files create structure, but once we start connecting them across domains and datasets,
we step into the world of Knowledge Graphs and ontologies.&lt;/p&gt;
&lt;hr&gt;
&lt;h3 id=&#34;knowledge-graphs-and-ontologies&#34;&gt;Knowledge Graphs and Ontologies&lt;/h3&gt;
&lt;p&gt;Knowledge Graphs are where identifiers and authority files come together,
enabling connections across datasets and disciplines.
At CoRDI, I saw many examples of both the opportunities
and the challenges of building such ecosystems.&lt;/p&gt;
&lt;p&gt;In the humanities session,
&lt;a href=&#34;https://orcid.org/0000-0002-3246-3531&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Florian Thiery&lt;/a&gt;
and
&lt;a href=&#34;https://orcid.org/0000-0002-5190-1867&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Lozana Rossenova&lt;/a&gt;
presented
challenges of federated queries using the Wikiverse and OpenStreetMap
within the NFDI Knowledge Graph Ecosystem
(&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736048&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736048&lt;/a&gt;).
While doing federated SPARQL qeries across Wikibase instances,
including FactGrid,
they noticed technical and conceptual hurdles,
such as namespace collisions, property mappings or
allow lists.
The latter specifying which other Wikibase instances can be accessed by federated queries.&lt;/p&gt;
&lt;p&gt;On the last day of the main CoRDI conference,
&lt;a href=&#34;https://orcid.org/0000-0002-5190-1867&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Lozana Rossenova&lt;/a&gt;
also presented the results of a survey about
how NFDI consortia are using Knowledge Graphs (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736078&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736078&lt;/a&gt;).
The survey was conducted between March and May 2025
and covered 27 Knowledge Graphs from 11 consortia.
















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Common challenges about using Knowledge Graphs in NFDI consortia&#34; srcset=&#34;
               /media/cordi-2025/2025-08-28_cordi_kg-survey-challenges_hu9366ffdbd8d7b28c27120ebf5ca9f236_1254047_909cbc3a21ead5e7c7301ad5acf98cce.webp 400w,
               /media/cordi-2025/2025-08-28_cordi_kg-survey-challenges_hu9366ffdbd8d7b28c27120ebf5ca9f236_1254047_61cc1e6f8558b80ecc53a65c7018c6c3.webp 760w,
               /media/cordi-2025/2025-08-28_cordi_kg-survey-challenges_hu9366ffdbd8d7b28c27120ebf5ca9f236_1254047_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-28_cordi_kg-survey-challenges_hu9366ffdbd8d7b28c27120ebf5ca9f236_1254047_909cbc3a21ead5e7c7301ad5acf98cce.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;For their presentation,
the challenges were grouped into four larger categories.
To me, it looks very familiar.
Anyone who starts working with Knowledge Graphs will eventually face one of these challenges.
We cannot overcome every challenge in a generic fashion.
I believe we should curate best practices related to these challenges
and create teaching materials from them.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://orcid.org/0000-0002-5149-603X&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Volker Hofmann&lt;/a&gt;
presented the Helmholtz Knowledge Graph (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736336&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736336&lt;/a&gt;).
I liked the used images of the presentation, such as the following about
separated islands of data. Apparently only a small percentage of all the entities.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The Helmholtz Knowledge Graph with many separated islands of data&#34; srcset=&#34;
               /media/cordi-2025/2025-08-28_cordi_helmholtz-graph_hueea98758f13c38b275b7af79feb418f4_1471399_722ab4a42129f38d4581682c25e489ee.webp 400w,
               /media/cordi-2025/2025-08-28_cordi_helmholtz-graph_hueea98758f13c38b275b7af79feb418f4_1471399_9fd5255ee11d4495e4026f84f19f8791.webp 760w,
               /media/cordi-2025/2025-08-28_cordi_helmholtz-graph_hueea98758f13c38b275b7af79feb418f4_1471399_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-28_cordi_helmholtz-graph_hueea98758f13c38b275b7af79feb418f4_1471399_722ab4a42129f38d4581682c25e489ee.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;This made me think about the many authority records in &lt;a href=&#34;https://www.wikidata.org/wiki/Q119717964&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;MetaBelgica&lt;/a&gt;
for which we only have a name and no linked birth/death dates or birth/death places.
In our source data we have linked publications though,
instead of leaving them out, we maybe should add some linked PIDs, ISBNs or publication places to MetaBelgica
to not create authority islands&amp;hellip;&lt;/p&gt;
&lt;p&gt;In
&lt;a href=&#34;https://orcid.org/0009-0003-7945-6704&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sarah Rebecca Ondraszek&amp;rsquo;s&lt;/a&gt;
presentation on the NFDI4Memory Knowledge Graph (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736125&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736125&lt;/a&gt;),
I found two things interesting.
First, they do not yet have a link with PROV-O, because of the more difficult alignment with BFO (yet there could be mappings).
Second, they mentioned Sampo, a Linked Data publishing tool from Finland,
that I have recently seen at various conferences.
We are testing it for the BELTRANS project as well,
and I know that the Sampo team is collaborating
with the CLARIAH-VL+ project in Flanders/Belgium to improve the software
(&lt;a href=&#34;https://github.com/GhentCDH/sampo-ui&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sampo fork&lt;/a&gt; from the Ghent-Centre for Digital Humanities that will be rebased to the original Sampo).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Related work to the memO ontology: ArCo, Europeana and Sampo&#34; srcset=&#34;
               /media/cordi-2025/2025-08-28_cordi_memo_hudc3b14a7e2fa41186f5aa0d965da0ea1_1667061_661fa0b44d1520e572292f7c79c55973.webp 400w,
               /media/cordi-2025/2025-08-28_cordi_memo_hudc3b14a7e2fa41186f5aa0d965da0ea1_1667061_6dc57404c0f5633bdeee464ea870cb12.webp 760w,
               /media/cordi-2025/2025-08-28_cordi_memo_hudc3b14a7e2fa41186f5aa0d965da0ea1_1667061_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-28_cordi_memo_hudc3b14a7e2fa41186f5aa0d965da0ea1_1667061_661fa0b44d1520e572292f7c79c55973.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;During the Q&amp;amp;A of
&lt;a href=&#34;https://orcid.org/0000-0002-5467-871X&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Julia Thoennissen&amp;rsquo;s&lt;/a&gt;
presentation
about FAIR and scalable access to large image data in the range of petabytes (among others in neuro science) (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736220&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736220&lt;/a&gt;)
I wondered if their workflow, which contains 3D images, could be extended with the &lt;a href=&#34;https://www.wikidata.org/wiki/Q22682088&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;International Image Interoperability Framework (IIIF)&lt;/a&gt; for data publishing.
This would provide a standard software solution instead of a custom one.
The authors were not familiar with IIIF, it&amp;rsquo;s probably less well-known outside of the cultural heritage domain.
Afterwards I had a look myself, there is a 3D community working group at IIIF.
However, based on the community group&amp;rsquo;s current description it does not seem that this feature already exists.&lt;/p&gt;
&lt;hr&gt;
&lt;p&gt;At the poster of the NFDIcore 3.0 ontology,
I also had a small chat with
&lt;a href=&#34;https://orcid.org/0000-0001-7192-7143&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Jörg Waitelonis&lt;/a&gt;
about the BFO basis of the ontology (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736251&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736251&lt;/a&gt;).
We talked about the possibilities of the ontology to also model temporal aspects, such as names that change over time.
Which comes with a lot of extra triples (in case this is done for many entities).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The poster of the NFDIcore 3.0 ontology based on BFO&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_poster-nfdi-core-ontology_hu037b29d83e98ba03dc76f3504366d79a_2734764_7dc44ac8659061cb62065d017e98e804.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_poster-nfdi-core-ontology_hu037b29d83e98ba03dc76f3504366d79a_2734764_124bd00ef9a29d04564abc8d325681d7.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_poster-nfdi-core-ontology_hu037b29d83e98ba03dc76f3504366d79a_2734764_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_poster-nfdi-core-ontology_hu037b29d83e98ba03dc76f3504366d79a_2734764_7dc44ac8659061cb62065d017e98e804.webp&#34;
               width=&#34;573&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In the presentation of
&lt;a href=&#34;https://orcid.org/0000-0003-3986-0510&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Leyla Jael Castro&lt;/a&gt;
about
Bioschemas and Schemas.science at NFDI (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16735850&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16735850&lt;/a&gt;),
I have learned about Bioschemas and Schemas.science.
The first is a is a community-driven initiative to build upon Schema.org for the life science domain.
The latter is a sister website in which they broaden the Bioschemas efforts to a more general research-related extension for Schema.org.
In both cases to enhance findability, as Schema.org is mainly used by search engines.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;ai&#34;&gt;AI&lt;/h2&gt;
&lt;p&gt;It&amp;rsquo;s 2025, of course there has to be at least one session about AI!
&lt;a href=&#34;https://orcid.org/0000-0002-8970-6282&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sandra Geisler&lt;/a&gt;
introduced the AI session,
by mentioning current barriers for FAIR Research Data Management,
and for which barriers AI can help.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Barriers for FAIR Research Data Management from which two relate to AI&#34; srcset=&#34;
               /media/cordi-2025/2025-08-27_cordi_ai-session_hud89936364b09e2519f3b6bb523fdb74c_260973_28e6066b074e3642fc303b1619407318.webp 400w,
               /media/cordi-2025/2025-08-27_cordi_ai-session_hud89936364b09e2519f3b6bb523fdb74c_260973_92c3f0839a26191212573cd1c1e673c3.webp 760w,
               /media/cordi-2025/2025-08-27_cordi_ai-session_hud89936364b09e2519f3b6bb523fdb74c_260973_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-27_cordi_ai-session_hud89936364b09e2519f3b6bb523fdb74c_260973_28e6066b074e3642fc303b1619407318.webp&#34;
               width=&#34;760&#34;
               height=&#34;573&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;For me it’s also not only about how AI can help RDM,
but also how RDM can help AI.
I forgot who said it, but someone at CoRDI remarked that &lt;em&gt;Ontologies and the Semantic Web are the ground truth to make AI creative&lt;/em&gt;.
I find this a very spot-on statement,
because machine learning without annotated data is just pattern-finding.
To move beyond that, we need humans and semantics:
structured knowledge that gives AI something meaningful to work with.&lt;/p&gt;
&lt;p&gt;Other AI-related talks and posters at the conference
focused on different aspects to annotate machine learning models.
To this regard I have learned about things like
FAIR4ML, a vocabulary to describe Machine/Deep Learning models (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16735334&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16735334&lt;/a&gt;), presented by &lt;a href=&#34;https://orcid.org/0000-0003-3986-0510&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Leyla Jael Castro&lt;/a&gt;,
or MLentory, a registry for Machine/Deep Learning models making use of FAIR Digital Objects (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16735316&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16735316&lt;/a&gt;).&lt;/p&gt;
&lt;p&gt;One presentation also mentioned an AI product,
that can generate a conversational podcast out of a research paper.
So you hear artificial voices having a chat about the topic.
The presenter said that after listening to the podcast,
one interesting thing was that the AI podcast
contained examples that were not in the paper,
but which the author deemed interesting examples as well.
That sounds amazing to me,
I have the feeling that this can revolutionize scientific communication.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;infrastructure&#34;&gt;Infrastructure&lt;/h2&gt;
&lt;p&gt;One part that is utterly important,
and leans towards the impressing keynote,
is &lt;em&gt;actual&lt;/em&gt; physical infrastructure.
With all the existing clouds, data spaces and services,
I&amp;rsquo;m often wondering where the stuff is actually hosted.
Who pays for it and for how long will the services stay online?&lt;/p&gt;
&lt;p&gt;At the very basis we have the NRENs from GÉANT, thus for example BELNET in Belgium.
And via procurement one can get access to virtual machines for example via the OCRE programme
at data centers of partners.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The base4nfdi poster featuring among others jupyter4nfdi&#34; srcset=&#34;
               /media/cordi-2025/2025-08-27_cordi_poster-base4nfdi_huf59147f27528b46103a1135aff084c1b_4003357_a1a536e185730890da36b17bf32f1865.webp 400w,
               /media/cordi-2025/2025-08-27_cordi_poster-base4nfdi_huf59147f27528b46103a1135aff084c1b_4003357_85d4cdc8b88324c4a67d85b76fc757d6.webp 760w,
               /media/cordi-2025/2025-08-27_cordi_poster-base4nfdi_huf59147f27528b46103a1135aff084c1b_4003357_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-27_cordi_poster-base4nfdi_huf59147f27528b46103a1135aff084c1b_4003357_a1a536e185730890da36b17bf32f1865.webp&#34;
               width=&#34;573&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In the context of the NFDI, base4nfdi provides higher level services such as jupyter4nfdi.
At the poster of base4nfdi, &lt;a href=&#34;https://orcid.org/0000-0002-1213-5135&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Martin Reinhardt&lt;/a&gt;
explained to me how the jupyter4nfdi works.
It is essentially a JupyterHub hosted at &lt;em&gt;Forschungszentrum Jülich&lt;/em&gt; and another data center.
Research institutions can host their own Jupyter services,
creating a decentralized network
that researchers can access via the central JupyterHub and Single-Sign-On.
This way, research institutions can configure their settings
so that only their own researchers
can use the Jupyter resources
from the institution&amp;rsquo;s servers, for example.
However in theory, researchers from other institutions could also use the provided resources.
According to Martin, this is also in the interest of cloud providers,
because they usually want at least a certain percentage of server load;
otherwise the infrastructure just sits there and costs money.&lt;/p&gt;
&lt;p&gt;Eventually,
research institutions still provide the actual infrastructure,
or at least pay for virtual machines in the cloud.
This brings us back to the initial question:
Who pays for it, and for how long will the services stay online?
This obviously depends on the research institutions,
but I think it will likely be project-based funding after all.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Sustainable research data infrastructure is project funding after all? - imgflip&#34; srcset=&#34;
               /media/cordi-2025/research-data-infrastructure-meme_hufdb039fb42de6696ed5ba98bed2b047e_92135_32c53cf773dd8666c7acd144c01d95ba.webp 400w,
               /media/cordi-2025/research-data-infrastructure-meme_hufdb039fb42de6696ed5ba98bed2b047e_92135_fb8a81d32b61e345b705c841f774037c.webp 760w,
               /media/cordi-2025/research-data-infrastructure-meme_hufdb039fb42de6696ed5ba98bed2b047e_92135_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/research-data-infrastructure-meme_hufdb039fb42de6696ed5ba98bed2b047e_92135_32c53cf773dd8666c7acd144c01d95ba.webp&#34;
               width=&#34;500&#34;
               height=&#34;666&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;While not ideal,
there are real-life examples of successful infrastructures
that prove such funding can work.
For example, historians can use &lt;a href=&#34;https://database.factgrid.de/wiki/Main_Page&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FactGrid&lt;/a&gt;
to store their data even after their funding period ends.
In Flanders,
I came across the &lt;a href=&#34;https://www.odis.be/hercules/_en_overODIS.php&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;ODIS platform&lt;/a&gt;,
which also offers a home for research data within its scope.
For the latter I also have heard that researchers
already started to mention this platform (or repository)
in their mandatory data management plans.&lt;/p&gt;
&lt;p&gt;Technical infrastructures are impressive, but they rely on more than cables and servers.
Agreements, policies, and shared responsibilities - in short, governance - are what make these systems sustainable.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;governance&#34;&gt;Governance&lt;/h2&gt;
&lt;p&gt;When we talk about infrastructure,
we inevitably end up talking about governance:
the agreements, rules, and decisions that make collaboration possible.
Networks don’t just run on cables and servers,
they run on policies, contracts, and trust.
This theme surfaced repeatedly,
from the inspiring keynote to lively panel debates and practical project presentations.&lt;/p&gt;
&lt;h3 id=&#34;global-governance&#34;&gt;Global governance&lt;/h3&gt;
&lt;p&gt;They keynote of Cathrin about global connectivity
obviously is also related to global governance,
which is not an easy topic in today&amp;rsquo;s geopolitical reality.
This was the main subject of the keynote&amp;rsquo;s Q&amp;amp;A,
with questions relating to why certain parts of the world are not connected
(GÉANT&amp;rsquo;s headquarters are in Amsterdam,
so it falls under Dutch and European law and is forced to follow sanctions)
and why satellite networks are not used instead of terrestrial networks
(there is no low-orbit capacity yet owned by Europe,
and high orbit does not provide enough bandwidth).&lt;/p&gt;
&lt;p&gt;A related perspective came up in the resilience panel:
the reminder that researchers are citizens of science, not citizens of a country.
Infrastructure cannot be shielded from geopolitics,
but the research community aspires to build connections that transcend national borders.&lt;/p&gt;
&lt;p&gt;This perspective was underlined by the example of PANGAEA during the resilience panel.
It stepped in to safeguard research data that risked being lost due to shifts in U.S. policy.
Here governance became a matter of resilience:
ensuring that research data remain accessible for the global community, even when political frameworks change.&lt;/p&gt;
&lt;h3 id=&#34;governance-in-practice&#34;&gt;Governance in practice&lt;/h3&gt;
&lt;p&gt;Governance is also a topic on a much more practical level, related to data stewardship.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://orcid.org/0000-0002-1355-5043&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Robert Herrenbrück&lt;/a&gt;
presented a poster about the establishment of a central helpdesk for KonsortSWD (NFDI4Society),
(&lt;a href=&#34;https://doi.org/10.5281/zenodo.16735332&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.16735332&lt;/a&gt;)
illustrating governance through support:
coordination and service structures that help researchers navigate responsibilities without losing sight of their own work.&lt;/p&gt;
&lt;p&gt;Somewhere between the crowded poster stands,
I found myself in a memorable discussion that perfectly captured
some practical governance challenges.
The discussion was aboutt templates for data sharing agreements.
Someone mentioned that when adapting the template to their needs
and referencing the original template,
they often get feedback such as &amp;ldquo;the template is different, why did you deviate from it?&amp;rdquo;.
I think this highlights that while templates provide a much-needed baseline,
they shouldn’t become rigid dogma.
Real collaborations need room for adaptation, and governance should support flexibility rather than punish it.&lt;/p&gt;
&lt;p&gt;If governance defines the rules and structures that make infrastructures sustainable,
communities are the people who put these rules into practice,
train new researchers, and ensure that collaboration thrives.&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;training--communities&#34;&gt;Training &amp;amp; Communities&lt;/h2&gt;
&lt;p&gt;While governance sets the framework,
it’s the communities of researchers, data stewards, and trainers who bring it to life.
From workshops to collaborative projects, these people ensure that research infrastructures are not just functional,
but actually usable and sustainable.&lt;/p&gt;
&lt;p&gt;As an outsider I must say that I am impressed about all the different initiatives
represented at CoRDI.
Through discussions at the conference I have learned that there are also other
non-NFDI initiatives in Germany, some of them older than NFDI.
For example the different federal state RDM initiatives
like bwFDM (shout out to &lt;em&gt;The Länd&lt;/em&gt;) which also had posters this time.
They jointly presented themselves on the last CoRDI in 2023 (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.242&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.52825/cordi.v1i.242&lt;/a&gt;,
see also my previous conference report &lt;a href=&#34;https://doi.org/10.59350/pg3xj-4z449&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.59350/pg3xj-4z449&lt;/a&gt;).&lt;/p&gt;
&lt;p&gt;Other initiatives are the &lt;em&gt;Fachinformationsdienste (FIDs)&lt;/em&gt;,
that in its current form receive funding since 2014.
I personally know the thematic &lt;a href=&#34;https://www.wikidata.org/wiki/Q63858763&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FIDBenelux&lt;/a&gt; as it relates to Belgium and
because they are also in the follow-up committee of the MetaBelgica project.
At CoRDI,
&lt;a href=&#34;https://orcid.org/0000-0001-8274-780X&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Reinhard Altenhöner&lt;/a&gt;
and
Jana Fabrizius
presented about the Role of FIDs Between Disciplinary Research and National Infrastructure (&lt;a href=&#34;https://doi.org/10.5281/zenodo.16736302&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;10.5281/zenodo.16736302&lt;/a&gt;).
I have learned that since 2018 FIDs are self organized and currently have long-term funding via FIDPlus.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Collaboration between FIDs and the NFDI&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_FID_hu76e88e6dd0cf4991e3fa35ac652d133a_177071_e9fc2de23f147e7024b7a37b19b59aac.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_FID_hu76e88e6dd0cf4991e3fa35ac652d133a_177071_c75a5d745f42ed4167b05dd91dce6529.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_FID_hu76e88e6dd0cf4991e3fa35ac652d133a_177071_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_FID_hu76e88e6dd0cf4991e3fa35ac652d133a_177071_e9fc2de23f147e7024b7a37b19b59aac.webp&#34;
               width=&#34;760&#34;
               height=&#34;572&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Regarding RDM training,
two posters stood out to me as memorable and insightful.
One covered the general aspects of collaboration, and the other covered the CARE principles.&lt;/p&gt;
&lt;p&gt;The poster &lt;em&gt;Lost in Collaboration&lt;/em&gt; really hit the nail with the identified difficulty for effective collaboration:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;Turning Collaboration into Concrete Outcomes.&amp;rdquo;&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Lost in collaboration poster about soft infrastructures and facilitators&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_poster-collaboration_hue2a4467c2a2f845036461e821c7427c2_5176033_123a9109765199ef93a8b9c71c604d66.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_poster-collaboration_hue2a4467c2a2f845036461e821c7427c2_5176033_176874314a88b64e87bac0b358966c01.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_poster-collaboration_hue2a4467c2a2f845036461e821c7427c2_5176033_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_poster-collaboration_hue2a4467c2a2f845036461e821c7427c2_5176033_123a9109765199ef93a8b9c71c604d66.webp&#34;
               width=&#34;573&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;At the second poster I joined an interesting discussion about the CARE principles.
&lt;a href=&#34;https://orcid.org/0000-0002-8774-0321&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Josef Jeschke&lt;/a&gt;
explained that apparently there is no clear definition in CARE about &lt;em&gt;indigenous&lt;/em&gt;,
for example, it does not have to be a minority.
In general I find the ethics aspect of CARE interesting,
so I am wondering if it could be applied even in non-indigenous use cases.
More info on their poster &lt;em&gt;The Informed CARE Data Steward. Ways to CAREification&lt;/em&gt;: &lt;a href=&#34;https://doi.org/10.5281/zenodo.16736131&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.16736131&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The informed CARE steward - ways to care for indigenous data sovereignity&#34; srcset=&#34;
               /media/cordi-2025/2025-08-26_cordi_poster-care_hu753840fed078d76be0ac6a2cf53bc0df_4583125_19bfab43fca2e5db6fcc184af9d89fc4.webp 400w,
               /media/cordi-2025/2025-08-26_cordi_poster-care_hu753840fed078d76be0ac6a2cf53bc0df_4583125_51e800d0dec22a3b1989e22eeec1399e.webp 760w,
               /media/cordi-2025/2025-08-26_cordi_poster-care_hu753840fed078d76be0ac6a2cf53bc0df_4583125_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2025/2025-08-26_cordi_poster-care_hu753840fed078d76be0ac6a2cf53bc0df_4583125_19bfab43fca2e5db6fcc184af9d89fc4.webp&#34;
               width=&#34;573&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;hr&gt;
&lt;h2 id=&#34;reflections--takeaways&#34;&gt;Reflections &amp;amp; Takeaways&lt;/h2&gt;
&lt;p&gt;Looking back at CoRDI 2025,
I was struck by how often the same themes surfaced across very different tracks.
Whether we were talking about workflows in the life sciences,
authority files in the humanities,
or large-scale infrastructures,
the same questions kept coming back: &lt;strong&gt;how do we connect, how do we sustain, and how do we collaborate?&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;I&amp;rsquo;m glad that this community exists.
PIDs and standards are probably not the most exciting topics,
but they are key and enable so much more!
So far we are preaching to the choir,
but according to me,
research infrastructures must enter main stream discussions.
&lt;strong&gt;It&amp;rsquo;s not just about some tools or consortia,
it&amp;rsquo;s nothing less than a shift in how research and research training works.&lt;/strong&gt;
And due to the scaling this also should be a shift in how funding schemes work
and how we train new generations of students and researchers.&lt;/p&gt;
&lt;p&gt;If there is one thing missing at CoRDI,
it is more European and international contributions.
But I think we are on a good way.
During the open panel discussion about how to measure the success of NFDI,
&lt;a href=&#34;https://orcid.org/0000-0002-0738-7661&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Joanne Yeomans&lt;/a&gt; from the &lt;a href=&#34;https://tdcc.nl/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Dutch Thematic Digital Compentence Centre&lt;/a&gt;,
publicly brought up something that I was thinking as well:
the success of NFDI is that NFDI is there in the first place!&lt;/p&gt;
&lt;p&gt;We, from other European institutions can reference to the NFDI/CoRDI, simply because it exists.
Even if it would be only a small initiative it would already be great,
but the scale of NFDI makes it even better:
all applied research and lessons learned stem from
collaborations of all major universities and research institutions of the
most populous country of the EU.&lt;/p&gt;
&lt;p&gt;Or in the words of the keynote speaker
&lt;a href=&#34;https://www.wikidata.org/wiki/Q136008224&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Cathrin Stöver&lt;/a&gt;&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;a community needs it conference.&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;&lt;strong&gt;I hope CoRDI will become &lt;em&gt;the&lt;/em&gt; research infrastructure conference,
beyond NFDI.
Hopefully see you all at the next CoRDI!&lt;/strong&gt;&lt;/p&gt;
&lt;h2 id=&#34;references&#34;&gt;References&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;Fischer, B. K., Eckart, T., Körner, E., Buddenbohm, S., Al-Eryani, S., &amp;amp; Gradl, T. (2025). &lt;i&gt;Authority Files and the Text+ Data Space&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736146&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736146&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Simons, O., &amp;amp; Moeller, K. (2025). &lt;i&gt;Wikibase - the best software to communicate with the upcoming knowledge graphs?&lt;/i&gt; Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16892677&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16892677&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Thiery, F., Rossenova, L., Mietchen, D., Homburg, T., &amp;amp; Thiery, P. (2025). &lt;i&gt;Distributed Research Data Knowledge Graphs - Challenges of federated queries using the Wikiverse and OpenStreetMap within the NFDI Knowledge Graph Ecosystem&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736048&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736048&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Rossenova, L., Limani, F., Fortmann-Grote, C., Kaplan, A., Thiery, F., Latif, A., Shigapov, R., Zapilko, B., Schubotz, M., Fliegl, H., &amp;amp; Tietz, T. (2025). &lt;i&gt;How are NFDI consortia using Knowledge Graphs?: An overview of common functions and challenges by the Working Group &amp;ldquo;Knowledge Graphs&amp;rdquo;&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736078&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736078&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Hofmann, V., Soylu, M., Preuß, G., Fathalla, S., Kulla, L., Dmello, F., &amp;amp; Sandfeld, S. (2025). &lt;i&gt;The Helmholtz Knowledge Graph - towards a Helmholtz FAIR data space&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736336&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736336&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Ondraszek, S. R., Tietz, T., Posthumus, E., Fliegl, H., Waitelonis, J., &amp;amp; Sack, H. (2025). &lt;i&gt;Indexing Historical Research Data: MemO and the NFDI4Memory Knowledge Graph&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736125&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736125&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Thönnißen, J., Oliveira, S., Oberstrass, A., Kropp, J.-O., Gui, X., Schiffer, C., &amp;amp; Dickscheid, T. (2025). &lt;i&gt;A Perspective on FAIR and Scalable Access to Large Image Data&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736220&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736220&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Waitelonis, J., Tietz, T., Beygi Nasrabadi, H., Bruns, O., Posthumus, E., Fliegl, H., &amp;amp; Sack, H. (2025). &lt;i&gt;NFDIcore 3.0: A Mid-Level BFO2020-Compliant Ontology for Sustainable Research Data Interoperability Across Consortia&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16736251&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16736251&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Castro, L. J., Gaignard, A., Juty, N., Schnitzer, H., Reed, P., &amp;amp; Goble, C. (2025). &lt;i&gt;Bioschemas and Schemas.science at NFDI&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16735850&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16735850&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Solanki, D., Ciuciu-Kiss, J. T., Quiñones, N., Ravinder, R., Venkatesh, S., Rebholz-Schuhmann, D., Garijo, D., &amp;amp; Castro, L. J. (2025). &lt;i&gt;FAIR4ML, a vocabulary to describe Machine/Deep Learning models&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16735334&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16735334&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Quiñones, N., Geist, L., Ravinder, R., Solanki, D., Venkatesh, S., Rebholz-Schuhmann, D., &amp;amp; Castro, L. J. (2025). &lt;i&gt;MLentory: NFDI4DS registry for machine learning models and related artifacts&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16735316&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16735316&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Herrenbrück, R., &amp;amp; Bayer, S. (2025). &lt;i&gt;Establishing a Central Helpdesk for KonsortSWD - NFDI4Society: Goals, Challenges, and Solutions&lt;/i&gt;. Zenodo. &lt;a href=&#34;https://doi.org/10.5281/ZENODO.16735332&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/ZENODO.16735332&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Brand, O., Bruhn, K., Cyra, M., Fingerhuth, M., Gerlach, R., Jacob, B., … Weiner, B. (2023). The Federal State Initiatives for RDM as Intermediaries in a Dynamic Landscape of RDM Infrastructures and Services. Proceedings of the Conference on Research Data Infrastructure , 1. &lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.242&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.52825/cordi.v1i.242&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Lieber, S. (2023, September 17). CoRDI 2023. Posts | Sven Lieber. &lt;a href=&#34;https://doi.org/10.59350/pg3xj-4z449&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.59350/pg3xj-4z449&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Altenhöner, R., &amp;amp; Fabrizius, J. (2025, August 4). Bridging Communities: The Role of FIDs Between Disciplinary Research and National Infrastructure. 2nd Conference on Research Data Infrastructure (CoRDI), Aachen, Germany. &lt;a href=&#34;https://doi.org/10.5281/zenodo.16736302&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.16736302&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;</description>
    </item>
    
    <item>
      <title>Clustering Book editions</title>
      <link>https://sven-lieber.org/en/2023/10/16/clustering-book-editions/</link>
      <pubDate>Mon, 16 Oct 2023 19:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2023/10/16/clustering-book-editions/</guid>
      <description>&lt;p&gt;What do the books &amp;ldquo;The invention of Nature&amp;rdquo; and &amp;ldquo;De uitvinder van de natuur&amp;rdquo; have in common?
Well, they are both different versions of the same &lt;em&gt;work&lt;/em&gt; &amp;ldquo;The invention of nature&amp;rdquo; by Andrea Wulf,
whether it is in a different format or a different language.
In this blog post I will briefly introduce the advantages of keeping
work-level records in library catalogs.
Furthermore, I will introduce &lt;strong&gt;a fast Python implementation&lt;/strong&gt; (&lt;a href=&#34;https://zenodo.org/doi/10.5281/zenodo.10011416&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.10011416&lt;/a&gt;)
which we used in the BELTRANS project to identify the works in a corpus of book translations #FRBRization.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;Library catalogs usually contain records about different versions of a book.
In technical terms, these records are so-called manifestations,
but for simplicity I will refer to them simply as book versions or book editions.
Such a conceptual division between different levels of description
is part of the Functional Requirements for Bibliographic Records (FRBR)
or how it is referred to nowadays, the IFLA Library Reference Model (IFLA-LRM).&lt;/p&gt;
&lt;h2 id=&#34;better-user-experience-and-better-data-quality-due-to-identified-works&#34;&gt;Better user experience and better data quality due to identified works&lt;/h2&gt;
&lt;p&gt;Imagine that you are searching for a specific book in a library catalog.
You remember that it had the word &amp;ldquo;invention&amp;rdquo; in its title.
Instead of seeing two search results &lt;strong&gt;&amp;ldquo;The invention of nature&amp;rdquo; by Andrea Wulf&lt;/strong&gt;
and &lt;strong&gt;&amp;ldquo;Invention: A Life of Learning through Failure&amp;rdquo; by James Dyson&lt;/strong&gt;,
you are overwhelmed with hundreds of versions of both books and others
such as ebooks, paperbacks, reprints, etc.
Actually you are interested in finding out which works the library has that contain the title.
Which specific version maybe is of less priority to you.
In such a scenario it might be better to only show works in the search results.
After clicking on the work you looked for,
you still can be presented with all the different versions of the book.&lt;/p&gt;
&lt;p&gt;There is also an advantage for the library personnel that works with the book data.
Creating catalog records often is a manual process, or at least it was for a long time.
Some of the records might contain more information than others,
for example the genre according to a specific categorization.
Identifying all the different versions of a book in your own catalog
or in a collection of (other) catalogs can help to fill gaps in the data.
If one book record is classified as a Comic and mentions a specific person as its illustrator,
another record of a different version of that book which maybe has no such information,
can be completed/enriched with this information.&lt;/p&gt;
&lt;h2 id=&#34;alright-you-convinced-me-how-can-i-do-it&#34;&gt;Alright, you convinced me, how can I do it?&lt;/h2&gt;
&lt;p&gt;There are different ways to do this.
I have read &lt;a href=&#34;https://orkg.org/list/R608176&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;a bit of literature&lt;/a&gt; and it seems that there is one very common way to do this.
Let me introduce you to &lt;strong&gt;the work-set clustering algorithm&lt;/strong&gt;, initially devised by OCLC.
The idea is simple: create several possible descriptive keys of a book version
based on a combination of work-level information such as title and author,
then check if two book versions have at least one key in common.
Sounds complicated? Check out the image below which hopefully makes it clear.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The English and Dutch version of a book that can be matched based on a common descriptive key&#34; srcset=&#34;
               /media/clustering-book-editions/work-set-clustering-descriptive-keys_hu2914f170801a4fd4098bf76695b64d75_88986_cb855f95b36e2a39a2dc1626a0ca7cee.webp 400w,
               /media/clustering-book-editions/work-set-clustering-descriptive-keys_hu2914f170801a4fd4098bf76695b64d75_88986_b7e29052741b4312d88a720154a6e9c0.webp 760w,
               /media/clustering-book-editions/work-set-clustering-descriptive-keys_hu2914f170801a4fd4098bf76695b64d75_88986_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/clustering-book-editions/work-set-clustering-descriptive-keys_hu2914f170801a4fd4098bf76695b64d75_88986_cb855f95b36e2a39a2dc1626a0ca7cee.webp&#34;
               width=&#34;760&#34;
               height=&#34;260&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In this example you can see two book records (red),
each describing a book version.
For each book version descriptive keys (blue) are generated
to assist in finding matches.
As you can see, both records have at least one key in common.
We found a match!
In this case only because we also considered a possible original title as part of the key.
Similarly we would find a match between the shown Dutch version of the printed book
and a hypothetical ebook version: the ISBN would be different,
but there would be a &lt;strong&gt;match based on the title-author or title-translator combination&lt;/strong&gt;.&lt;/p&gt;
&lt;h2 id=&#34;nice-idea-how-do-i-write-a-program-to-do-this&#34;&gt;Nice idea, how do I write a program to do this?&lt;/h2&gt;
&lt;p&gt;That’s the one million dollar question.
Different scientific papers talk about the algorithm, a few share code.
But a code implementation always depends on the context it was written for,
for example by using XSLT rules for XML records.
I tried to write a program in Python that is as generic as possible
and came up with three solutions.
Why three? Well, it turned out the first two were not fast enough
for the amount of translations we have in the BELTRANS project.&lt;/p&gt;
&lt;h3 id=&#34;first-attempt-reusing-the-python-package-sklearn&#34;&gt;First attempt: reusing the Python package sklearn&lt;/h3&gt;
&lt;p&gt;Most of the papers I have read are already a few years old,
they mainly use software stacks that are less common nowadays.
I needed a solution that I can quickly run without a large setup.
Currently a lot of code libraries for common languages such as Python exist.
So why not reuse an existing one?!
The class &lt;code&gt;AgglomerativeClustering&lt;/code&gt; from the Python package sklearn implements a hierarchical clustering algorithm.&lt;/p&gt;
&lt;p&gt;Actually I am not interested in different levels of clusters,
only the highest level with the least number of clusters.
So I did what a lot of people do nowadays, &lt;strong&gt;I had a chat with ChatGPT to brainstorm&lt;/strong&gt;
possible solutions with AgglomerativeClustering.
And it came up with a smart idea!&lt;/p&gt;
&lt;p&gt;The input for the clustering is a so-called distance matrix:
basically a table listing all elements as rows
and all elements as columns,
where the value in a cell indicates how similar an element in row X is with the element in column Y.
We don&amp;rsquo;t need a sophisticated similarity measure,
a simple &amp;ldquo;yes&amp;rdquo; for an overlap would suffice.&lt;/p&gt;
&lt;p&gt;ChatGPTs idea was to use negative distance values.
Like this, the whole clustering just takes two lines of code
where I instruct the AgglomerativeClustering library to return clusters with distance threshold zero.&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-Python&#34; data-lang=&#34;Python&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;kn&#34;&gt;from&lt;/span&gt; &lt;span class=&#34;nn&#34;&gt;sklearn.cluster&lt;/span&gt; &lt;span class=&#34;kn&#34;&gt;import&lt;/span&gt; &lt;span class=&#34;n&#34;&gt;AgglomerativeClustering&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;n&#34;&gt;model&lt;/span&gt; &lt;span class=&#34;o&#34;&gt;=&lt;/span&gt; &lt;span class=&#34;n&#34;&gt;AgglomerativeClustering&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;(&lt;/span&gt;&lt;span class=&#34;n&#34;&gt;n_clusters&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;=&lt;/span&gt;&lt;span class=&#34;kc&#34;&gt;None&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;,&lt;/span&gt; \
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;                                &lt;span class=&#34;n&#34;&gt;affinity&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;=&lt;/span&gt;&lt;span class=&#34;s1&#34;&gt;&amp;#39;precomputed&amp;#39;&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;,&lt;/span&gt; \
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;                                &lt;span class=&#34;n&#34;&gt;linkage&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;=&lt;/span&gt;&lt;span class=&#34;s1&#34;&gt;&amp;#39;single&amp;#39;&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;,&lt;/span&gt; \
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;                                &lt;span class=&#34;n&#34;&gt;distance_threshold&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;=&lt;/span&gt;&lt;span class=&#34;mi&#34;&gt;0&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;)&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;c1&#34;&gt;# get a list where each index &lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;c1&#34;&gt;# corresponds to an element in elementIDs&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;c1&#34;&gt;# e.g. [0,2,0,1] the first element is in cluster 0, &lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;c1&#34;&gt;# the second in cluster 2, &lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;c1&#34;&gt;# the third in cluster 0 and the 4th in cluster 1&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&lt;span class=&#34;n&#34;&gt;clusterLabels&lt;/span&gt; &lt;span class=&#34;o&#34;&gt;=&lt;/span&gt; &lt;span class=&#34;n&#34;&gt;model&lt;/span&gt;&lt;span class=&#34;o&#34;&gt;.&lt;/span&gt;&lt;span class=&#34;n&#34;&gt;fit_predict&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;(&lt;/span&gt;&lt;span class=&#34;n&#34;&gt;distanceMatrix&lt;/span&gt;&lt;span class=&#34;p&#34;&gt;)&lt;/span&gt;
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;This went well for a relevant subset of roughly 20 thousand records in our BELTRANS project.
After a few minutes we got correct and directly reusable results!
However, when we wanted to achieve the same for more than 150 thousand records,
this implementation of the clustering was not performant enough.
Even on a powerful server with more than 32 GB of RAM,
the computation of the distance matrix takes too much time and space.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    One element likely has very few or no matches with all the other elements.
A sparse matrix can be used, where not &lt;em&gt;every&lt;/em&gt; value is stored, only not-zero values.
Unfortunately the &lt;code&gt;AgglomerativeClustering&lt;/code&gt; module does not support this.
  &lt;/div&gt;
&lt;/div&gt;
&lt;h3 id=&#34;second-attempt-implementing-the-algorithm-myself&#34;&gt;Second attempt: implementing the algorithm myself&lt;/h3&gt;
&lt;p&gt;A little disappointed from the last solution,
I thought I will try implementing it myself,
all I need: a few data structures and loops.
The idea of the algorithm is to start with every element in its own cluster.
Then comparing it to other clusters until we find a match.
When a match is found the loop is stopped,
the clusters are merged and we loop again over the (now updated) list of clusters.
We do all this until no more merges are possible.&lt;/p&gt;
&lt;p&gt;This algorithm does not need a large distance matrix to start with.
Testing if there is an overlap between two clusters can be reduced to a single intersection operation
between two Python sets that represent the descriptive keys of the respective clusters.&lt;/p&gt;
&lt;p&gt;Unfortunately all that takes too long.
With a few example data I could verify that it worked as expected,
but as soon as there are a few thousand elements the algorithm keeps running &amp;hellip;&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    I tried to remedy the situation by putting parts of the computation
in a reusable function that I can call in parallel by using
&lt;code&gt;ThreadPoolExecutor&lt;/code&gt; or &lt;code&gt;ProcessPoolExecutor&lt;/code&gt;, but it did not help much.
  &lt;/div&gt;
&lt;/div&gt;
&lt;h3 id=&#34;third-attempt-inverted-index-for-clustering-in-no-time&#34;&gt;Third attempt: inverted index for clustering in no time&lt;/h3&gt;
&lt;p&gt;Even more frustrated I went back to the drawing board, or pen and paper in my case.
What I need is a solution that ideally only runs once over all elements
that need to be clustered.
But how do I find out possible overlaps without comparing everything?&lt;/p&gt;
&lt;p&gt;The disadvantages of the previous solution were
that I had to perform set intersections to detect overlaps
while iterating over all clusters,
stop the iteration to update the clusters and start iterating again.&lt;/p&gt;
&lt;p&gt;Actually the data already includes the matches between elements implicitly,
and I used this to compute clusters in no time.
Every element has one or more descriptive keys.
But instead of only storing the mapping element-&amp;gt;descriptive keys in a Python dictionary,
I also store descriptive key-&amp;gt;elements in another dictionary.
&lt;strong&gt;This is my &lt;em&gt;inverted index&lt;/em&gt; which I can use to look up overlaps.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;Eventually I start with zero clusters and I simply iterate over the inverted index.
If one of the elements from a key already is part of a cluster,
I add the other elements to it.
If no cluster was found, I create a new one and add the elements.
Basically all elements of a descriptive key I check,
either go to an existing cluster or end up in a new cluster.
If different elements for a descriptive key are already in different clusters,
a new merged cluster is created and the old ones deleted.&lt;/p&gt;
&lt;p&gt;This solution works as well, and wow, it is fast too.
Interestingly &lt;strong&gt;it takes longer to upload the computed cluster information to our database
than performing the clustering itself.&lt;/strong&gt;
Inverted indexes can do a lot of heavy lifting!&lt;/p&gt;
&lt;p&gt;You can find the source code on GitHub (&lt;a href=&#34;https://zenodo.org/doi/10.5281/zenodo.10011416&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.10011416&lt;/a&gt;).&lt;/p&gt;
&lt;h3 id=&#34;why-is-it-so-fast&#34;&gt;Why is it so fast?&lt;/h3&gt;
&lt;p&gt;It all has to do with scalability.
I only iterate &lt;em&gt;once&lt;/em&gt; over all elements to build the inverted index dictionary.
Then I iterate &lt;em&gt;once&lt;/em&gt; over all elements in the inverted index to compute the clusters.
For 100 input elements this roughly means iterating 200 times,
for 100,000 input elements roughly 200,000 times.
An algorithm that &lt;strong&gt;scales linear to its input&lt;/strong&gt;.&lt;/p&gt;
&lt;p&gt;Compare this to compute a distance matrix of 100*100 elements
(compute and store 10 thousand results)
and 100,000*100,000 elements (compute and store 10 billion results).
An algorithm that &lt;strong&gt;scales quadratic to its input&lt;/strong&gt;.&lt;/p&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;Having work-level information brings advantages for both librarians and users.
The presented solution is &lt;em&gt;just one possible&lt;/em&gt; implementation of &lt;em&gt;one possible&lt;/em&gt; algorithm.
Which data fields you use to create descriptive keys
and how you extract them from your bibliographic data explicitly was not covered in this post.
I also did not cover how you could use the information of the clusters to actually achieve the advantages.
The main purpose of this blog post was to &lt;strong&gt;introduce and share the generic Python implementation&lt;/strong&gt;,
because I had the feeling there is not much out there that can be reused easily.&lt;/p&gt;
&lt;p&gt;For the BELTRANS project we extracted descriptive keys
with a SPARQL query from an Resource Description Framework (RDF) representation of bibliographic information.
Similarly, we used SPARQL INSERT queries to explicitly create RDF representations
for each work cluster and link its manifestations to it using the fabio ontology.
&lt;a href=&#34;https://github.com/kbrbe/beltrans-data-integration/tree/main/data-integration/sparql-queries/clustering&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Here&lt;/a&gt; you can have a look at our SPARQL queries.&lt;/p&gt;
&lt;p&gt;Curious about more content like this or some behind-the-scenes?
Consider subscribing to my bi-weekly Newsletter &lt;em&gt;FAIR Data Digest&lt;/em&gt;
to receive more interesting content about Linked Data every other Tuesday!&lt;/p&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;</description>
    </item>
    
    <item>
      <title>CoRDI 2023</title>
      <link>https://sven-lieber.org/en/2023/09/17/cordi-2023/</link>
      <pubDate>Sun, 17 Sep 2023 09:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2023/09/17/cordi-2023/</guid>
      <description>&lt;p&gt;It is difficult to foster innovation and perform high quality research
when the underlying data is not accessible and usable.
Data and data services should be offered via so-called &lt;em&gt;research data infrastructures&lt;/em&gt;,
think of user-friendly web platforms, repositories, and databases.
But how does it work in practice, and how important is the human aspect of it?
I have been at the &lt;strong&gt;1st Conference on Research Data Infrastructure&lt;/strong&gt;,
12 – 14 September 2023, and got some pretty good insights.
In this blog post I will reflect on the conference and cover the keynotes,
the poster sessions, our own work around Knowledge Graphs and Wikibase,
as well as a selection of other interesting talks.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;(Research) data is produced in various fields,
ranging from cultural heritage and agriculture to physics and biology for example.
Such data should be managed and preserved in a systematic way
to accelerate scientific progress and promote collaboration,
while enhancing public trust and fostering innovation.&lt;/p&gt;
&lt;p&gt;Several large research (data) infrastructure initiatives
emerged on European Level in the past decade.
Those initiatives, including the &lt;em&gt;European Strategy Forum on Research Infrastructures (ESFRI)&lt;/em&gt;
or the &lt;em&gt;European Open Science Cloud (EOSC)&lt;/em&gt;,
have brought hundreds of institutional partners together
to build and maintain research infrastructure within numerous research projects.
National initiatives have emerged as well.&lt;/p&gt;
&lt;p&gt;Under the theme &lt;em&gt;&lt;strong&gt;Connecting Communities&lt;/strong&gt;&lt;/em&gt;,
the association related to the &lt;em&gt;German National Research Data Infrastructure (NFDI)&lt;/em&gt;
organized the 1st Conference on Research Data Infrastructure (CoRDI).
More than &lt;strong&gt;680 persons&lt;/strong&gt; attended this international multi-track conference
which hosted &lt;strong&gt;84 presentations&lt;/strong&gt; and more than a &lt;strong&gt;100 posters&lt;/strong&gt;.
Next to inspiring keynotes, there were thematic tracks at the first day
and tracks related to enabling, connecting, linking,
harmonizing,  securing, and spreading of Research Data Management (RDM)
on the remaining days.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;My main takeaways&lt;/strong&gt;&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;It’s mostly about human problems that need to be fixed (connecting communities, data governance, etc)&lt;/li&gt;
&lt;li&gt;Knowledge Graphs and especially Wikibase gain attention in the context of research (data) infrastructures&lt;/li&gt;
&lt;li&gt;There is an increasing need for sustainably funded data management careers&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;In the rest of this post I will focus on my &lt;em&gt;main highlights&lt;/em&gt;,
not on all the presentations that I have seen.
Also, since the conference was organized in the frame of NFDI,
a lot of networking and promotion was devoted to the different NFDI consortia.
I am not affiliated with any of these consortia
and because I haven’t worked with any of those consortia in the past,
this blog post is limited to my own personal (outsider) experiences at the conference.
I am glad to update the post and link relevant other blog posts that provide other insides.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;


&lt;details class=&#34;toc-inpage d-print-none  &#34; open&gt;
  &lt;summary class=&#34;font-weight-bold&#34;&gt;Table of Contents&lt;/summary&gt;
  &lt;nav id=&#34;TableOfContents&#34;&gt;
  &lt;ul&gt;
    &lt;li&gt;&lt;a href=&#34;#context&#34;&gt;Context&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#keynotes&#34;&gt;Keynotes&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#christine-borgman--knowledge-infrastructures&#34;&gt;Christine Borgman – Knowledge Infrastructures&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#julia-janssen--data-art-installations&#34;&gt;Julia Janssen – data art installations&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#mark-d-wilkinson--fair-principles&#34;&gt;Mark D. Wilkinson – FAIR principles&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#poster-sessions&#34;&gt;Poster session(s)&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#metabelgica-wikibase-and-nfdi4culture&#34;&gt;MetaBelgica, Wikibase and NFDI4Culture&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#knowledge-graphs-at-nfdi&#34;&gt;Knowledge Graphs at NFDI&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#selection-of-presentations&#34;&gt;Selection of presentations&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#social-sciences&#34;&gt;Social sciences&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#data-spaces&#34;&gt;Data spaces&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#physics&#34;&gt;Physics&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#scholarly-communication&#34;&gt;Scholarly Communication&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#research-data-management&#34;&gt;Research Data Management&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#networking&#34;&gt;Networking&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#final-remarks&#34;&gt;Final remarks&lt;/a&gt;&lt;/li&gt;
  &lt;/ul&gt;
&lt;/nav&gt;
&lt;/details&gt;

&lt;h2 id=&#34;keynotes&#34;&gt;Keynotes&lt;/h2&gt;
&lt;p&gt;Three excellent keynotes were given which gave different insights into data:
&lt;strong&gt;Christine Borgman&lt;/strong&gt; zoomed into the bigger picture of
Knowledge Infrastructures,
&lt;strong&gt;Julia Janssen&lt;/strong&gt; presented how data and scientific insights
can be translated into art installations,
and &lt;strong&gt;Mark D. Wilkinson&lt;/strong&gt; talked about his regrets
when co-designing the first FAIR principles and the road ahead.&lt;/p&gt;
&lt;h3 id=&#34;christine-borgman--knowledge-infrastructures&#34;&gt;Christine Borgman – Knowledge Infrastructures&lt;/h3&gt;
&lt;blockquote&gt;
&lt;p&gt;“May all your problems be technical” – Jim Gray&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;&lt;a href=&#34;https://orcid.org/0000-0002-9344-1029&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Christine Borgman&lt;/a&gt;
ended her keynote (&lt;a href=&#34;https://doi.org/10.5281/zenodo.8344854&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.8344854&lt;/a&gt;) with this quote from Jim Gray.
And after listening to her presentation one really grasps how much truth lies in this statement.
From a research data infrastructure point of view
we have already a lot of platforms and initiatives.
But what about data governance?
To which extent do different communities agree on data management practices?
According to her, we should focus on Knowledge Infrastructures,
&lt;strong&gt;including also the social component!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Needed data governance according to Christine Borgman&#34; srcset=&#34;
               /media/cordi-2023/2023-09-12_cordi-keynote-governance_hubbea43a6af2f587e21d58d0e5d400e66_4091735_0d51e1baf6c429357a3e774623e6befd.webp 400w,
               /media/cordi-2023/2023-09-12_cordi-keynote-governance_hubbea43a6af2f587e21d58d0e5d400e66_4091735_90b2f3f0f0725a8bd89c513902bc03ed.webp 760w,
               /media/cordi-2023/2023-09-12_cordi-keynote-governance_hubbea43a6af2f587e21d58d0e5d400e66_4091735_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-12_cordi-keynote-governance_hubbea43a6af2f587e21d58d0e5d400e66_4091735_0d51e1baf6c429357a3e774623e6befd.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;One really interesting point she made
is the difference between knowledge production as done by researchers
and data production for others as done by data creators.
Christine Borgman pledges for a data management workforce:
&lt;strong&gt;domain data scientists, data librarians and data archivists&lt;/strong&gt;
who cover the middle ground between knowledge and data production.
I fully agree! More such careers are needed (and sustainably funded) at different research institutions.&lt;/p&gt;
&lt;h3 id=&#34;julia-janssen--data-art-installations&#34;&gt;Julia Janssen – data art installations&lt;/h3&gt;
&lt;p&gt;Data, .. this anyway is something really abstract.
The researcher and artist &lt;a href=&#34;https://studiojuliajanssen.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Julia Janssen&lt;/a&gt;
has shown us in her keynote how data can be made tangible.
She introduced us to some of her art installations and why and how she came up with it.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A physical data visualization, data to touch&#34; srcset=&#34;
               /media/cordi-2023/2023-09-12_cordi-keynote-data-visualization_hu7eac12269bb55eabf03bcc4e7f650671_2920215_74fee0d60b34b22268fda828a49c6b8d.webp 400w,
               /media/cordi-2023/2023-09-12_cordi-keynote-data-visualization_hu7eac12269bb55eabf03bcc4e7f650671_2920215_fd56bbde95b2dd6bc2b3119fe3bcaddd.webp 760w,
               /media/cordi-2023/2023-09-12_cordi-keynote-data-visualization_hu7eac12269bb55eabf03bcc4e7f650671_2920215_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-12_cordi-keynote-data-visualization_hu7eac12269bb55eabf03bcc4e7f650671_2920215_74fee0d60b34b22268fda828a49c6b8d.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;For example, according to her data from Meta (the company behind Facebook, Instagram and WhatApp),
she likes the color green.
She built a whole art installation with more than 4000 small balls
where each ball stands for one of her interests, according to Meta.
Just to try figuring out why she might like the color green.
Other people literally crawled under it to get to know her better,
looking at her interests trying to find out why she might like the color green.&lt;/p&gt;
&lt;h3 id=&#34;mark-d-wilkinson--fair-principles&#34;&gt;Mark D. Wilkinson – FAIR principles&lt;/h3&gt;
&lt;p&gt;A common theme at the conference were of course the FAIR principles.
The third and last keynote about those principles
was given by the main author of the original FAIR principles paper,
&lt;a href=&#34;https://orcid.org/0000-0001-6960-357X&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Mark D. Wilkinson&lt;/a&gt; (&lt;a href=&#34;https://doi.org/10.5281/zenodo.8353153&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.8353153&lt;/a&gt;).
Besides presenting his most recent work, he also talked about his regrets
after almost 10 years since the initial version of the principles from a 2014 workshop.
One of his &lt;strong&gt;regrets&lt;/strong&gt; is the term &lt;strong&gt;data&lt;/strong&gt;, apparently the term FAIR Digital Objects or FAIR Research Objects seems to capture their ideas better with less confusion. Also &lt;strong&gt;the term license&lt;/strong&gt; turned out to be an issue, &lt;em&gt;access&lt;/em&gt; or &lt;em&gt;usage policy&lt;/em&gt; would have been better, avoiding the heavy legislative meaning.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Mark D. Wilkinson presenting his keynote at CoRDI&#34; srcset=&#34;
               /media/cordi-2023/2023-09-14_cordi-fair_hu4b9726025aed6ba702068e5c3e08e25d_2946085_75a289999222f05fed8ebd28147c2bfe.webp 400w,
               /media/cordi-2023/2023-09-14_cordi-fair_hu4b9726025aed6ba702068e5c3e08e25d_2946085_41d63d6fc05907863fddb2f132a879b2.webp 760w,
               /media/cordi-2023/2023-09-14_cordi-fair_hu4b9726025aed6ba702068e5c3e08e25d_2946085_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-14_cordi-fair_hu4b9726025aed6ba702068e5c3e08e25d_2946085_75a289999222f05fed8ebd28147c2bfe.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;But looking forward, Mark D. Wilkinson talked about
(not so) autonomous agents and FAIRification in different use cases.
It was nice to see that they are also using the RDF Mapping language (RML)
and the YARRRML serialization format to accomplish FAIRification.
Those specifications and implementations for Knowledge Graph Construction were developed by ex-colleagues of mine and I am also still using it a lot.&lt;/p&gt;
&lt;h2 id=&#34;poster-sessions&#34;&gt;Poster session(s)&lt;/h2&gt;
&lt;p&gt;One of the most engaging events during a conference is of course the poster session.
&lt;strong&gt;And CoRDI had two of it!&lt;/strong&gt;
Many, many posters and as much motivated researchers to present it.
Topics covered among others anything around data management plans,
terminologies, persistent identifiers, FAIR initiatives
and of course posters presenting the different NFDI consortia.
From the plenty of posters I will focus on two in this blog post:
one more technical poster and one highlighting international collaboration.&lt;/p&gt;
&lt;p&gt;There have been posters from various domains,
and it became apparent that there is an overlap between the domains
in terms of problems and solutions.
For example, there was the poster from &lt;a href=&#34;https://orcid.org/0000-0002-7899-7192&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Steffen Neumann&lt;/a&gt;
and colleagues about repository federations in the chemical domain (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.202&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.202&lt;/a&gt;).
Among others, they work with the
&lt;em&gt;Open Archives Initiatives Protocol for Metadata Harvesting (OAI-PMH)&lt;/em&gt;,
which is also a common interface in the library domain where I am currently in.
There I have learned that it is possible to embed JSON-LD in OAI-PMH and retrieve it via Xpath.
A very handy way to publish Linked Data using already available interfaces.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A poster about repository federations of the chemical domain&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-repository-federation_huc8633d13b720bac3bf9fda5dab5749f3_4141544_7c39d0866e15888354e58652d971570b.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-repository-federation_huc8633d13b720bac3bf9fda5dab5749f3_4141544_5f80dac51dc229f2904cab12c736c898.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-repository-federation_huc8633d13b720bac3bf9fda5dab5749f3_4141544_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-repository-federation_huc8633d13b720bac3bf9fda5dab5749f3_4141544_7c39d0866e15888354e58652d971570b.webp&#34;
               width=&#34;570&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Considering the Humanities domain, there was an interesting poster from &lt;a href=&#34;https://orcid.org/0000-0002-0719-9003&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Nanette Rißler-Pipka&lt;/a&gt; and colleagues, including my colleague &lt;a href=&#34;https://orcid.org/0000-0002-2430-475X&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sally Chambers&lt;/a&gt;. (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.274&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.274&lt;/a&gt;).
The poster focused on synergies between national and European research (data) infrastructures
based on the example of the SSHOC market place.
In the &lt;a href=&#34;https://fair-data-digest.org/archive/8&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;8th edition&lt;/a&gt;
of the &lt;a href=&#34;https://fair-data-digest.org&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FAIR Data Digest&lt;/a&gt;,
I have covered with an example how differently national ERIC nodes
might be linked with national initiatives (what is an ERIC? &lt;a href=&#34;https://fair-data-digest.org/archive/6&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;check here&lt;/a&gt;).
Nice to see it and talk about it in real life.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A poster about the pathway between national and European research data infrastructure based on the SSHOC marketplace&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-sshoc_hu47134499e6b9cb61f6bf5168237522ee_3428218_7a6c5fa2e8b61cb3ed3902810dc5977e.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-sshoc_hu47134499e6b9cb61f6bf5168237522ee_3428218_e7159d0926ff12f5003c274874673c63.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-sshoc_hu47134499e6b9cb61f6bf5168237522ee_3428218_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-sshoc_hu47134499e6b9cb61f6bf5168237522ee_3428218_7a6c5fa2e8b61cb3ed3902810dc5977e.webp&#34;
               width=&#34;570&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h2 id=&#34;metabelgica-wikibase-and-nfdi4culture&#34;&gt;MetaBelgica, Wikibase and NFDI4Culture&lt;/h2&gt;
&lt;p&gt;I participated at CoRDI to present our envisioned platform &lt;a href=&#34;https://www.kbr.be/en/projects/metabelgica/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;MetaBelgica&lt;/a&gt;
which aims to become
the &lt;strong&gt;single source of truth for cultural heritage metadata about Belgian entities&lt;/strong&gt;
of type persons, organizations, time/events and locations.&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Abstract: &lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.381&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.52825/cordi.v1i.381&lt;/a&gt;&lt;/li&gt;
&lt;li&gt;Presentation: &lt;a href=&#34;https://doi.org/10.5281/zenodo.8337301&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://doi.org/10.5281/zenodo.8337301&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;For those of you subscribed to my &lt;a href=&#34;https://fair-data-digest.org&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FAIR Data Digest newsletter&lt;/a&gt;,
you already read about the main idea behind MetaBelgica
in the &lt;a href=&#34;https://fair-data-digest.org/archive/15&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;last edition&lt;/a&gt;.
The following image summarizes the problem we aim to solve.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The problem that MetaBelgica aims to solve: providing a single source of truth for Belgian entities related to cultural heritage, to remove duplicate efforts on the side of the data curators and the users&#34; srcset=&#34;
               /media/cordi-2023/2023-09-12_cordi-metabelgica_huf1ce54f87b4c10c540741a3ec2cd776d_152690_8462c1cfcef11df7b431d233f5339744.webp 400w,
               /media/cordi-2023/2023-09-12_cordi-metabelgica_huf1ce54f87b4c10c540741a3ec2cd776d_152690_9fd9eb5eb0c24f7a0dcf3c6effd07be6.webp 760w,
               /media/cordi-2023/2023-09-12_cordi-metabelgica_huf1ce54f87b4c10c540741a3ec2cd776d_152690_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-12_cordi-metabelgica_huf1ce54f87b4c10c540741a3ec2cd776d_152690_8462c1cfcef11df7b431d233f5339744.webp&#34;
               width=&#34;760&#34;
               height=&#34;417&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Each participating institution (left in the image) manages metadata about their own collections
using domain-specific data standards.
This is totally fine, because usually there aren’t any conflicts,
because there is no overlap in the collections.
But as part of the cataloging process,
the same related data about entities such as persons or events are created by each institution separately. Over and over again.
This results in &lt;strong&gt;duplicate efforts for the institutions and for the users&lt;/strong&gt;.&lt;/p&gt;
&lt;p&gt;We want to integrate the data of the partnering institutions
by using the Resource Description Framework (RDF)
and use the Wikibase software as a platform to manage the integrated data in a Knowledge Graph.&lt;/p&gt;
&lt;h3 id=&#34;knowledge-graphs-at-nfdi&#34;&gt;Knowledge Graphs at NFDI&lt;/h3&gt;
&lt;p&gt;With our submission related to Knowledge Graphs we are not alone!&lt;/p&gt;
&lt;p&gt;Under high pressure of continuously ringing alarm bells during the national-wide sirens and (phone) alarm system test, &lt;a href=&#34;https://orcid.org/0000-0002-5190-1867&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Lozana Rossenova&lt;/a&gt;
presented work about Knowledge Graph Infrastructure based on Wikibase
(&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.266&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.266&lt;/a&gt;).
Wikibase instances are already deployed by the German National Library of Science and Technology (TIB) in the context of NFDI4Culture.&lt;/p&gt;
&lt;p&gt;Speaking of which, &lt;a href=&#34;https://orcid.org/0000-0001-7069-9804&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Harald Sack&lt;/a&gt;
also presented work that originated from NFDI4Culture
(&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.371&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.371&lt;/a&gt;).
Their original NFDI4Culture ontology was extended to cover other NFDI consortia.
In that adaptation process, the ontology was modularized into a core and several extensions.&lt;/p&gt;
&lt;p&gt;This is great! Plenty of lessons learned and possible networking and collaboration ahead
with respect to MetaBelgica.&lt;/p&gt;
&lt;h2 id=&#34;selection-of-presentations&#34;&gt;Selection of presentations&lt;/h2&gt;
&lt;p&gt;There were tons of interesting presentations,
but unfortunately I could not split myself to see all.
Here is a small selection of some presentations from various domains
such that you will get an idea of research data infrastructure solutions at CoRDI.&lt;/p&gt;
&lt;h3 id=&#34;social-sciences&#34;&gt;Social sciences&lt;/h3&gt;
&lt;p&gt;I have learned about the term &lt;em&gt;microdata&lt;/em&gt;:
information at the level of individual response of surveys
(used in social sciences or official statistics).
Dana Müller presented work about geographically separated safe rooms as part of the International Data Access Network (IDAN).
In those rooms, data can be accessed without the having to leave the host institution (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.269&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.269&lt;/a&gt;).
When looking at the following picture I am remembered of my MetaBelgica presentation,
I am wondering if with respect to safe rooms there is a Belgian gap in social sciences too :-)&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Safe rooms for the International Data Access Network, but not in Belgium&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-saferoom_hu011811e9e7fd9529280b530948ddcc0f_4128653_3a1ce3c6d07d6e4b1316021b5baf121f.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-saferoom_hu011811e9e7fd9529280b530948ddcc0f_4128653_a977268b0791250c35deebebd504d6ae.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-saferoom_hu011811e9e7fd9529280b530948ddcc0f_4128653_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-saferoom_hu011811e9e7fd9529280b530948ddcc0f_4128653_3a1ce3c6d07d6e4b1316021b5baf121f.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;data-spaces&#34;&gt;Data spaces&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://orcid.org/0000-0003-3538-0106&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Peter Wittenburg&lt;/a&gt;
presented the idea of FAIR Data Objects (FDOs),
which according to his vision should become similarly to TCP/IP
a standard protocol for data exchange
(&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.263&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.263&lt;/a&gt;).
Like this, &lt;strong&gt;FAIR data can flow through the emerging data spaces&lt;/strong&gt;
similar to how TCP/IP packages flow through the internet.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Peter Wittenburg presenting FAIR Data Objects at CoRDI&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-fdo_hu80b87f035b05197ec681f947dd7bbbe8_4518016_3244e8e9395437c83b594885e473089f.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-fdo_hu80b87f035b05197ec681f947dd7bbbe8_4518016_7239e1e3633d98fb345375b8da8b351f.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-fdo_hu80b87f035b05197ec681f947dd7bbbe8_4518016_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-fdo_hu80b87f035b05197ec681f947dd7bbbe8_4518016_3244e8e9395437c83b594885e473089f.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;physics&#34;&gt;Physics&lt;/h3&gt;
&lt;p&gt;Not my cup of tea,
but apparently there are a lot of physical sciences data infrastructures in the UK.
And as we have learned further in the presentation
given by &lt;a href=&#34;https://orcid.org/0000-0001-8286-3835&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Nicola Knight&lt;/a&gt;,
there are also alignments with the NFDI (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.339&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.339&lt;/a&gt;).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Nicola Knight presenting UK physical sciences data infrastructures&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-uk_huf8f72ecd046890f57c31fba8977c16bb_3943463_f6d378f76accb0de83faea1a1ff4d3e4.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-uk_huf8f72ecd046890f57c31fba8977c16bb_3943463_8c9c8a7e0dc3176c4406fba28bca5786.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-uk_huf8f72ecd046890f57c31fba8977c16bb_3943463_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-uk_huf8f72ecd046890f57c31fba8977c16bb_3943463_f6d378f76accb0de83faea1a1ff4d3e4.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;scholarly-communication&#34;&gt;Scholarly Communication&lt;/h3&gt;
&lt;p&gt;Even though we have all the open and FAIR data and code repositories,
the actual way of scholarly communication
via research papers has not change much over the centuries.
&lt;a href=&#34;https://orcid.org/0000-0002-0698-2864&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sören Auer&lt;/a&gt;
presented the &lt;em&gt;Open Research Knowledge Graph (ORKG)&lt;/em&gt; and how they plan to change that
(&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.272&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.272&lt;/a&gt;).
One adds structured information about the contributions of a specific paper to the Knowledge Graph
via a web interface.
&lt;strong&gt;Comparisons and visualizations can be created on the fly&lt;/strong&gt;,
everything gets DOIs and is thus citable.
They also experimented already to promote the ORKG earlier in the publication process,
for example as preparation for the submission to the SEMANTiCS conference
according to  Sören Auer.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Scholarly Communication via publications has not change much over the centuries, the Open Research Knowledge Graph wants to change that&#34; srcset=&#34;
               /media/cordi-2023/2023-09-14_cordi-orkg_hu1af4bfa65a138588679f486af0b79faf_2943545_6c942d0d5daf1becca7bfb93752ae227.webp 400w,
               /media/cordi-2023/2023-09-14_cordi-orkg_hu1af4bfa65a138588679f486af0b79faf_2943545_bc444af27c3d768c82f4bce7e0e8ff1e.webp 760w,
               /media/cordi-2023/2023-09-14_cordi-orkg_hu1af4bfa65a138588679f486af0b79faf_2943545_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-14_cordi-orkg_hu1af4bfa65a138588679f486af0b79faf_2943545_6c942d0d5daf1becca7bfb93752ae227.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;research-data-management&#34;&gt;Research Data Management&lt;/h3&gt;
&lt;p&gt;A whole team of people presented the state initiatives for Research Data Management (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.242&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.242&lt;/a&gt;).
The following image shows some German states that are already participating in this initiative.
They do some necessary field-work, providing support to researchers directly,
or support smaller members such as Universities of Applied Sciences in grant applications.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Supporting researchers with Research Data Management&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-rdm_hu471d30b10e8bda80792103da7528cb5e_3826730_f194bad0d0778acad4c6435947e3af45.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-rdm_hu471d30b10e8bda80792103da7528cb5e_3826730_7f55556b4dd33cd758955a4f3fc9b7b8.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-rdm_hu471d30b10e8bda80792103da7528cb5e_3826730_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-rdm_hu471d30b10e8bda80792103da7528cb5e_3826730_f194bad0d0778acad4c6435947e3af45.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;networking&#34;&gt;Networking&lt;/h3&gt;
&lt;p&gt;Last but not least,
&lt;a href=&#34;https://orcid.org/0000-0001-6967-9443&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Nina Leonie Weisweiler&lt;/a&gt; presented the engagement of the Helmholtz association in NFDI (&lt;a href=&#34;https://doi.org/10.52825/cordi.v1i.402&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.52825/cordi.v1i.402&lt;/a&gt;).
The following image nicely shows which Helmholtz centre is active in which NFDI consortium.
Internally, the different NFDI representatives of Helmholtz regularly sit together
to discuss their involvement, great initiative!&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Connecting the dots: the engagement of different Helmholtz centres in the NFDI&#34; srcset=&#34;
               /media/cordi-2023/2023-09-13_cordi-helmholtz-links_hue78778ae3ab2af4f8018aab217d50136_4025007_6c01953dc253f60b4217d3b2001000b8.webp 400w,
               /media/cordi-2023/2023-09-13_cordi-helmholtz-links_hue78778ae3ab2af4f8018aab217d50136_4025007_8f36cd882d818efd305c5004136933be.webp 760w,
               /media/cordi-2023/2023-09-13_cordi-helmholtz-links_hue78778ae3ab2af4f8018aab217d50136_4025007_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-13_cordi-helmholtz-links_hue78778ae3ab2af4f8018aab217d50136_4025007_6c01953dc253f60b4217d3b2001000b8.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;

Connecting the dots as the title mentions, or connecting communities, the theme of the conference!&lt;/p&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;This is it, I hope you enjoyed reading this blog post about CoRDI 2023.
Let me know if you have any questions, feedback or remarks.&lt;/p&gt;
&lt;p&gt;Thanks to the whole CoRDI team for making this edition happen!
I am looking forward to a possible 2025 edition of CoRDI to learn about the progress of the community.&lt;/p&gt;
&lt;p&gt;In the meantime, there will be another NFDI event in 2024, but apparently slightly smaller, because the combined NFDI and EOSC communities are too big for a suitable venue.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;2024: NFDI meets EOSC&#34; srcset=&#34;
               /media/cordi-2023/2023-09-14_cordi-eosc-2024_hu4c17a49f0edf32237d7036ccc3425e48_4639399_8e665f5d8fe7420581a847d9103298b6.webp 400w,
               /media/cordi-2023/2023-09-14_cordi-eosc-2024_hu4c17a49f0edf32237d7036ccc3425e48_4639399_1a5dc3723a8b0a148ee9a47b0c52a557.webp 760w,
               /media/cordi-2023/2023-09-14_cordi-eosc-2024_hu4c17a49f0edf32237d7036ccc3425e48_4639399_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/cordi-2023/2023-09-14_cordi-eosc-2024_hu4c17a49f0edf32237d7036ccc3425e48_4639399_8e665f5d8fe7420581a847d9103298b6.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Don&amp;rsquo;t forget to share this post on social media and in your networks!
Please also consider subscribing to my bi-weekly Newsletter &lt;em&gt;FAIR Data Digest&lt;/em&gt; to receive more interesting content every other Tuesday!&lt;/p&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;</description>
    </item>
    
    <item>
      <title>RDF named graphs</title>
      <link>https://sven-lieber.org/en/2023/06/26/rdf-named-graphs/</link>
      <pubDate>Mon, 26 Jun 2023 19:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2023/06/26/rdf-named-graphs/</guid>
      <description>&lt;p&gt;With the Resource Description Framework (RDF) you can represent Linked Data as subject-predicate-object triples.
But what if you have to represent additional context of those triples?
In this Blog post I will briefly introduce the advanced data management feature &amp;ldquo;named graphs&amp;rdquo;
with two examples: different data source named graphs in the BELTRANS project
and different story-contexts for Digital Humanities research.
Furthermore I will graphically illustrate the different ways you can query named graphs with SPARQL!&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;Linked Data represented using the Resource Description Framework (RDF) consists of subject, predicate and object triples!
This is what you will learn in any introduction to Linked Data.&lt;/p&gt;
&lt;p&gt;However, did you know that with RDF you can also have &lt;em&gt;Quads&lt;/em&gt;?
Statements with four components: subject, predicate, object and &lt;em&gt;context&lt;/em&gt;.
Continue reading for some practical examples and graphical illustrations.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter FAIR Data Digest.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h2 id=&#34;example-of-named-graphs-for-provenance-information&#34;&gt;Example of named graphs for provenance information&lt;/h2&gt;
&lt;p&gt;According to the &lt;em&gt;Data Management&lt;/em&gt; chapter of the &lt;a href=&#34;https://patterns.dataincubator.org/book/named-graphs.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Linked Data Patterns&lt;/a&gt; website,
named graphs can be used for different purposes such as expressing the provenance of triples, use the named graph for access control, or in general for data management.&lt;/p&gt;
&lt;p&gt;In the &lt;a href=&#34;https://www.kbr.be/en/projects/beltrans/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELTRANS&lt;/a&gt; project where we collect data from various sources,
we store the data of each data source in a separated named graph.
One the one hand, this allows us to update a specific data source independent of other data sources,
and on the other hand, we can refer to a triple by its data source as the following example illustrates.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Named graph example from the BELTRANS project: one named graph per data source and one for the integrated project database&#34; srcset=&#34;
               /media/named-graphs/named-graphs-beltrans_hu3a2c8d3d87314a1d76db6e8af5bdfc9f_35217_609b73e2782b320418eab04671214674.webp 400w,
               /media/named-graphs/named-graphs-beltrans_hu3a2c8d3d87314a1d76db6e8af5bdfc9f_35217_c436830011034685f45051372ef7686b.webp 760w,
               /media/named-graphs/named-graphs-beltrans_hu3a2c8d3d87314a1d76db6e8af5bdfc9f_35217_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/named-graphs/named-graphs-beltrans_hu3a2c8d3d87314a1d76db6e8af5bdfc9f_35217_609b73e2782b320418eab04671214674.webp&#34;
               width=&#34;519&#34;
               height=&#34;408&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;We have a named graph containing RDF triples that describe book translations in our project database (&lt;code&gt;&amp;lt;http://project-db&amp;gt;&lt;/code&gt;).
Each of this book descriptions might be available in data
that we obtained from different national libraries (our data sources &lt;code&gt;&amp;lt;http://kbr-data&amp;gt;&lt;/code&gt; and &lt;code&gt;&amp;lt;http://bnf-data&amp;gt;&lt;/code&gt;).
This is represented by having a &lt;code&gt;schema:sameAs&lt;/code&gt; link from the book description
in the project named graph to the different data source named graphs.&lt;/p&gt;
&lt;p&gt;Without named graphs, the generic &lt;code&gt;schema:sameAs&lt;/code&gt; link would not allow us
to distinguish the descriptions of the book in one particular data source by using SPARQL.&lt;/p&gt;
&lt;p&gt;&lt;em&gt;A side note: it might be possible by using a FILTER expression on the string representation of the URI.
For example, if URI starts with &lt;code&gt;http://data.bnf.fr&lt;/code&gt; it probably is a book from the National Library of France.
However, this is a bit against the idea of the Semantic Web,
because now we assume semantics based on the name of the thing and not via explicitly expressed semantics with RDF terms. This is considered a bad practice.&lt;/em&gt;&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;With named graphs, however, we have a way to semantically express the context of a book description using explicit RDF statements.&lt;/strong&gt;
In a SPARQL query we could do the following if named graphs are used for data of each data source:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-fallback&#34; data-lang=&#34;fallback&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;SELECT ?bookURI ?kbrURI ?bnfURI
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;WHERE {
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  # we are interested in book descriptions from our project database
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  graph &amp;lt;http://project-db&amp;gt; { ?bookURI a schema:CreativeWork . }
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  # based on the named graph context, identify the KBR book description
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  OPTIONAL { 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;    graph &amp;lt;http://project-db&amp;gt; { ?bookURI schema:sameAs ?kbrURI . }
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;    graph &amp;lt;http://kbr-data&amp;gt; { ?kbrURI a schema:CreativeWork . }
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  }
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt; 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  # based on the named graph context, identify the BnF book description 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  OPTIONAL {  
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;    graph &amp;lt;http://project-db&amp;gt; { ?bookURI schema:sameAs ?bnfURI . } 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;    graph &amp;lt;http://bnf-data&amp;gt; { ?bnfURI a schema:CreativeWork . } 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;  }
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;}
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;In this example, the specific URI linked via &lt;code&gt;schema:sameAs&lt;/code&gt;
also has to be  a book description in a specific named graph.
Like this, we can distinguish in the SPARQL query based on the named graph context,
in which &lt;code&gt;schema:sameAs&lt;/code&gt; we are interested in.&lt;/p&gt;
&lt;h2 id=&#34;illustrations-of-different-ways-to-query-named-graph-data&#34;&gt;Illustrations of different ways to query named graph data&lt;/h2&gt;
&lt;p&gt;You have already seen that I have used the &lt;code&gt;graph&lt;/code&gt; keyword in the SPARQL query above.
But there are more possibilities!
To illustrate these possibilities, namely the keywords &lt;code&gt;FROM&lt;/code&gt; and &lt;code&gt;FROM NAMED&lt;/code&gt;, I will consider a slightly different example,
inspired from a presentation at the &lt;a href=&#34;https://sven-lieber.org/en/2023/06/05/dhbenelux-2023/#linked-open-data-for-greek-mythology&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DHBenelux conference 2023&lt;/a&gt;
about Greek Mythology.&lt;/p&gt;
&lt;p&gt;In the following example, we are interested in the fictional character john and his place of birth according to different novels.
Each novel is represented as a different named graph, i.e. &lt;code&gt;&amp;lt;http://story-1&amp;gt;&lt;/code&gt;, &lt;code&gt;&amp;lt;http://story-2&amp;gt;&lt;/code&gt;, and &lt;code&gt;&amp;lt;http://story-3&amp;gt;&lt;/code&gt;.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;An illustration of three named graphs&#34; srcset=&#34;
               /media/named-graphs/named-graphs_huf978d9de39b9eb3a618b65bdc8eb1029_20709_13fd338bbb6ef8a1c877693eda169649.webp 400w,
               /media/named-graphs/named-graphs_huf978d9de39b9eb3a618b65bdc8eb1029_20709_ddd9a8ae9e1ba294c3db6de7bc09a94b.webp 760w,
               /media/named-graphs/named-graphs_huf978d9de39b9eb3a618b65bdc8eb1029_20709_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/named-graphs/named-graphs_huf978d9de39b9eb3a618b65bdc8eb1029_20709_13fd338bbb6ef8a1c877693eda169649.webp&#34;
               width=&#34;519&#34;
               height=&#34;408&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Somewhere in the first novel, it is mentioned that john was born in &lt;code&gt;ex:town1&lt;/code&gt;.
The second novel even specifies that &lt;code&gt;ex:town1&lt;/code&gt; is located in country &lt;code&gt;ex:country2&lt;/code&gt;.
However, the second novel does also state that john was born in &lt;code&gt;ex:town2&lt;/code&gt;.
Last but not least,
according to the third novel, the city &lt;code&gt;ex:town1&lt;/code&gt; is in country &lt;code&gt;ex:country1&lt;/code&gt;.&lt;/p&gt;
&lt;p&gt;This example should illustrate that each novel provides its own context.
The john of novel 1 might not be the same of the john in novel 2.
Similarly, one could say that &lt;code&gt;ex:town1&lt;/code&gt; is the same as &lt;code&gt;ex:town2&lt;/code&gt;
or that &lt;code&gt;ex:country1&lt;/code&gt; is the same as &lt;code&gt;ex:country2&lt;/code&gt;.&lt;/p&gt;
&lt;p&gt;But, are these entities actually &lt;em&gt;the same&lt;/em&gt;?
This might be a complex and even philosophical discussion.
Instead of explicitly indicating that something is the same as something else,
we can simply keep the context of each story.
What we know is that &lt;code&gt;ex:john&lt;/code&gt; appears in all three novels.&lt;/p&gt;
&lt;p&gt;Let&amp;rsquo;s assume, that the fact that john is a (fictional) Person (e.g. by indicating &lt;code&gt;a schema:Person&lt;/code&gt;) is specified without any specific context (named graph).
The tricky part is now how to query the data.
Depending on your use case you have different options which are visualized below.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Different SPARQL queries on named graphs: SELECT, SELECT FROM, SELECT FROM NAMED&#34; srcset=&#34;
               /media/named-graphs/named-graphs-queries_hu3fcbd5ed3e13daf3be49113ecc5160c8_202148_ccbdca5b5bab45d86e8ebff8205ebb80.webp 400w,
               /media/named-graphs/named-graphs-queries_hu3fcbd5ed3e13daf3be49113ecc5160c8_202148_52a13e1c570c5b6e138bf33dc8408d88.webp 760w,
               /media/named-graphs/named-graphs-queries_hu3fcbd5ed3e13daf3be49113ecc5160c8_202148_1200x1200_fit_q75_h2_lanczos_3.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/named-graphs/named-graphs-queries_hu3fcbd5ed3e13daf3be49113ecc5160c8_202148_ccbdca5b5bab45d86e8ebff8205ebb80.webp&#34;
               width=&#34;760&#34;
               height=&#34;269&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In the illustration we see that each query returns different results.
Let&amp;rsquo;s have a look what each of the queries does.
One remark: The values in the output that are strikethrough will not be part of the result,
because not every variable of the triple pattern will be bound.
If you would wrap the triple pattern to fetch the country into an &lt;code&gt;OPTIONAL&lt;/code&gt;,
you would get those results too.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;(1)&lt;/strong&gt; In the first case no graph is explicitly specified. Thus the &lt;em&gt;default graph&lt;/em&gt; will be queried.
Depending on the triple store (RDF database) you are using,
the &lt;em&gt;default graph&lt;/em&gt; might be &lt;em&gt;yet another&lt;/em&gt; named graph which in our example is almost empty.
Remember, it only contains the triple that &lt;code&gt;ex:john&lt;/code&gt; is a person.
However, in this example we assume that the &lt;em&gt;default graph&lt;/em&gt;
is a union of all named graphs: an inclusive strategy.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;(2)&lt;/strong&gt; In the second case, we explicitly specify with the &lt;code&gt;FROM&lt;/code&gt; clause,
which named graphs the query should consider.
If you do not specify any context in the triple patterns of the query,
the query looks into the union of the specified named graphs.
As expected, data of &lt;code&gt;&amp;lt;http://story-3&amp;gt;&lt;/code&gt; is not queried and will thus not be part of the output.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;(3)&lt;/strong&gt; In the third case, we similarly specify named graphs explicitly,
this case with the &lt;code&gt;FROM NAMED&lt;/code&gt; clause.
Like this the query loops over each of the mentioned named graphs.
This means, that when producing one result set (one row of the output),
the variable for the named graph is bound to only one named graph.
Therefore it is not possible to query &lt;em&gt;across&lt;/em&gt; named graphs.
And this is exactly why for the given data,
no results will be provided.&lt;/p&gt;
&lt;p&gt;As you can see, a named graph is also just a URI!
Therefore we can use properties to describe the context as well.
For example&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-fallback&#34; data-lang=&#34;fallback&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;&amp;lt;http://story-1&amp;gt; a schema:Book ; 
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;                 schema:about ex:john ;
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;                 schema:author ex:author .
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;p&gt;And where you store this information is up to you.
Now you know how you can retrieve it with SPARQL!&lt;/p&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;Named graphs can be a powerful data management technique.
Yet you have to be aware of different implementations,
especially if you change your software stack at some point.
Additionally, it makes SPARQL queries slightly more complex which may look scary for beginners.&lt;/p&gt;
&lt;p&gt;Besides the RDF and SPARQL specifications, you can find more information about named graphs
at &lt;a href=&#34;https://web.archive.org/web/20230124111048/https://blog.metaphacts.com/the-default-graph-demystified&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;this Blog post&lt;/a&gt; from the company metaphacts,
or at &lt;a href=&#34;https://web.archive.org/web/20220704093443/https://www.stardog.com/labs/blog/from-vs-from-named-in-sparql/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;this Blog post&lt;/a&gt; from the company Stardog.&lt;/p&gt;
&lt;p&gt;Curious about more content like this or some behind-the-scenes?
Consider subscribing to my bi-weekly Newsletter &lt;em&gt;FAIR Data Digest&lt;/em&gt;
to receive more interesting content about Linked Data every other Tuesday!&lt;/p&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;</description>
    </item>
    
    <item>
      <title>Knowledge Graphs for Data Integration - UHasselt 2023</title>
      <link>https://sven-lieber.org/en/2023/06/12/knowledge-graphs-for-data-integration-uhasselt-2023/</link>
      <pubDate>Mon, 12 Jun 2023 09:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2023/06/12/knowledge-graphs-for-data-integration-uhasselt-2023/</guid>
      <description>&lt;p&gt;In June 2023 I attended a workshop in Hasselt, Belgium, organized by a research network on Knowledge Graphs and Data Integration.
Linked Data-based systems and declarative SPARQL queries somehow magically do all the work efficiently, but how?
For many people, presentations about &lt;em&gt;using&lt;/em&gt; Linked Data and SPARQL already seem very technical.
But this event went further, it covered the fundamental algorithms that power the Linked Data systems!
In this post I try my best to briefly summarize some of the talks and highlight the added value of those fundamental and often behind-the-scenes work.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;The &lt;a href=&#34;https://www.fwo.be/en/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FWO&lt;/a&gt;-founded research network &lt;a href=&#34;https://u0152642.pages.gitlab.kuleuven.be/kg4di-fwo-network/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;&lt;em&gt;Knowledge Graphs and Data Integration&lt;/em&gt;&lt;/a&gt;
organized a &lt;a href=&#34;https://www.uhasselt.be/knowledge_graphs_kg4di&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;workshop&lt;/a&gt;
at the &lt;a href=&#34;https://www.uhasselt.be/en/instituten-en/dsi&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Data Science Institute&lt;/a&gt; of Hasselt University in Belgium.
As part of this workshop, a keynote about &lt;em&gt;a logical approach to model interpretability&lt;/em&gt; was given
by &lt;a href=&#34;https://marceloarenas.cl/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marcelo Arenas&lt;/a&gt; from the &lt;em&gt;Pontificia Universidad Catolica de Chile&lt;/em&gt;.
Afterwards, several papers were presented in three different sessions: foundations, experience and exemplars, and processing and mining.
I will group my summaries in these categories as well.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;My overall takeaway&lt;/strong&gt; is that there is still plenty of research for Knowledge Graphs!
Not only for how to use them or how to encode knowledge in them,
but also for how to &lt;em&gt;actually implement&lt;/em&gt; Knowledge Graph software
based on formally sound techniques.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;


&lt;details class=&#34;toc-inpage d-print-none  &#34; open&gt;
  &lt;summary class=&#34;font-weight-bold&#34;&gt;Table of Contents&lt;/summary&gt;
  &lt;nav id=&#34;TableOfContents&#34;&gt;
  &lt;ul&gt;
    &lt;li&gt;&lt;a href=&#34;#context&#34;&gt;Context&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#keynote&#34;&gt;Keynote&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#explaining-predictions-of-ml-models&#34;&gt;Explaining predictions of ML models&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#a-logic-based-query-language-for-ml-models&#34;&gt;A logic-based query language for ML models&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#foundations&#34;&gt;Foundations&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#graph-compression&#34;&gt;Graph compression&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#effective-reasoning&#34;&gt;Effective reasoning&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#reliable-query-languages&#34;&gt;Reliable query languages&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#experience-and-exemplars&#34;&gt;Experience and exemplars&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#data-dependencies&#34;&gt;Data dependencies&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#rapid-prototyping&#34;&gt;Rapid Prototyping&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#technology-stacks-knowledge-graph&#34;&gt;Technology stacks Knowledge Graph&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#law-as-code&#34;&gt;Law as code&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#processing-and-mining&#34;&gt;Processing and mining&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#mining-software-engineering-patterns&#34;&gt;Mining software engineering patterns&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#more-efficient-sparql-queries&#34;&gt;More efficient SPARQL queries&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#more-on-query-optimization&#34;&gt;More on query optimization&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#final-remarks&#34;&gt;Final remarks&lt;/a&gt;&lt;/li&gt;
  &lt;/ul&gt;
&lt;/nav&gt;
&lt;/details&gt;

&lt;h2 id=&#34;keynote&#34;&gt;Keynote&lt;/h2&gt;
&lt;p&gt;Before I focus on the keynote itself I would like to start with an example.
Recently in a show of the Flemish comedian &lt;a href=&#34;https://www.lievenscheire.be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Lieven Scheire&lt;/a&gt;,
I heard about a Machine Learning (ML) algorithm that was trained to distinguish images
from wolves and from huskies. The algorithm seemed to perform very well, but not always.
Later, by trying out a few things, it turned out the algorithm apparently &amp;ldquo;learned&amp;rdquo; that the distinguishing factor is a white background (snow).
So if there was white background, the animal on the image was classified as a husky image by the ML model.&lt;/p&gt;
&lt;p&gt;This is just an example to illustrate that we don&amp;rsquo;t actually know what a ML algorithm/model has learned and thus &lt;em&gt;why&lt;/em&gt; a certain output is generated.
That&amp;rsquo;s why such models are usually called a &lt;em&gt;black box&lt;/em&gt;. But in a world in which more and more ML enters our daily lives,
&lt;strong&gt;it is absolutely fundamental that we are able to explain why the model behaves in a certain way.&lt;/strong&gt;&lt;/p&gt;
&lt;h3 id=&#34;explaining-predictions-of-ml-models&#34;&gt;Explaining predictions of ML models&lt;/h3&gt;
&lt;p&gt;In simple terms, the keynote of &lt;a href=&#34;https://marceloarenas.cl/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marcelo Arenas&lt;/a&gt;
presented some fundamental research in early stages that tries to specify a declarative query language (think of SQL or SPARQL) for ML models.
The keynote lecture had some very nice examples to illustrate the graph algorithms.
But why graphs actually?&lt;/p&gt;
&lt;p&gt;Machine Learning models are usually based on so-called &amp;ldquo;neural networks&amp;rdquo;
that are inspired by neurons in the human brain.
These neural networks  that may consist of several layers of neurons
&amp;ldquo;learn&amp;rdquo; patterns by adjusting how information flows through the connections between the nodes.
This is not a Knowledge Graph in a Linked Data-based sense
in which you explicitly encode information that you can query later.
&lt;strong&gt;But it is a graph and hence one can perform graph algorithms that include some sort of queries.&lt;/strong&gt;&lt;/p&gt;
&lt;h3 id=&#34;a-logic-based-query-language-for-ml-models&#34;&gt;A logic-based query language for ML models&lt;/h3&gt;
&lt;p&gt;In his talk, Marcelo explained how they define theoretical logics for operations on such graphs. Their overarching goal is to develop a declarative query language, such as SPARQL, but for neural networks.
This is not purely theoretical as they also implemented their &amp;ldquo;FOIL&amp;rdquo; logic using a special logic programming language that is different to languages such as Python or Java that you may know. Like this the theories can be tested with real life data and it can be observed how effective they are with specific data (besides the theoretical complexity).
If you are interested in more, here you can find a preprint of an earlier paper: &lt;a href=&#34;https://arxiv.org/abs/2110.02376&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://arxiv.org/abs/2110.02376&lt;/a&gt;&lt;/p&gt;
&lt;h2 id=&#34;foundations&#34;&gt;Foundations&lt;/h2&gt;
&lt;p&gt;From this session, I will focus on three of the talks. Two that focus on effective computations
and one about reliable query languages.&lt;/p&gt;
&lt;h3 id=&#34;graph-compression&#34;&gt;Graph compression&lt;/h3&gt;
&lt;p&gt;The work of &lt;a href=&#34;https://twitter.com/jbinero&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Jeroen Bollen&lt;/a&gt;
is motivated by the fact that memory is often a bottleneck for so-called Graph Neural Networks.
There are ways to compress the graphs.
But so far there haven&amp;rsquo;t been any formal proofs showing that the accuracy of the algorithm
remains stable with differently compressed graphs.&lt;/p&gt;
&lt;p&gt;In his talk, Jeroen presented his work where he showed that
they could minimize computational and memory requirements
while keeping a similar accuracy.
Very interesting work that demonstrates that &lt;strong&gt;fundamental research can make a difference!&lt;/strong&gt;&lt;/p&gt;
&lt;h3 id=&#34;effective-reasoning&#34;&gt;Effective reasoning&lt;/h3&gt;
&lt;p&gt;Other more formal work inspired by a real world use case was presented by
&lt;a href=&#34;https://be.linkedin.com/in/robin-de-vogelaere-50b898100&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Robin De Vogelaere&lt;/a&gt;.
You probably have used &lt;em&gt;glue&lt;/em&gt; (a cohesive) at least once in you life.
Selecting the right adhesive in the right situation is not trivial.
It depends on the material and in which conditions the product will be used (heat, shock, etc).&lt;/p&gt;
&lt;p&gt;Robin presented the &lt;em&gt;FO(.)&lt;/em&gt; first-order logic
that they use in their &lt;a href=&#34;https://www.idp-z3.be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;IDP/Z3&lt;/a&gt; Knowledge Base system
to perform reasoning (on cohesives).
In his presentation he highlighted why the Knowledge Graph Community should care (see image below),
&lt;strong&gt;A perfect example of the knowledge transfer the workshop aimed for!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Robin De Vogelaere presents reasons why the Knowledge Graph community should care about the FO dot logic&#34; srcset=&#34;
               /media/kg4di-2023/2023-06-09_kg4di-why-kg-should-care_hu23c6e23ad05e1683207e6e4894ea6de3_3329862_a6878a6e4e74172711318478b9213a91.webp 400w,
               /media/kg4di-2023/2023-06-09_kg4di-why-kg-should-care_hu23c6e23ad05e1683207e6e4894ea6de3_3329862_32cc66685c5a8f76fbf503e852a1caa2.webp 760w,
               /media/kg4di-2023/2023-06-09_kg4di-why-kg-should-care_hu23c6e23ad05e1683207e6e4894ea6de3_3329862_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kg4di-2023/2023-06-09_kg4di-why-kg-should-care_hu23c6e23ad05e1683207e6e4894ea6de3_3329862_a6878a6e4e74172711318478b9213a91.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;reliable-query-languages&#34;&gt;Reliable query languages&lt;/h3&gt;
&lt;p&gt;Staying in the Linked Data RDF world:
You might have used SPARQL already to query something from Wikidata via the &lt;a href=&#34;https://query.wikidata.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Wikidata Query Service&lt;/a&gt;.
But using some baked-in SPARQL functionality, you can actually write a &lt;em&gt;federated SPARQL Query&lt;/em&gt;
that fetches different parts of the data from different servers.&lt;/p&gt;
&lt;p&gt;But what if one crucial FILTER part of the query relies on the data of a server
that is currently offline?
Should the whole query fail or would you accept data from the other servers
that could not have been filtered (and which are wrong for your use case)?&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://baccaert.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Tim Baccaert&lt;/a&gt; presented some early work on exactly those problems.
He is looking into different ways to make query languages more reliable.
For example by providing better machine-understandable error messages
such that clients can interpret partial results better.&lt;/p&gt;
&lt;p&gt;This is &lt;em&gt;one way&lt;/em&gt; of dealing better with the reality where servers are down sometimes.
Another way is &lt;a href=&#34;https://linkeddatafragments.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Linked Data Fragments (LDF)&lt;/a&gt;:
Instead of processing the whole possibly complex SPARQL query on the server,
the LDF server only accepts simple patterns
and the computation of the result is done on the server &lt;em&gt;and&lt;/em&gt; the client.&lt;/p&gt;
&lt;h2 id=&#34;experience-and-exemplars&#34;&gt;Experience and exemplars&lt;/h2&gt;
&lt;p&gt;This session is about experiences in setting up Knowledge Graphs
and in using the data.&lt;/p&gt;
&lt;h3 id=&#34;data-dependencies&#34;&gt;Data dependencies&lt;/h3&gt;
&lt;p&gt;Understanding relationships between attributes of relational data
is important to integrate data in a meaningful way.
&lt;a href=&#34;https://be.linkedin.com/in/marcel-parciak-b412171ab&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marcel Parciak&lt;/a&gt;
presented work on detecting (approximate) &lt;em&gt;functional dependencies&lt;/em&gt; in data.&lt;/p&gt;
&lt;p&gt;By performing a literature review before their experiments,
they could identify 12 measures to detect possible dependencies.
They performed an evaluation of the different measures
on benchmark datasets.
Before doing that they also had to annotate the benchmark data
because there were no labels for
when there is a functional dependency and when not.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://scholar.google.com/citations?user=FFZa8_QAAAAJ&amp;amp;hl=en&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Larissa C. Shimomura&lt;/a&gt;
is researching another kind of dependency: Graph Generating Dependencies.
Such dependencies are helpful to understand the data and possible correlations between different attributes.
In particular, Larissa investigates reasoning on property graphs.&lt;/p&gt;
&lt;h3 id=&#34;rapid-prototyping&#34;&gt;Rapid Prototyping&lt;/h3&gt;
&lt;p&gt;Declarative solutions following common standards are the goal
when working with Knowledge Graphs in RDF.
However, especially in the beginning of a project
the requirements are often vague
and different stakeholders need to build a common understanding.&lt;/p&gt;
&lt;p&gt;In his talk, &lt;a href=&#34;https://chrdebru.github.io/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Christophe Debruyne&lt;/a&gt;
presented work about the TOXIN Knowledge Graph.
The aim of this Knowledge Graph is to describe
existing safety data about cosmetic ingredients.
It therefore contributes to non-animal systemic toxicity assessments.&lt;/p&gt;
&lt;p&gt;In the beginning phase of the project many things were set up
with Python.
This helped in getting a shared understanding between the stakeholders
before setting up a complete infrastructure.
In later stages, R2RML and other declarative tools
such as SHACL for quality assessments were used.&lt;/p&gt;
&lt;h3 id=&#34;technology-stacks-knowledge-graph&#34;&gt;Technology stacks Knowledge Graph&lt;/h3&gt;
&lt;p&gt;While talking about infrastructure,
the technology stacks of cloud platforms are quite complex.
This means that switching between providers can become really cumbersome.&lt;/p&gt;
&lt;p&gt;In his presentation, &lt;a href=&#34;https://researchportal.vub.be/en/persons/johannes-h%C3%A4rtel&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Johannes Härtel&lt;/a&gt;
talked about how they used RDF to describe technology stacks.
Furthermore, they have implemented a tool that harvests data
about the software stacks from GitHub repositories.&lt;/p&gt;
&lt;p&gt;For example, if you have a docker file,
the script detects this and will add to the Knowledge Graph
that you at least use YAML files and
that you have at least one docker setup provided.&lt;/p&gt;
&lt;p&gt;The goal of their work is to allow for a smoother switch
between technology stacks.
For example, to identify Open Source technologies that fit
the current (proprietary) technology stack you are using.
At the core, they used Apache Jena to store RDF.
Johannes listed four reasons for why to use Knowledge Graphs (see image below).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Johannes Härtel presents 4 reasons why to use Knowledge Graphs&#34; srcset=&#34;
               /media/kg4di-2023/2023-06-09_kg4di-kg-reasons_hu59cd17f98f51697bd6d1df042207d800_3283321_1498a350d6ce592f19eea97396f43c3d.webp 400w,
               /media/kg4di-2023/2023-06-09_kg4di-kg-reasons_hu59cd17f98f51697bd6d1df042207d800_3283321_89c0b29d4e25eb583184d9198b3e7eb7.webp 760w,
               /media/kg4di-2023/2023-06-09_kg4di-kg-reasons_hu59cd17f98f51697bd6d1df042207d800_3283321_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kg4di-2023/2023-06-09_kg4di-kg-reasons_hu59cd17f98f51697bd6d1df042207d800_3283321_1498a350d6ce592f19eea97396f43c3d.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h3 id=&#34;law-as-code&#34;&gt;Law as code&lt;/h3&gt;
&lt;p&gt;The transparency of legislation and the accountability of governments are very important
for a democracy.
In his presentation,
&lt;a href=&#34;https://tomdenies.be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Tom De Nies&lt;/a&gt;
presented the tools &lt;a href=&#34;https://themis.vlaanderen.be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Themis&lt;/a&gt;
and &lt;a href=&#34;https://kaleidos.vlaanderen.be/aanmelden&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Kaleidos&lt;/a&gt;
which are used to support the legislation process in
Flanders, Belgium with Linked Data.&lt;/p&gt;
&lt;p&gt;The clue?
There is no data conversion needed, their system works with RDF at its core!
Data is stored, accessed and published from a RDF triple store.
Input forms and other website components are created
by using an &lt;a href=&#34;https://emberjs.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;EmberJS&lt;/a&gt; frontend
that relies on a rich microservice infrastructure in the backend.&lt;/p&gt;
&lt;h2 id=&#34;processing-and-mining&#34;&gt;Processing and mining&lt;/h2&gt;
&lt;p&gt;Last but not least,
the last session focused on mining information
and effective query execution.&lt;/p&gt;
&lt;h3 id=&#34;mining-software-engineering-patterns&#34;&gt;Mining software engineering patterns&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://cu.linkedin.com/in/yunior-pacheco-correa&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Yunior Pacheco Correa&lt;/a&gt;
presented work on mining software engineering patterns from code repositories.
This can help to support a developer with auto-completion
when writing code, such as &lt;a href=&#34;https://github.com/features/copilot&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;GitHub Copilot&lt;/a&gt;.
Another possibility would be to query the mined patterns
to for example look for security vulnerabilities.&lt;/p&gt;
&lt;h3 id=&#34;more-efficient-sparql-queries&#34;&gt;More efficient SPARQL queries&lt;/h3&gt;
&lt;p&gt;Three presentations from the IDLab &lt;a href=&#34;https://knows.idlab.ugent.be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;KNoWS group&lt;/a&gt;
focused on the behind-the-scenes of querying with SPARQL.
SPARQL is a declarative query language: the user specifies &lt;em&gt;what&lt;/em&gt; to retrieve and not &lt;em&gt;how&lt;/em&gt;.
Internally, the used SPARQL query engine
processes the query and builds a query plan.
For example, it determines how to join result patterns.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://ca.linkedin.com/in/bryanelliotttam/en&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Bryan-Elliott Tam&lt;/a&gt;
presented work related to &lt;em&gt;guided link traversal query processing&lt;/em&gt;.
So basically, how a RDF query interface can already provide hints
about its data via the &lt;a href=&#34;https://treecg.github.io/specification/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;TREE hypermedia specification&lt;/a&gt;,
such that the query engine can make better choices in query planning.
Similar research was presented by Jonni Hansi.&lt;/p&gt;
&lt;p&gt;Bryan&amp;rsquo;s research has shown that with SPARQL FILTER expressions and the TREE specification,
they were able to prune links to Linked Data fragments that will certainly not contribute to the SPARQL query.
This means the query will be faster!&lt;/p&gt;
&lt;p&gt;Ruben H. Eschauzier presented early work on
a ML model to learn effective joins in query planning.
As a modular part of the &lt;a href=&#34;https://comunica.dev/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Comunica&lt;/a&gt; query engine,
the presented model can be used to predict the execution time of a join
and tries to select the most efficient one.
Currently this model only outperforms the Comunica templates
in 7 out of 18 templates.
There is thus still more research needed!&lt;/p&gt;
&lt;h3 id=&#34;more-on-query-optimization&#34;&gt;More on query optimization&lt;/h3&gt;
&lt;p&gt;The research of &lt;a href=&#34;https://research.tue.nl/en/persons/wilco-van-leeuwen&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Wilco van Leeuwen&lt;/a&gt;
centers around &lt;em&gt;cardinality estimation&lt;/em&gt;:
a technique used to estimate the number of items returned by a query or subquery.
He is interested in obtaining the most accurate answers using
the least amount of resources (time and space).&lt;/p&gt;
&lt;p&gt;Compared to relational data, graph data introduce additional challenges
for cardinality estimation.
His research helps understanding different cardinality estimation techniques and their trade-offs.
This could for example be applied to improve query planning on graph data,
such as for SPARQL query planning.&lt;/p&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;Thanks to the organizers for planning this event
and to the researchers to present their work!
I think this has been a fruitful exchange of knowledge.
I am looking forward to the following events.&lt;/p&gt;
&lt;p&gt;Don&amp;rsquo;t forget to share this post on social media and in your networks!&lt;/p&gt;
&lt;p&gt;Please also consider subscribing to my bi-weekly newsletter &lt;em&gt;FAIR Data Digest&lt;/em&gt;
to receive more interesting content every other Tuesday!&lt;/p&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;</description>
    </item>
    
    <item>
      <title>DHBenelux 2023</title>
      <link>https://sven-lieber.org/en/2023/06/05/dhbenelux-2023/</link>
      <pubDate>Mon, 05 Jun 2023 09:00:00 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2023/06/05/dhbenelux-2023/</guid>
      <description>&lt;p&gt;This year we had the honor to host the 10th anniversary edition of the DHBenelux conference at the Royal Library of Belgium (KBR).
In this blog post I will reflect on the different sessions I&amp;rsquo;ve attended at this year&amp;rsquo;s multitrack edition of DHBenelux 2023. I will group my notes around the following four recurring topics of the talks that I have seen: (1) data platforms and their usability, (2) how digital methods such as Linked Open Data are used for DH research, the other way around, (3) how DH research studies digital methods and tools, and finally (4) research with (historical) geospatial data in cultural heritage.&lt;/p&gt;
&lt;h2 id=&#34;context&#34;&gt;Context&lt;/h2&gt;
&lt;p&gt;Digital Humanities (DH) is a discipline in which computational methods are combined with humanities research, both by systematically using digital methods and artifacts for humanities research and by studying the application of said methods.&lt;/p&gt;
&lt;p&gt;To foster DH research in Belgium, the Netherlands and Luxembourg (Benelux), the &lt;a href=&#34;https://dhbenelux.org/about/founding-statement/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DHBenelux conference&lt;/a&gt; was founded in 2014. Although the focus lies on these three countries, the conference is open to anyone interested and the conference language is English.&lt;/p&gt;
&lt;p&gt;This year’s edition had several tracks and because I have a technical background and mainly work with data and research data infrastructure,
I mainly followed more technical sessions.
However, if you are interested in other talks, the abstracts of the conference are openly available in the &lt;a href=&#34;https://zenodo.org/communities/dhbenelux2023&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DHBenelux 2023 community&lt;/a&gt; on the research repository Zenodo (persistent), or on the &lt;a href=&#34;https://2023.dhbenelux.org/schedule/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DHBenelux 2023 website&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;My main takeaways are&lt;/strong&gt;&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Digital methods support the humanities research, but humanities methodologies can also positively influence digital tools and methods&lt;/li&gt;
&lt;li&gt;the use cases of the diverse humanities research provide excellent food for thought for data management practices, for example what kind of biases exist and what good practices are&lt;/li&gt;
&lt;li&gt;the conference aims at community-building and experience exchange which I find is the greatest distinction to the more competitive Computer Science conferences I’ve attended in the past (which have a much lower acceptance rate of for example 25% compared to the 90% of DHBenelux)&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Because the conference took place at my workplace, I did not travel to the conference. This also meant that I commuted home and did not join most of the social programme (#partypooper). Therefore, and also because I volunteered during the conference for the organizers that are my colleagues, I will focus in this blog post mainly on the scientific presentations and discussions and less on the venue.&lt;/p&gt;
&lt;p&gt;Instead of listing the talks chronologically, I will group my notes under the following recurring topics I have encountered in the talks that I have attended.&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Data platforms and usability&lt;/li&gt;
&lt;li&gt;Using humanities methodologies for digital resources&lt;/li&gt;
&lt;li&gt;Using digital methods for humanities research&lt;/li&gt;
&lt;li&gt;Geospatial research&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;There were also a few workshop the day before the conference, but I did not participate in any of the workshops as non of the workshops sparked my immediate interest and I anyway had some work to finish 🙂
This was only my second DHBenelux conference, but it seems that a combination of workshops about DH tools and Linked Open Data are common. This is great!&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;


&lt;details class=&#34;toc-inpage d-print-none  &#34; open&gt;
  &lt;summary class=&#34;font-weight-bold&#34;&gt;Table of Contents&lt;/summary&gt;
  &lt;nav id=&#34;TableOfContents&#34;&gt;
  &lt;ul&gt;
    &lt;li&gt;&lt;a href=&#34;#context&#34;&gt;Context&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#data-platforms-and-usability&#34;&gt;Data platforms and usability&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#stakeholder-readiness-level-in-the-clariah-media-suite&#34;&gt;Stakeholder Readiness Level in the CLARIAH Media Suite&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#difference-between-scholar-and-journalist-users&#34;&gt;Difference between scholar and journalist users&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#lessons-learned-when-building-a-consortium-of-stakeholders---luxtime-project&#34;&gt;Lessons learned when building a consortium of stakeholders - LuxTIME project&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#turn-photos-into-items&#34;&gt;Turn photos into items&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#collection-centric-approach&#34;&gt;Collection-centric approach&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#visualizations-and-user-interactions&#34;&gt;Visualizations and user interactions&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#using-humanities-methodologies-for-digital-resources&#34;&gt;Using humanities methodologies for digital resources&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#keynote-about-cosmovision&#34;&gt;Keynote about Cosmovision&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#the-expressiveness-of-wikidata&#34;&gt;The expressiveness of Wikidata&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#uncertainty-when-labeling-images&#34;&gt;Uncertainty when labeling images&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#using-digital-methods-for-humanities-research&#34;&gt;Using digital methods for humanities research&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#linked-open-data-for-greek-mythology&#34;&gt;Linked Open Data for Greek mythology&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#linked-data-thesauri-to-study-the-dutch-east-india-company&#34;&gt;Linked Data thesauri to study the Dutch East India Company&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#persistent-isni-identifiers-for-research-and-legal-deposit&#34;&gt;Persistent ISNI identifiers for research and legal deposit&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#geospatial-research&#34;&gt;Geospatial research&lt;/a&gt;
      &lt;ul&gt;
        &lt;li&gt;&lt;a href=&#34;#closing-keynote-about-geographic-information-systems&#34;&gt;Closing keynote about Geographic Information Systems&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#historical-belgian-places&#34;&gt;Historical Belgian places&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#historical-greek-places&#34;&gt;Historical Greek places&lt;/a&gt;&lt;/li&gt;
        &lt;li&gt;&lt;a href=&#34;#visualizing-location-data-in-books&#34;&gt;Visualizing location data in books&lt;/a&gt;&lt;/li&gt;
      &lt;/ul&gt;
    &lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#poster-session&#34;&gt;Poster session&lt;/a&gt;&lt;/li&gt;
    &lt;li&gt;&lt;a href=&#34;#final-remarks&#34;&gt;Final remarks&lt;/a&gt;&lt;/li&gt;
  &lt;/ul&gt;
&lt;/nav&gt;
&lt;/details&gt;

&lt;h2 id=&#34;data-platforms-and-usability&#34;&gt;Data platforms and usability&lt;/h2&gt;
&lt;p&gt;Humanities research revolves around a lot of valuable and super interesting data such as Libraries’ collections, digitized documents, or other media data. Different platforms and tools exist to work with this data in a user-friendly way. I would say even more than in Computer Science, user-friendly interfaces are a must for less-technical users in the Humanities. Therefore it is nice to see different efforts around usability.&lt;/p&gt;
&lt;p&gt;There have been talks about usability of the &lt;a href=&#34;https://web.archive.org/web/20230507112751/https://mediasuite.clariah.nl/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;CLARIAH Media Suite&lt;/a&gt; in the Netherlands,
the upcoming &lt;a href=&#34;https://web.archive.org/web/20230315161436/https://www.kbr.be/en/projects/data-kbr-be/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DATA-KBR&lt;/a&gt; platform in Belgium,
as well as the &lt;a href=&#34;https://web.archive.org/web/20230327210711/https://www.timemachine.eu/ltm-projects/luxtime/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;LuxTIME&lt;/a&gt;
and the &lt;a href=&#34;https://web.archive.org/web/20230603010341/https://www.tropy.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Tropy&lt;/a&gt; projects in Luxembourg.
On top of that, other talks focused on data visualization.&lt;/p&gt;
&lt;h3 id=&#34;stakeholder-readiness-level-in-the-clariah-media-suite&#34;&gt;Stakeholder Readiness Level in the CLARIAH Media Suite&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://nl.linkedin.com/in/roelandordelman&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Roeland Ordelman&lt;/a&gt; presented User-related research around the CLARIAH Media Suite.
What I really liked is the idea to build something similar to the
&lt;a href=&#34;https://web.archive.org/web/20230601115210/https://en.wikipedia.org/wiki/Technology_readiness_level&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Technology Readiness Level (TRL)&lt;/a&gt;
but for usability by different stakeholders: the so-called Stakeholder Readiness Level (SRL).
I find this a very interesting and holistic approach that, on the one hand, can improve the usability of systems, and on the other hand could be used as a benchmark to compare the usability of systems.&lt;/p&gt;
&lt;p&gt;My own research on usability has been a while, but I can clearly see how it could be connected or extended. For example, how a Stakeholder Readiness Level could be used to measure the customer journey when using public services (&lt;a href=&#34;https://doi.org/10.5281/zenodo.8001638&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.8001638&lt;/a&gt;, &lt;a href=&#34;https://sven-lieber.org/en/project/fast/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;FAST project&lt;/a&gt;) or to define requirements for a user interface that uses visual notations to visualize constraints of data constraint languages (&lt;a href=&#34;https://content.iospress.com/articles/semantic-web/sw210450&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.3233/SW-210450&lt;/a&gt;).&lt;/p&gt;
&lt;h3 id=&#34;difference-between-scholar-and-journalist-users&#34;&gt;Difference between scholar and journalist users&lt;/h3&gt;
&lt;p&gt;Besides the presentation of Roeland about the general Stakeholder Readiness Level, &lt;a href=&#34;https://www.uu.nl/staff/WSanders&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Willemien Sanders&lt;/a&gt; talked about differences between scholars and journalists  when creating data stories with the media suite.&lt;/p&gt;
&lt;p&gt;There are differences with respect to speed, transparency, reflection, data/service provider involvement and sensitivity/copyright issues. More details about the differences are in the abstract. In my opinion this kind of research is very important. Not only to understand the user needs better, but also to improve the digital product and to come up with suitable tools and business models.&lt;/p&gt;
&lt;h3 id=&#34;lessons-learned-when-building-a-consortium-of-stakeholders---luxtime-project&#34;&gt;Lessons learned when building a consortium of stakeholders - LuxTIME project&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/petrosaposto&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Petros Apostolopoulos&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/stakats&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sean Takats&lt;/a&gt; shared their experiences from the creation of a consortium to analyze and interpret historical big data within the Luxembourg Time machine (LuxTIME) project &lt;a href=&#34;https://doi.org/10.5281/zenodo.7986361&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7986361&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;They performed interviews that focused on the following four aspects: historical data, research initiatives, collaborations with other institutions, and future collaboration with the project and participation in the consortium.&lt;/p&gt;
&lt;p&gt;These interviews were necessary, because detailed information about what historical data is held at the institutions
or in which research initiatives the institutions are involved is often not available on their websites. The objective is to map the state of the art of historical data in Luxembourg.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Some lessons learned from the LuxTIME project about setting up a consortium&#34; srcset=&#34;
               /media/dhbenelux-2023/2023-06-02_dhbenelux-luxtime_hu3f27aa7dbdf84d57c1dcb44f075bdfd8_4415755_934e95dadc204d13c1e2f001b6a4a6fd.webp 400w,
               /media/dhbenelux-2023/2023-06-02_dhbenelux-luxtime_hu3f27aa7dbdf84d57c1dcb44f075bdfd8_4415755_0d598e390757a0148fde2895eb95db5c.webp 760w,
               /media/dhbenelux-2023/2023-06-02_dhbenelux-luxtime_hu3f27aa7dbdf84d57c1dcb44f075bdfd8_4415755_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/dhbenelux-2023/2023-06-02_dhbenelux-luxtime_hu3f27aa7dbdf84d57c1dcb44f075bdfd8_4415755_934e95dadc204d13c1e2f001b6a4a6fd.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Interestingly, some institutions do not consider their existing digitization initiatives as research projects,
because, according to these institutions, these initiatives do not involve research.&lt;/p&gt;
&lt;p&gt;One objective is not to complicate collaboration even further by needing to select representatives,
but rather create a peer-to-peer collaboration. Additionally, regular meetings should be organized alternately at the different consortium members locations to strengthen the collaboration.&lt;/p&gt;
&lt;h3 id=&#34;turn-photos-into-items&#34;&gt;Turn photos into items&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/alucchesi/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Anita Lucchesi&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/stakats&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sean Takats&lt;/a&gt; presented the &lt;a href=&#34;https://tropy.org&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Tropy&lt;/a&gt; platform. It helps in organizing research collections: imagine that you have a bunch of images in a folder, with Tropy you can annotate and organize the images. Among others, this allows at a later stage to export the items in standardized data formats.&lt;/p&gt;
&lt;h3 id=&#34;collection-centric-approach&#34;&gt;Collection-centric approach&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/schambers3&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sally Chambers&lt;/a&gt; presented updates on the DATA-KBR project that aims to optimize KBRs ICT infrastructure to stimulate sustainable data access. This contribution is therefore not only about an upcoming platform (for now internally a Nextcloud instance is used), but also about the collections themselves: collections as data.&lt;/p&gt;
&lt;p&gt;The updates she presented focus on a researcher-centered and iterative approach to gather requirements.
If you are interested to read more about how to publish collections as data,
she referred to the pre-print &lt;em&gt;A Checklist to Publish Collections as Data in GLAM Institutions&lt;/em&gt; &lt;a href=&#34;https://doi.org/10.48550/arXiv.2304.02603&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.48550/arXiv.2304.02603&lt;/a&gt;&lt;/p&gt;
&lt;p&gt;Focussing on collections and data instead of being too application-centric made me think of the amazing book &lt;a href=&#34;https://www.goodreads.com/en/book/show/38744362&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Software Wasteland&lt;/a&gt; from &lt;a href=&#34;https://www.linkedin.com/in/davemccomb/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Dave McComb&lt;/a&gt;. After all we want to study cultural heritage collections!&lt;/p&gt;
&lt;p&gt;I got a similar vibe from a question that was asked during the talk of &lt;a href=&#34;https://twitter.com/lorenverreyen&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Loren Verreyen&lt;/a&gt;
about her previous work on Radio’s ephemerality in DH.
There she used textual data from the &lt;a href=&#34;https://genome.ch.bbc.co.uk/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BBC programme index&lt;/a&gt; to study BBC producers of the Third Programme (1946 - 1967). Digitized documents from a broadcast listing magazine were used as a basis.&lt;/p&gt;
&lt;p&gt;A question from the audience, whether corrected OCR data of the research project was fed back to the BBC portal, got my attention. According to Loren, this has not yet happened but would be very useful. I like this idea of a more collection-centric approach where different smaller or larger research projects continuously improve or extend a collection.&lt;/p&gt;
&lt;p&gt;This made me also think of the &lt;a href=&#34;https://www.kb.nl/en/onderzoeken-%26-vinden/researcher-residence&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Researcher-in-residence&lt;/a&gt; programme of the Royal Library of The Netherlands.
In this programme, researchers have to use existing digital collections to perform research and improve or extend these collections (or create new collections/datasets). Like this, the collections get richer over time. By the way, as of writing this post, they are looking for a new researcher-in-residence, check out their website!&lt;/p&gt;
&lt;p&gt;Another very interesting talk from &lt;a href=&#34;https://www.douwezeldenrust.nl/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Douwe Zeldenrust&lt;/a&gt; about the Records Continuum Model highlighted the evolution of collections.&lt;/p&gt;
&lt;p&gt;In his talk, he presented meta research about the Dutch language in the United States that originally was performed by &lt;a href=&#34;https://www.wikidata.org/wiki/Q3296598&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Jo Daan&lt;/a&gt;. More precisely, Douwe focused on the archiving process of research data. He did so based on the example of Jo Daan’s collection and several other collections that were added later during other PhD projects which partially built their collection on top of the one from Jo or criticized the original collection.&lt;/p&gt;
&lt;h3 id=&#34;visualizations-and-user-interactions&#34;&gt;Visualizations and user interactions&lt;/h3&gt;
&lt;p&gt;When talking about users and usability, visualizations are always important!&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/aida_hor&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Aida Horaniet Ibañez&lt;/a&gt; investigated the current situation of data visualization in DH (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7963466&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7963466&lt;/a&gt;).
She collected a corpus of more than 200 figures from research papers.
After noticing that there was a predominance of statistical graphs, they extended the corpus with 500 extra figures from online magazines, social media, etc.&lt;/p&gt;
&lt;p&gt;Basically the presented research differentiates between the following three types of visualizations:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Statistical graphics (reducing cognitive load by visualizing reliable data sources)&lt;/li&gt;
&lt;li&gt;Data humanism (in-depth details about context where design and topic stimulates the user’s interest)&lt;/li&gt;
&lt;li&gt;Humanistic interpretation (visualizations with user interaction)&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Apparently, only a few tools support the latter two of the types.
One interesting finding was that many visualizations from personal blogs are unavailable after a while.
In my opinion, yet another good reason for heritage institutions to invest in web and social media harvesting
(like we did at KBR with the &lt;a href=&#34;https://web.archive.org/web/20230505105310/https://www.kbr.be/en/projects/besocial/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BESOCIAL project&lt;/a&gt;, or like the &lt;a href=&#34;https://netpreserve.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;International Internet Preservation Consortium (IIPC)&lt;/a&gt; does on a large scale).&lt;/p&gt;
&lt;h2 id=&#34;using-humanities-methodologies-for-digital-resources&#34;&gt;Using humanities methodologies for digital resources&lt;/h2&gt;
&lt;p&gt;One fascinating thing about DH is that digital methods are not only used within DH, but humanities methodologies are also used to study digital resources. I think this provides new and interesting ways to look at the technology we are building and using!&lt;/p&gt;
&lt;p&gt;There have been presentations about post colonial heritage, Wikidata’s data model and inclusivity, and best practices when annotating human characters in children’s literature according to ethnicity and gender.&lt;/p&gt;
&lt;h3 id=&#34;keynote-about-cosmovision&#34;&gt;Keynote about Cosmovision&lt;/h3&gt;
&lt;p&gt;This keynote from &lt;a href=&#34;https://twitter.com/patymurrieta&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Patricia Murrieta-Flores&lt;/a&gt; focused on decolonial praxis in the Digital Humanities. She talked about languages in the Mesoamerican area. Apparently, there existed more than 68 languages before the “Conquest of America”! Yet most we know was written by Europeans.&lt;/p&gt;
&lt;p&gt;My main takeaway is the following paraphrased quote:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;Be aware of the role of our work as DH researchers, there are also other perspectives in the world that are equally wonderful and which should be taken into account.&amp;rdquo; - Patricia Murrieta-Flores (paraphrased)&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;Most of the existing texts in her field are written from one perspective only: written by the colonizers and not the indigenous people. Partricia made the following comparison: Imagine all books about France and its culture would have been written by the British and in English. This would probably give a different view on things.&lt;/p&gt;
&lt;p&gt;According to Patricia, this is also a crucial aspect for machine learning: there is a lack of training material for certain topics and definitely from different worldviews, which raises the question if AI is safe to use!&lt;/p&gt;
&lt;p&gt;And if you wonder what Cosmovision is (I have to admit I did not know it):
According to &lt;a href=&#34;https://web.archive.org/web/20230531211528/https://www.lawinsider.com/dictionary/cosmovision&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Law Insider&lt;/a&gt;,&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;Cosmovision means the conception that indigenous peoples have, both collectively and individually, of the physical and spiritual world and the environment in which they conduct their lives.&amp;rdquo; - Law Insider&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;So in simple terms, it is about different worldviews.&lt;/p&gt;
&lt;h3 id=&#34;the-expressiveness-of-wikidata&#34;&gt;The expressiveness of Wikidata&lt;/h3&gt;
&lt;p&gt;Ammandeep K Mahal and &lt;a href=&#34;https://twitter.com/j_w_baker&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;James Baker&lt;/a&gt; presented their work on &lt;em&gt;histories of women active in archaeology, history and heritage&lt;/em&gt;
in the &lt;a href=&#34;https://web.archive.org/web/20230603043349/https://beyondnotability.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BeyondNotability&lt;/a&gt; project. They curated and publish a dataset in a &lt;a href=&#34;https://beyond-notability.wikibase.cloud/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Wikibase instance&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;Their presentation focuses on the free and open knowledge base &lt;a href=&#34;https://www.wikidata.org/wiki/Wikidata:Main_Page&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Wikidata&lt;/a&gt;.
Mainly they criticize the following three points (summarized in my understanding):&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;Wikidata statements do not accommodate for complex historical phenomena, such as (historical) women life’s combining paid work and care work&lt;/li&gt;
&lt;li&gt;Ethnicity and citizenship are likewise not represented in a sufficient detail, even worse, bots often assume citizenship based on other properties and change the values&lt;/li&gt;
&lt;li&gt;Sex and gender in Wikidata are &lt;a href=&#34;https://www.wikidata.org/wiki/Property:P21&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;sexOrGender&lt;/a&gt;, (a single property)&lt;/li&gt;
&lt;/ol&gt;
&lt;p&gt;My opinion: Without considering any context about the meaning of the property itself, I have to say that the third point seems already questionable: a property with the word “or” in the name should, according to me, be split into several properties!&lt;/p&gt;
&lt;p&gt;A key message of their talk was that Wikidata has a clean and simple interface which suggests simple and authoritative facts. However, the talk pages behind the interface may provide more controversial discussions that the average user might not be aware of.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;I think:&lt;/strong&gt;
Data models can be made as complex as necessary. In the vision of the Semantic Web and Linked Open Data,
anyone can create their own ontology to accommodate any world view in detail.
The practice, however, has taught us that a few broad and simplified solutions often get the most attention.
For data models, for example the generic &lt;a href=&#34;https://web.archive.org/web/20230605093502/https://schema.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;schema.org vocabulary&lt;/a&gt;
and for data, apparently, the Wikidata platform that does not even have a formal ontology.&lt;/p&gt;
&lt;p&gt;This simplicity makes such solutions appealing but indeed amplifies certain world views and leaves out important context.
I think that setting up a domain-specific Wikibase with its own vocabulary,
as the authors did, is for now the best solution.
A solution that resonates the most with the original Semantic Web vision of many decentralized ontologies (even though in this case it is not a formal ontology).&lt;/p&gt;
&lt;h3 id=&#34;uncertainty-when-labeling-images&#34;&gt;Uncertainty when labeling images&lt;/h3&gt;
&lt;p&gt;The talk of &lt;a href=&#34;https://be.linkedin.com/in/paavo-van-der-eecken&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Paavo van der Eecken&lt;/a&gt; about the use of digital tools for uncertain humanities data was very interesting. In a nutshell, the talk covered considerations when creating labels used to annotate and categorize human characters on images in children&amp;rsquo;s books. I have learned that it is very important to use labels that are as specific as possible when dealing with ethnicity and gender to avoid biased generalizations.&lt;/p&gt;
&lt;p&gt;But I also learned that, when labeling images,
there is a difference between &lt;em&gt;ontological doubt&lt;/em&gt; and &lt;em&gt;epistemic doubt&lt;/em&gt;.
The former refers to unclear boundaries and the latter to incomplete knowledge.
Knowing these theoretical foundations helps, on a very practical level, to plan, schedule and execute projects with human annotators and to achieve the best possible outcome.&lt;/p&gt;
&lt;h2 id=&#34;using-digital-methods-for-humanities-research&#34;&gt;Using digital methods for humanities research&lt;/h2&gt;
&lt;p&gt;Many talks actually covered this kind of topics as it is a core part of DH research. However, since it was a multi track conference, I only attended a few talks that applied digital methods for humanities research.&lt;/p&gt;
&lt;p&gt;I have attended talks that apply Linked Data for Greek mythology narratives or the Dutch East India Company, and persistent identifiers for library catalogs.&lt;/p&gt;
&lt;h3 id=&#34;linked-open-data-for-greek-mythology&#34;&gt;Linked Open Data for Greek mythology&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/dareiadareia&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Daria Kondakova&lt;/a&gt; and Jakob Kohler talked about their use of Linked Open Data to study methodological narratives.&lt;/p&gt;
&lt;p&gt;They have built a tool with a user interface to create narratives in the form of RDF triples, with the aim to compare mythological narratives. Currently they are investigating how they can express relationships between the narratives. For example, Zeus in story A is not the same as Zeus in story B.&lt;/p&gt;
&lt;p&gt;Thinking about it, one part of the solution could be to use the &lt;a href=&#34;https://www.w3.org/TR/prov-o/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;W3C PROV ontology&lt;/a&gt;
and specify that the different Zeus instances are a specialization of a more broad Zeus concept (PROV agents can also be PROV entities I believe).
Yet this would not solve the initial issue as there are still no relationships between the Zeus specializations.&lt;/p&gt;
&lt;p&gt;Maybe named-graphs might be a solution to encapsulate the different narratives. In such a scenario there should be only one Zeus URI, but it would be used in different named-graphs. One would not have to explicitly specify the relations between different Zeus URIs, each is embedded in the context of the narrative named-graph.&lt;/p&gt;
&lt;h3 id=&#34;linked-data-thesauri-to-study-the-dutch-east-india-company&#34;&gt;Linked Data thesauri to study the Dutch East India Company&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://www.huygens.knaw.nl/medewerkers/brecht-nijman/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Brecht Nijman&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/kw_pepping&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Kay Pepping&lt;/a&gt; presented their work in building a SKOS taxonomy &lt;a href=&#34;https://doi.org/10.5281/zenodo.7973694&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7973694&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;They want to unlock a key resource of the Dutch East India Company archive: the &lt;em&gt;Overgekomen Brieven en Papieren&lt;/em&gt; which are roughly 5mio digitized documents. These handwritten documents went through a Handwritten Text Recognition (HTR) step.&lt;/p&gt;
&lt;p&gt;However, due to text recognition errors or different spellings as well as synonyms, a simple text search for a concept is not sufficient.
Their example was the herb Clove (Dutch: kruidnagel) that could be written as &lt;em&gt;kruidnagel&lt;/em&gt;, &lt;em&gt;kruytnagel&lt;/em&gt;, &lt;em&gt;garioffelnagel&lt;/em&gt;, &lt;em&gt;moernagel&lt;/em&gt;, etc.&lt;/p&gt;
&lt;p&gt;Because existing thesauri are not detailed enough they created their own &lt;a href=&#34;https://web.archive.org/web/20230528043826/https://www.w3.org/2004/02/skos/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;SKOS&lt;/a&gt;
and &lt;a href=&#34;https://web.archive.org/web/20230511002253/https://www.w3.org/TR/skos-reference/skos-xl.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;SKOS-XL&lt;/a&gt; taxonomy with the tool &lt;a href=&#34;https://web.archive.org/web/20230531224650/https://www.poolparty.biz/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;PoolParty&lt;/a&gt;.&lt;/p&gt;
&lt;h3 id=&#34;persistent-isni-identifiers-for-research-and-legal-deposit&#34;&gt;Persistent ISNI identifiers for research and legal deposit&lt;/h3&gt;
&lt;p&gt;My colleague &lt;a href=&#34;https://twitter.com/annvancamp&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Ann Van Camp&lt;/a&gt; and our volunteer &lt;a href=&#34;https://www.linkedin.com/in/sergio-alonso-mislata-24646113b/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sergio Alonso Mislata&lt;/a&gt; presented the benefits of using a persistent ISNI identifier for contributors to creative works in the catalog of the Royal Library of Belgium.&lt;/p&gt;
&lt;p&gt;This helps not only in the research projects &lt;a href=&#34;https://www.kbr.be/en/projects/beltrans/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELTRANS&lt;/a&gt; and &lt;a href=&#34;https://www.kbr.be/en/projects/camille/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;CAMille&lt;/a&gt;, but also to identify relevant works for the &lt;a href=&#34;https://www.kbr.be/en/tag/legal-deposit/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;legal deposit&lt;/a&gt; in Belgium.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;A side note:&lt;/strong&gt;
Sergio addressed us after a conference last year, saying that he would like to know more about Linked Open Data.
We are always open to host volunteers or job students at KBR!&lt;/p&gt;
&lt;h2 id=&#34;geospatial-research&#34;&gt;Geospatial research&lt;/h2&gt;
&lt;p&gt;One particular type of DH data is geospatial data, think of digitized old maps or location information in metadata.
This again is something where DH shines, beautiful old maps are digitized and used.&lt;/p&gt;
&lt;p&gt;There was a keynote about geo metadata and large language models, and talks around historical location data in Belgium and in Greece as well as location data in written works.&lt;/p&gt;
&lt;h3 id=&#34;closing-keynote-about-geographic-information-systems&#34;&gt;Closing keynote about Geographic Information Systems&lt;/h3&gt;
&lt;p&gt;The closing keynote from &lt;a href=&#34;https://twitter.com/pirayendepiraye&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Piraye Hacıgüzeller&lt;/a&gt; provided thoughts on typological thinking and traditional Geographic Information System (GIS).
It mainly focused on the complexity of encoding geographic information on archaeological sites.&lt;/p&gt;
&lt;p&gt;They tried to encode this kind of information using &lt;a href=&#34;https://web.archive.org/web/20230601130243/https://cidoc-crm.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;CIDOC-CRM&lt;/a&gt;, but it was not expressive enough.
Furthermore, some geographical information is difficult to express on maps, but more easily in text due too ambiguity and fuzziness such as for directions.&lt;/p&gt;
&lt;p&gt;Piraye highlighted that with &lt;a href=&#34;https://web.archive.org/web/20230605115732/https://huggingface.co/alexbrandsen/ArcheoBERTje&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;ArcheoBERTje&lt;/a&gt;,
a sophisticated language model in Dutch language exists (based on the Dutch language BERT model BERTje).
According to research, this language model could not further be improved with domain-specific thesauri as it already covered all the relevant information.&lt;/p&gt;
&lt;p&gt;She ended the talk with the open question if there is still a need for metadata.
Unfortunately there was not much time for the QA.
But after the conference I checked again the presentation with the title &lt;a href=&#34;https://www.youtube.com/watch?v=WqYBx2gB6vA&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;The future of Knowledge Graphs in the World of Language Models&lt;/a&gt;
by one of Wikidata’s founders &lt;a href=&#34;https://twitter.com/vrandezo&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Denny Vrandečić&lt;/a&gt;.
And I think there are a few good points why there still is a need for curated metadata.
As Denny phrased it:&lt;/p&gt;
&lt;blockquote&gt;
&lt;p&gt;&amp;ldquo;In a world of infinite content, knowledge becomes valuable.&amp;rdquo; - Denny Vrandečić&lt;/p&gt;
&lt;/blockquote&gt;
&lt;p&gt;However, I get her point, especially after the explanation of the issue with fuzziness in geographic data in text. Thus who knows, for certain applications and use cases a language model might still be helpful and even preferable to metadata following a (too) strict schema.&lt;/p&gt;
&lt;h3 id=&#34;historical-belgian-places&#34;&gt;Historical Belgian places&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/LeaHermenault&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Léa Hermenault&lt;/a&gt; presented the Belgian Historical Gazetteer (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7997173&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7997173&lt;/a&gt;. Their work will create a database of historical place names that will facilitate the mapping of place names found in historical documents.&lt;/p&gt;
&lt;p&gt;They are using a PostgreSQL database with PostGIS extension to store the data. Furthermore, they are using the tool OpenRefine to enrich the data and create RDF. Eventually the data will reside in a &lt;a href=&#34;https://druid.datalegend.net/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Druid&lt;/a&gt; database hosted by &lt;a href=&#34;https://triply.cc/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Triply&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A CIDOC-CRM based model for the Belgian historical Gazetteer, credits to Vincent Ducatteeuw&#34; srcset=&#34;
               /media/dhbenelux-2023/2023-06-01_dhbenelux-belgian-gazetteer-model_hued9ed5b4e480695769a31152dc578150_2893084_df94481cdbedeacd2f1cea0af00e2f72.webp 400w,
               /media/dhbenelux-2023/2023-06-01_dhbenelux-belgian-gazetteer-model_hued9ed5b4e480695769a31152dc578150_2893084_8911d6f2d9c2992808cf02377fa2a4bb.webp 760w,
               /media/dhbenelux-2023/2023-06-01_dhbenelux-belgian-gazetteer-model_hued9ed5b4e480695769a31152dc578150_2893084_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/dhbenelux-2023/2023-06-01_dhbenelux-belgian-gazetteer-model_hued9ed5b4e480695769a31152dc578150_2893084_df94481cdbedeacd2f1cea0af00e2f72.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;This is a great project which also facilitates (upcoming) work at KBR. In the recently approved project MetaBelgica, KBR and three other Federal Scientific Institutes will combine their person, organization, events/time and location information into a Wikibase instance. The presented historical gazetteer is a great reference dataset.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/VDucatteeuw&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Vincent Ducatteeuw&lt;/a&gt; presented another project that focuses on historical Belgian geo data (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7945196&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7945196&lt;/a&gt;), more precisely on historical street names in the city of Ghent. Street names change over time and there is an ontological issue: if a street name has changed, is it still the same street or a new one?&lt;/p&gt;
&lt;p&gt;In his presentation, Vincent explained that they are using an identification methodology to distinguish streets by their &lt;em&gt;proper name&lt;/em&gt; and their location or type of mereology (&lt;a href=&#34;https://content.iospress.com/articles/applied-ontology/ao200235&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.3233/AO-200235&lt;/a&gt;).
Furthermore, they are applying this methodology on the &lt;em&gt;Wegwijzer der stad Gent&lt;/em&gt;, a city directory with scanned documents from 1875-1932.&lt;/p&gt;
&lt;h3 id=&#34;historical-greek-places&#34;&gt;Historical Greek places&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/laure_nmile&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Laura Soffiantini&lt;/a&gt; is pursuing the research question &lt;em&gt;How Greece is described in Natural History? How, why, and in relation to what does Pliny refer to Greek places?&lt;/em&gt;&lt;/p&gt;
&lt;p&gt;Therefore, she is using a multi-step workflow to extract place names from &lt;a href=&#34;https://topostext.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;TOPOSText&lt;/a&gt;, enriches is with the tool &lt;a href=&#34;https://github.com/flairNLP/flair&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;flair&lt;/a&gt; and disambiguates places with the help of an authoritative list from &lt;a href=&#34;https://www.trismegistos.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Trismegistos&lt;/a&gt; and the tool &lt;a href=&#34;https://github.com/seatgeek/thefuzz&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;TheFuzz&lt;/a&gt; (formerly known as FuzzyWuzzy).&lt;/p&gt;
&lt;h3 id=&#34;visualizing-location-data-in-books&#34;&gt;Visualizing location data in books&lt;/h3&gt;
&lt;p&gt;&lt;a href=&#34;http://baciu.online/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Dan Costa Baciu&lt;/a&gt; presented a prototype to visualize location data of well-known institutions, infrastructures, etc that are found in books.&lt;/p&gt;
&lt;p&gt;He argues that &lt;em&gt;geospatial discovery&lt;/em&gt; such as in GoogleFlights or AirBnB does not yet exist for library catalogs. But also that it would be useful! However with their platform it is possible. They are using different color to distinguish between location data of the book metadata (such as publisher location) and location data extracted from the content of the book.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The prototype presented by Dan Costa Bacui, that visualizes location data from book metdata and location data in the book&#34; srcset=&#34;
               /media/dhbenelux-2023/2023-06-02_dhbenelux-book-locations_hu2ae886869c9011d4659a78aae15a32c6_2738529_80d9834474ea8b573f1a117b54c019d8.webp 400w,
               /media/dhbenelux-2023/2023-06-02_dhbenelux-book-locations_hu2ae886869c9011d4659a78aae15a32c6_2738529_27225538e8696c5b7048c4d54fb32605.webp 760w,
               /media/dhbenelux-2023/2023-06-02_dhbenelux-book-locations_hu2ae886869c9011d4659a78aae15a32c6_2738529_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/dhbenelux-2023/2023-06-02_dhbenelux-book-locations_hu2ae886869c9011d4659a78aae15a32c6_2738529_80d9834474ea8b573f1a117b54c019d8.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;This is a very interesting and cool idea.
For the &lt;a href=&#34;https://web.archive.org/web/20230604163451/https://www.kbr.be/en/projects/beltrans/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELTRANS&lt;/a&gt; project I am working on,
which covers FR-NL / NL-FR translations,
this would be a really interesting way to visualize data.
Especially, since for certain comics, locations in the book are &lt;em&gt;culturally translated&lt;/em&gt; as well.
For example, Belgian cities of the Wallonian region in Belgium that got translated to Dutch cities for the translation for the Dutch-language market (translations that mainly target the larger Netherlands rather than Flanders).&lt;/p&gt;
&lt;p&gt;However, this would require extracting the location information of the books. Hence this is a bit out of scope of the quantitative study in BELTRANS. But it would be an interesting addition to refine the BELTRANS corpus in the future (see also my notes on collection-centric thinking here in this blog post).&lt;/p&gt;
&lt;h2 id=&#34;poster-session&#34;&gt;Poster session&lt;/h2&gt;
&lt;p&gt;There was of course also a poster session with many interesting posters.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A photo of the DHBenelux 2023 poster session&#34; srcset=&#34;
               /media/dhbenelux-2023/2023-06-01_dhbenelux-poster-session_hu90fd198076b7b5a814f55ae24fa9a40c_3813299_ca53057419c4e574d936c122a9d63827.webp 400w,
               /media/dhbenelux-2023/2023-06-01_dhbenelux-poster-session_hu90fd198076b7b5a814f55ae24fa9a40c_3813299_22729c65de73e72c64077a7e3745a523.webp 760w,
               /media/dhbenelux-2023/2023-06-01_dhbenelux-poster-session_hu90fd198076b7b5a814f55ae24fa9a40c_3813299_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/dhbenelux-2023/2023-06-01_dhbenelux-poster-session_hu90fd198076b7b5a814f55ae24fa9a40c_3813299_ca53057419c4e574d936c122a9d63827.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;One poster (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7976748&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7976748&lt;/a&gt;) was about the research in the &lt;a href=&#34;https://web.archive.org/web/20230604163451/https://www.kbr.be/en/projects/beltrans/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BELTRANS&lt;/a&gt; project for which I also work.
It was presented by two of the project&amp;rsquo;s PhDs students: &lt;a href=&#34;https://twitter.com/clarafo98&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Clara Folie&lt;/a&gt; and &lt;a href=&#34;https://be.linkedin.com/in/timothy-sirjacobs-89587520b&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Timothy Sirjacobs&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;Another interesting poster (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7986819&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7986819&lt;/a&gt;) was about the organization of a Hackathon at KU Leuven, presented by &lt;a href=&#34;https://twitter.com/leahcb&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Leah Budke&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;My ex &lt;a href=&#34;https://web.archive.org/web/20230505105310/https://www.kbr.be/en/projects/besocial/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BESOCIAL&lt;/a&gt;
colleague &lt;a href=&#34;https://twitter.com/FMessens&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Fien Messens&lt;/a&gt; presented a poster (&lt;a href=&#34;https://doi.org/10.5281/zenodo.7956417&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;DOI: 10.5281/zenodo.7956417&lt;/a&gt;)
on web archiving of religious websites at KU Leuven.&lt;/p&gt;
&lt;h2 id=&#34;final-remarks&#34;&gt;Final remarks&lt;/h2&gt;
&lt;p&gt;This is it, I hope you enjoyed reading this rather lengthy blog post about DHBenelux 2023.&lt;/p&gt;
&lt;p&gt;Thanks to my colleague &lt;a href=&#34;https://twitter.com/juliebirkholz&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Julie Birkholz&lt;/a&gt; and all the other organizers and volunteers for making this DHBenelux edition happen!
The presentations gave me a lot of food-for-thought and I enjoyed catching up with the community. See you at the next year’s edition at KU Leuven.&lt;/p&gt;
&lt;p&gt;Don&amp;rsquo;t forget to share this post on social media and in your networks!&lt;/p&gt;
&lt;p&gt;Please also consider subscribing to my bi-weekly Newsletter &lt;em&gt;FAIR Data Digest&lt;/em&gt; to receive more interesting content every other Tuesday!&lt;/p&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;</description>
    </item>
    
    <item>
      <title>My Phd explained</title>
      <link>https://sven-lieber.org/en/2020/06/21/my-phd-explained/</link>
      <pubDate>Sun, 21 Jun 2020 16:28:14 +0200</pubDate>
      <guid>https://sven-lieber.org/en/2020/06/21/my-phd-explained/</guid>
      <description>&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    A side note: This post describes my PhD work from the perspective of June 2020. The actual finished PhD in 2022 was slightly different in comparison.
  &lt;/div&gt;
&lt;/div&gt;
&lt;p&gt;It is particularly challenging to explain PhD research to family and friends,
especially in an abstract field such as Computer Science.
In the following blog post I will try to explain my PhD,
freed from academic precision and more from the big picture and with plenty of examples.&lt;/p&gt;
&lt;p&gt;The bullet points form one coherent story,
but each paragraph marked with an arrow can be clicked to read more details.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h1 id=&#34;the-context&#34;&gt;The context&lt;/h1&gt;
&lt;p&gt;Did you ever wonder how an assistance system like &lt;em&gt;Alexa&lt;/em&gt; or &lt;em&gt;Siri&lt;/em&gt;
can answer arbitrarily questions?
Therefore such a system needs to search data of different disciplines
also referred to as domains.&lt;/p&gt;
&lt;p&gt;To assist that, it makes sense to combine these data in a natural way,
you have to imagine it like a graph:
&lt;strong&gt;nodes&lt;/strong&gt; are things like concrete persons and &lt;strong&gt;connections&lt;/strong&gt;
between the nodes are the relationships between those things.
As text such a graph could look like this:&lt;/p&gt;
&lt;div class=&#34;highlight&#34;&gt;&lt;pre tabindex=&#34;0&#34; class=&#34;chroma&#34;&gt;&lt;code class=&#34;language-fallback&#34; data-lang=&#34;fallback&#34;&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;Sven hasBirthday 25.04.1988 .
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;Sven worksFor IDLab .
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;IDLab headQuartersIn Belgium .
&lt;/span&gt;&lt;/span&gt;&lt;span class=&#34;line&#34;&gt;&lt;span class=&#34;cl&#34;&gt;Belgium belongsTo Europe .
&lt;/span&gt;&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;details&gt;
&lt;summary&gt;
But first it has to be determined **what** can be connected **how** under which **restrictions**
and this is called a **data model**.
In this example my employer IDLab, which is an organization cannot have a date of birth
and similarly Sven cannot have headquarter as he is a person.
Generally speaking I&#39;m researching how such data models can be created.
&lt;/summary&gt;
Data modeling is nothing new, but in contrast to the past
where data was modeled for a single database or a single computer program
we make use of the world wide web.
Each concept and every relationship such as `person` and `headQuartersIn`
but also concrete data itself like `Sven` or `Belgium`
gets an own web address!
Thereby they are globally unique identifiable and computer programs
as well as users can look up the concept or the information!
This could look like on the following page: https://sven-lieber.org/profile
&lt;/details&gt;
&lt;h1 id=&#34;my-research---part-1-the-data-modeling-process&#34;&gt;My research - part 1: the data modeling process&lt;/h1&gt;
&lt;p&gt;Well, everyone can create such a data model in one afternoon,
this is not a big deal.
The challenge, however, is to make it professional so
in the end a qualitatively high data model pops out.
Ideally the data model should be constructed in a way
that it solves the problems it was created for in the first place!
For the question &amp;ldquo;How is the weather today in Ghent?&amp;rdquo;
the data model for example at least needs to represent
the concepts &lt;em&gt;weather&lt;/em&gt; and &lt;em&gt;town&lt;/em&gt; as well as &lt;em&gt;temporal information&lt;/em&gt;.&lt;/p&gt;
&lt;details&gt;
&lt;summary&gt;
Now a data model can be developed being easy and reusable,
or complex and tailored to solve a very concrete problem.
&lt;/summary&gt;
An example of an easy and reusable data model:
if you google any business, Google will show you on the right
within an infobox what the opening hours or the founding year of that business is.
Among others, Google can do this because the website owners used
a standardized data model to mark information in their website.
Not much precision is needed as the information is mainly shown to humans.
In contrast to this, in *biomedical* science very complex data models
following logical rules are created which are very precise such that also
computer programs can &#34;understand it&#34;.
&lt;/details&gt;
&lt;details&gt;
&lt;summary&gt;
Usually within the creation of a data model one has to
talk to officials and experts, analyze existing data
and have a clue about existing data models.
After all the data model should follow standards and
support seamless data exchange with other data models, called interoperability.
**All that is also a complex process which can be split into steps.**
&lt;/summary&gt;
By the way, this is the same for professional software engineering.
A measurable process subject to optimization distinguishes
software engineering from the simple act of *programming*.
&lt;/details&gt;
&lt;p&gt;The first part of my research revolves around this process.
I&amp;rsquo;m pursuing the question &lt;strong&gt;what is the best way to represent restrictions in a data model&lt;/strong&gt;,
so how one can create a data model which is not too simple but also not too complex for a given problem.
Restrictions could be on the one hand logical axioms like &lt;code&gt;all life forms have a birth date&lt;/code&gt;,
or on the other hand testable constraints important for data exchange and data quality such as &lt;code&gt;the birth date has to be smaller than the death date&lt;/code&gt;.&lt;/p&gt;
&lt;p&gt;Thereby I do not reinvent the wheel,
I carefully read the literature of knowledge modeling and I still do!
The goal is to reuse existing methods as much as possible
but to adapt them to the modeling of constraints.
Part of this is also to analyze which measurable pros and cons arise
if certain restrictions are either expressed as
logical axioms in the data model or as constraints.
One use case in my PhD are privacy regulations.
To check compliance with privacy regulations on the one hand
the data model needs to provide relevant information
and on the other hand constraints are possibly needed.&lt;/p&gt;
&lt;h1 id=&#34;my-research---part-2-user-friendly-visualizations&#34;&gt;My research - part 2: user-friendly visualizations&lt;/h1&gt;
&lt;p&gt;In my lab we mostly work in different projects together with industry, the government or institutions.
The experience of these projects revealed concrete problems regarding &lt;strong&gt;user-friendliness&lt;/strong&gt; of modeling languages:
A 2017 newly introduced language to represent constraints in graph data models is text-based,
user have to first learn it which might be tricky!&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Pictures are worth a thousand words.&lt;/strong&gt;
Humans possess an astonishing information processing system though, ..
no, I do not talk about the smartphone in the pocket but about the brain!
From cognitive science we know that text is processed slowly as we have to concentrate more.
In contrast to that shapes and colors are almost automatically processes
with no visible effort.
We work with knowledge graphs which are tangible, we take advantage of that!&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;We develop graphical languages&lt;/strong&gt; to describe constraints on our knowledge graphs.
Based on the assumption that such languages are &lt;em&gt;more intuitive&lt;/em&gt; compared to an alien text syntax
we examine in user studies which graphical languages are suited best.&lt;/p&gt;
&lt;h2 id=&#34;and-why-all-that&#34;&gt;And why all that?&lt;/h2&gt;
&lt;p&gt;&lt;strong&gt;Data without description is useless.&lt;/strong&gt;
Imagine you find a piece of paper full with numbers.
Without context you have no clue what these numbers mean.
Maybe somewhere a computer program exists which knows how to
read and interpret this data.
However, for you or other programs this piece of paper is useless!&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Adding descriptions provides context and makes the data self-describing
and hence interpretable.&lt;/strong&gt;
If you note on the piece of paper that the first column is &lt;em&gt;weight&lt;/em&gt; measured in &lt;em&gt;kilogram&lt;/em&gt;
and that the second column represents birth dates
and that both are characteristics of a person
whose name is noted in the third column you performed data modeling!
Additionally the data are now &lt;em&gt;self-described&lt;/em&gt; and
independent from specific computer programs.
By the way: This piece of paper could be representative for a file or a website.&lt;/p&gt;
&lt;p&gt;If data are properly described they can be combined in a &lt;strong&gt;meaningful way&lt;/strong&gt;
and such combined data can be searched quicker, easier and most importantly
in a uniform way &lt;strong&gt;which is then utilized by systems such as Alexa or Siri!&lt;/strong&gt;
A not-meaningful combination is for example the addition
of a temperature in degree celsius on my age in years.&lt;/p&gt;
&lt;h1 id=&#34;summarized&#34;&gt;Summarized&lt;/h1&gt;
&lt;p&gt;Based on what was read we can summarize:
&lt;strong&gt;I deal with data modeling.&lt;/strong&gt;
I pursue the question how we can build knowledge graphs which also comply with
requirements towards data quality and privacy.
In particular I investigate on the one hand methods of knowledge engineering,
so which steps need to be performed to create such knowledge graphs,
and on the other hand the user-friendly visualization of data constraints
such that also non-experts can create such constraints.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;So that&amp;rsquo;s it!
I hope I cleared some things out :-)&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;Below is a picture of the ISWC 2019 conference
where I presented my PhD plan both as presentation and as a poster (my trip report of this conference held in New Zealand is &lt;a href=&#34;https://sven-lieber.org/en/2019/11/05/iswc-2019/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Me presenting my poster at the ISWC doctoral consortium in 2019&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_ca54ba1c770e0164c0875be2c4ae89ff.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_726327b97a9c025dc3b6c69901e4f2f2.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_ca54ba1c770e0164c0875be2c4ae89ff.webp&#34;
               width=&#34;570&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>K-Cap 2019</title>
      <link>https://sven-lieber.org/en/2019/11/24/k-cap-2019/</link>
      <pubDate>Sun, 24 Nov 2019 21:09:40 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2019/11/24/k-cap-2019/</guid>
      <description>&lt;p&gt;At the end of November 2019 I could visit the sunny California to attend the Conference on Knowledge Capture (K-Cap).
A lot of interesting workshops, tutorials and talks covered topics such as the representation of knowledge in Wikidata,
approaches to transform tabular data to Linked Data, or the representation of scientific literature and processes as Knowledge Graphs.
Additionally I could present our work on MontoloStats. Continue reading for the full trip report!&lt;/p&gt;
&lt;p&gt;The &lt;a href=&#34;http://www.k-cap.org/2019/index.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;International Conference on Knowledge Capture&lt;/a&gt; (K-Cap) takes place every two years alternating with the EKAW conference.
This year it took place in Marina Del Rey, CA in the United States.
I traveled directly from New Zealand were &lt;a href=&#34;https://sven-lieber.org/en/2019/11/05/iswc-2019/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;I attended the ISWC conference&lt;/a&gt; a few weeks earlier.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I liked the most&lt;/strong&gt;
K-Cap is a single track conference and thus you can&amp;rsquo;t miss any talk because of another one.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I didn&amp;rsquo;t like&lt;/strong&gt;
It is interesting that K-Cap is not only about Semantic Web stuff but knowledge capturing from a general perspective also including NLP and machine learning.
However, for me some of these talks were hard to grasp and were not directly related to my research.&lt;/p&gt;
&lt;p&gt;Talks regarding &amp;ldquo;Concept embeddings&amp;rdquo; seemed to be very present, but since this topic is not too relevant for me I&amp;rsquo;ll skip those in this report 😃.
Instead in this post I will mainly focus on talks from the following topics which also seemed to be very prominent:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Wikidata&lt;/li&gt;
&lt;li&gt;Languages and approaches to map tabular data to Linked Data&lt;/li&gt;
&lt;li&gt;Knowledge Graphs of scientific practices or literature&lt;/li&gt;
&lt;/ul&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h1 id=&#34;sciknow-workshop&#34;&gt;SciKnow workshop&lt;/h1&gt;
&lt;p&gt;On the first day I attended the &lt;a href=&#34;https://sciknow.github.io/sciknow2019/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Third International Workshop on Capturing Scientific Knowledge (Sciknow 2019)&lt;/a&gt;.
&lt;a href=&#34;https://twitter.com/yolandagil&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Yolanda Gil&lt;/a&gt; gave a very interesting keynote about the capturing of hypotheses and scenarios in scientific research.
&amp;ldquo;Hypotheses are not finished after a paper is accepted&amp;rdquo; she said, the confidence value of the hypothesis changes over time.
New and/or more test subjects might be available later.
Additionally there are also other methods available to evaluate a hypothesis even though some labs might use the same or similar methods &lt;em&gt;just because they were always used&lt;/em&gt;.
&lt;strong&gt;I think we should definitely keep track of our hypotheses, methods and results in a systematic fashion: We are scientists in the end, future research from different angles might bring new insights on the current state of knowledge in our community!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Yolanda Gil giving a keynote about scientific hypotheses&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-yolanda-gil-keynote_hua4e836a21043ccf421a5bc18b2d5494a_1267762_93cf427ec055a3800ba5463da0976e8f.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-yolanda-gil-keynote_hua4e836a21043ccf421a5bc18b2d5494a_1267762_08b64a541dc54afbb666ba83039b65e9.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-yolanda-gil-keynote_hua4e836a21043ccf421a5bc18b2d5494a_1267762_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-yolanda-gil-keynote_hua4e836a21043ccf421a5bc18b2d5494a_1267762_93cf427ec055a3800ba5463da0976e8f.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://www.tib.eu/en/research-development/data-science-digital-libraries/staff/allard-oelen/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Allard Oelen&lt;/a&gt; talked about the annotation of scientific literature
using crowdsourcing to easier compare scientific contributions.
It is difficult to find crowd workers qualified for such a task, additionally it is very time consuming as a scientific paper first has to be understood.
Thus Allard and his colleagues aim for a system in which the authors annotate their papers while submitting.
Ideally this should take less than 10 minutes of the authors&amp;rsquo; time.
&lt;strong&gt;It is a very intersting approach, but authors might submit only the bare minimum information when provided with a template during paper submission.
I think a gamification approach could here be of help, i.e. showing additionall information based on the Knowledge Graph&amp;rsquo;s statistics,
such as &amp;ldquo;x% of authors also filled in this form field&amp;rdquo;.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Allard Oelen presenting work to easier compare scientific contributions due to annotation&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-allard-oelen-annotation_hu23272e4bfdd5e45e22e7fb9a09cdfb0b_1146532_b746241ba861456a6b289b84901bdfa4.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-allard-oelen-annotation_hu23272e4bfdd5e45e22e7fb9a09cdfb0b_1146532_6f0e36355e9bdcf9fb7a3082402294bf.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-allard-oelen-annotation_hu23272e4bfdd5e45e22e7fb9a09cdfb0b_1146532_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-allard-oelen-annotation_hu23272e4bfdd5e45e22e7fb9a09cdfb0b_1146532_b746241ba861456a6b289b84901bdfa4.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/ocorcho&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Oscar Corcho&lt;/a&gt; presented ongoing work from &lt;a href=&#34;http://mayor2.dia.fi.upm.es/oeg-upm/index.php/en/phdstudents/393-aiglesias/index.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Ana Iglesias Molina&lt;/a&gt;
about using spreadsheets to declaratively express Linked Data mappings which then can be translated to the most suitable language.
Starting point of the presented work is the Bio2RDF dataset which contains biomedical knowledge; &lt;em&gt;how is it generated?&lt;/em&gt;, &lt;em&gt;is it still up to date?&lt;/em&gt;.
Based on experience in working with different domain experts, spreadsheets have proven to be a good way to represent information.
&lt;strong&gt;It is an interesting approach and distantly reminds me a bit of the &lt;a href=&#34;https://spec.ottr.xyz/tabOTTR/0.3/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;tabOTTR&lt;/a&gt; representation of the ontology template language &lt;a href=&#34;https://ottr.xyz/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;OTTR&lt;/a&gt;.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Oscar Corcho presents the ongoing work of Ana Iglesias Molina about expressing RDF mapping in a tabular structure&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-oscar-corcho-choose-mapping-language_huf6c1229f47c35d5f47c6ccfad6811d6e_1365230_4ce3d537ae92d428a0b44d1683b339f6.webp 400w,
               /media/kcap-2019/2019-11-24-oscar-corcho-choose-mapping-language_huf6c1229f47c35d5f47c6ccfad6811d6e_1365230_a63c70a8bd767f5a8d4867c962adece3.webp 760w,
               /media/kcap-2019/2019-11-24-oscar-corcho-choose-mapping-language_huf6c1229f47c35d5f47c6ccfad6811d6e_1365230_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-oscar-corcho-choose-mapping-language_huf6c1229f47c35d5f47c6ccfad6811d6e_1365230_4ce3d537ae92d428a0b44d1683b339f6.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In classical Linked Open Data, Linked Data is generated and then linked.
&lt;a href=&#34;http://dgarijo.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Daniel Garijo&lt;/a&gt; presented WDPlus which follows a &lt;em&gt;link-first&lt;/em&gt; approach, different domain-specific satellites are created around the Wikidata core.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Daniel Garijo presents the WDPlus framework&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-wdplus_hudd8d531328144fc1dfefd272eea3f3c7_1178933_0bf1f4e9ba916e88af93bd01cc07adec.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-wdplus_hudd8d531328144fc1dfefd272eea3f3c7_1178933_c21da8dd5d81accc7d1b034ebb77d6be.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-wdplus_hudd8d531328144fc1dfefd272eea3f3c7_1178933_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-daniel-garijo-wdplus_hudd8d531328144fc1dfefd272eea3f3c7_1178933_0bf1f4e9ba916e88af93bd01cc07adec.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Some of these satellites might be based on tabular data.
Different languages such as R2RDF and RML exist to transform structured data to RDF.
However, especially in spreadsheets domain experts use a variety of layouts to represent the data which poses challenges.
Therefore WDPlus has the T2WML framework, using YAML to express the transformation from tabular data to Wikidata.
These mappings makes use of YAML which apparently is human-friendlier, something which I&amp;rsquo;ve seen in other approaches too.
&lt;strong&gt;I like the idea of the &amp;ldquo;link-first&amp;rdquo; approach, it reminds me of Ontology Engineering where existing classes and properties should be reused as much as possible.
In fact, also the challenges described by Daniel are similar, i.e. the identification of new properties which is currently performed manually by knowledge engineers, namespace issues and inter-satellite links.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Daniel Garijo presents the T2WML framework to map tabular data in different layouts to wikidata&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-t2wml_hu755c55016871c69c6c74bde1585e2473_1117296_d2819676b7e251319541ea4898937862.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-t2wml_hu755c55016871c69c6c74bde1585e2473_1117296_c6689a7ccd9a062a7f38b4c382d6c37b.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-daniel-garijo-t2wml_hu755c55016871c69c6c74bde1585e2473_1117296_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-daniel-garijo-t2wml_hu755c55016871c69c6c74bde1585e2473_1117296_d2819676b7e251319541ea4898937862.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;wikidata-tutorial&#34;&gt;Wikidata tutorial&lt;/h1&gt;
&lt;p&gt;In the afternoon of the first day I attended the &lt;a href=&#34;https://knowledgecaptureanddiscovery.github.io/Tutorials/T2WML-K-CAP2019/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Linking, Extending, Exploiting and Enhancing Tabular Data with Wikidata&lt;/a&gt;
tutorial given by &lt;a href=&#34;http://dgarijo.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Daniel Garijo&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/szeke&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Pedro Szekely&lt;/a&gt;.
I finally learned the basics of Wikidata and its data model which differs from e.g. the one from DBpedia.&lt;/p&gt;
&lt;p&gt;Wikidata is a collection of claims whereas each claim has related qualifiers and references.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A Wikidata page where the different components of a Wikidata statement are highlighted&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-data-model_hu46183c4c183995fb7c9610b9dde02de0_1121935_43b5b4123bd7de11204a930c0c72a467.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-data-model_hu46183c4c183995fb7c9610b9dde02de0_1121935_c3435e26c88a446f7181019c2e32decd.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-data-model_hu46183c4c183995fb7c9610b9dde02de0_1121935_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-pedro-wikidata-data-model_hu46183c4c183995fb7c9610b9dde02de0_1121935_43b5b4123bd7de11204a930c0c72a467.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Since Wikidata is using a role model it can represent information more accurrate as e.g. DBpedia.
In DBpedia Pluto is still a planet, whereas in Wikidata it is stated that it was considered a planet in a certain period but now not anymore.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;In DBpedia Pluto is still a planet, whereas Wikidata shows more detailed information based on qualifiers&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-pluto_huffbfb8c3861160717396194bacb92c5a_1096099_8cf7b10bce0dcadc2dbc8d19568f22f2.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-pluto_huffbfb8c3861160717396194bacb92c5a_1096099_09d19aaede3865ecf9a666920b92614f.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-pluto_huffbfb8c3861160717396194bacb92c5a_1096099_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-pedro-wikidata-pluto_huffbfb8c3861160717396194bacb92c5a_1096099_8cf7b10bce0dcadc2dbc8d19568f22f2.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;The following picture shows some stats about the concept hierarchy in Wikidata.
It also shows the very nice output of a commandline tool showing the hierarchy with attached stats.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Large class hierarchies in Wikidata, more than 2 million classes&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-hierarchy-stats_hubf0b8b727fdd7b8f5b972d817625ce9f_1184705_ef511140e80d3815cc1107991033814c.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-hierarchy-stats_hubf0b8b727fdd7b8f5b972d817625ce9f_1184705_bf52081ccbd1a876bafea574079030b5.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-pedro-wikidata-hierarchy-stats_hubf0b8b727fdd7b8f5b972d817625ce9f_1184705_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-pedro-wikidata-hierarchy-stats_hubf0b8b727fdd7b8f5b972d817625ce9f_1184705_ef511140e80d3815cc1107991033814c.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;However, the mapping of the Wikidata data model to RDF looks complex in the beginning, i.e. each Wikidata statement is expressed using multiple triples as reification has to be used.
Please note also the different namespaces.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A Wikidata statement mapped to RDF triples using reification&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-daniel-wikidata-rdf_hu20310bcf4b9f0a43af077486350e90da_1211314_9c10c3a7c417fc8dd163fd21aeec6b04.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-daniel-wikidata-rdf_hu20310bcf4b9f0a43af077486350e90da_1211314_8910c82c140d3fb0b75812a048aebe06.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-daniel-wikidata-rdf_hu20310bcf4b9f0a43af077486350e90da_1211314_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-daniel-wikidata-rdf_hu20310bcf4b9f0a43af077486350e90da_1211314_9c10c3a7c417fc8dd163fd21aeec6b04.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;According to GoogleTrends and statistics published by Wikidata it gained a lot of attention in the recent years.
The role-based model with qualifiers and references looks very interesting, also the fact that Wikidata is a huge crowdsourcing project.
I wonder how easy or hard it would be to adapt or develop certain Ontology Engineering activities suitable for Wikidata.&lt;/strong&gt;&lt;/p&gt;
&lt;h1 id=&#34;first-conference-day&#34;&gt;First conference day&lt;/h1&gt;
&lt;p&gt;&lt;a href=&#34;https://allenai.org/team/peterc/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Peter Clark&lt;/a&gt; from the Allen Institute of AI gave the keynote of the first conference day.
He introduced the project &lt;a href=&#34;https://allenai.org/aristo/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Aristo&lt;/a&gt; which is about reasoning and explanation.&lt;/p&gt;
&lt;p&gt;&amp;ldquo;Tables are a way to express systematic knowledge, e.g. in science&amp;rdquo; he said.
Peter talked about so called table knowledge systems which are also used in the backend of the Aristo system.
Knowledge is then extracted from these tables to execute table reasoning.
&lt;strong&gt;It is interesting to see that there are more approaches than only node-link-based Knowledge Graphs.
This table-based solution performs quite good in different challenges.
The Aristo question answering system outperforms an average 8th grader.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Peter clark explains a table knowledge system&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-peter-clark-table-knowledge_hu5b31072ea1d8913624dc843887af4f61_813526_60799cb9fab2014ccf42192f34a180a8.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-peter-clark-table-knowledge_hu5b31072ea1d8913624dc843887af4f61_813526_64bf23e66f96f212af22ef4607711efe.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-peter-clark-table-knowledge_hu5b31072ea1d8913624dc843887af4f61_813526_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-peter-clark-table-knowledge_hu5b31072ea1d8913624dc843887af4f61_813526_60799cb9fab2014ccf42192f34a180a8.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Later that day I had the chance to present our work on &lt;a href=&#34;https://zenodo.org/record/3343053#.Xd2gvvco-y4&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;MontoloStats&lt;/a&gt;
which apparently sparked some discussions among the reviewers and thus was marked a spotlight paper because it was controversial.&lt;/p&gt;
&lt;p&gt;Looking back at the first K-Cap 2001 I initially asked the question how we model knowledge nowadays were we have hundreds of ontologies and dozens of W3C recommendations available at our finger tips.
I presented our approach to create a statistical dataset about the use of different OWL axioms in ontologies extracted from &lt;a href=&#34;https://lov.linkeddata.es/dataset/lov/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;LOV&lt;/a&gt; and &lt;a href=&#34;https://bioportal.bioontology.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;BioPortal&lt;/a&gt;.
Concretely we are looking into how such statistical metadata can improve the process of ontology reuse, but beyond that, the dataset provides empirical data providing several insights and raises new questions to research.
&lt;strong&gt;I am lucky about the feedback I got. Specific tools like Protégé or visualizations like VOWL might focus too much on certain restriction types
and thus seduce users to use these more often than others.
Although marked as controversial, the audience seemed to appreciate the fact that there are now openly available empirical data about the use of axioms in ontologies.
For me such axioms are just one way to express restrictions, limiting conditions imposed by the real world or use case.
Let&amp;rsquo;s see where my research towards better modeling guidelines to express restrictions will lead me 😃.&lt;/strong&gt;&lt;/p&gt;
&lt;h1 id=&#34;town-hall-meeting&#34;&gt;Town Hall meeting&lt;/h1&gt;
&lt;p&gt;In the town hall meeting it was announced that Oscar Corcho becomes the new chair.
Furthermore the location and duration of the conference, as well as the review process and list of topics were discussed.&lt;/p&gt;
&lt;p&gt;The audience seemed to prefer larger cities over isolated areas as the traveling becomes easier.
Although it was pointed out that the sessions were quite packed with presentations, the 2-day duration is preferred over 2.5 days.
Also given the fact that the last half day of a conference some people leave earlier anyway.&lt;/p&gt;
&lt;p&gt;Whereas long papers were accompanied by a presentation, short papers had a poster.
It was pointed out that a short paper is not shorter in research and that the posters should also be visible for the duration of the entire conference.&lt;/p&gt;
&lt;p&gt;Since the review phase was in August (vacation period) it was difficult, but generally everybody was happy with the reviews.
Something like a best reviewer award was also proposed to give further incentive for good reviews.&lt;/p&gt;
&lt;p&gt;The shepherding process for certain papers was also appreciated although it gave an extra burden to reviewers and
meta reviewers as an answer letter and resubmitted paper needed to be checked again.&lt;/p&gt;
&lt;p&gt;For the future the list of topics might also undergo a small change.&lt;/p&gt;
&lt;p&gt;The venue for the next K-Cap will be &amp;hellip; announced.&lt;/p&gt;
&lt;h1 id=&#34;second-conference-day&#34;&gt;Second conference day&lt;/h1&gt;
&lt;p&gt;Due to a project meeting I unfortunately missed the keynote and most of the talks of the morning.
Even a single-track conference can&amp;rsquo;t help much in such a case 😃.&lt;/p&gt;
&lt;p&gt;In the afternoon &lt;a href=&#34;https://twitter.com/larahack&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Cristina-Iulia Bucur&lt;/a&gt; talked about the review process in science which didn&amp;rsquo;t really change much the past few hundred years.
Currently she and her colleagues are looking into ways to improve the reviewing process by using Semantic Web technologies.
Reviews and concerned text passages are expressed as Linked Data, additionally reviews can be further annotated to e.g. say if it concerns grammar or content.
A user study was performed to evaluate the proposed fine grained review data model.
&lt;strong&gt;Looking at how the review process works, I wonder why nobody ever changed it.
I think even in a pre-semantic-web era other tools could have been used to improve the review process.
In the here proposed solution I especially like the UI in which the review is directly attached to the text, so researchers can review while reading.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Cristina-Iulia Bucur is researching different interfaces for reviewers to make more detailed comments&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-cristina-review-comments_huade280314e5237f77ca0279974c4eb02_956687_5e3d0e66de6d0994261bb3053785e035.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-cristina-review-comments_huade280314e5237f77ca0279974c4eb02_956687_681b5abce135fb857c80f20fe7595b04.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-cristina-review-comments_huade280314e5237f77ca0279974c4eb02_956687_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-cristina-review-comments_huade280314e5237f77ca0279974c4eb02_956687_5e3d0e66de6d0994261bb3053785e035.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://github.com/binh-vu&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Binh Vu&lt;/a&gt; talked about the language D-REPR to map tabular data with different layout to RDF.
Based on openly available tabular data extracted from data.gov it was verified that the language meets stated requirements.
Additionally the performance was evaluated by comparing the execution of the Rust-based implementation of D-REPR against two other solutions on data in the nested relational model layout.
&lt;strong&gt;There are many languages and tools to transform tabular or any structured data to RDF.
A discussion after the talk revealed that instead of creating new languages and tools the existing recommendations should be updated, taking current limitations into account.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Binh Vu explains the evaluation of D-REPR used to transform differently layouted tables to RDF&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-binh-table-layout_hu2f6b06c283b8f3a483f9719fd23dd616_813873_de8eb77cf4febce1243dc7ba0a8197f1.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-binh-table-layout_hu2f6b06c283b8f3a483f9719fd23dd616_813873_5afaba449b13b0e51f2447653a927583.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-binh-table-layout_hu2f6b06c283b8f3a483f9719fd23dd616_813873_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-binh-table-layout_hu2f6b06c283b8f3a483f9719fd23dd616_813873_de8eb77cf4febce1243dc7ba0a8197f1.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;I&amp;rsquo;m currently looking into visualizations for RDF data, thus I found the talk of &lt;a href=&#34;http://sda.cs.uni-bonn.de/people/vitalis-wiens/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Vitalis Wiens&lt;/a&gt; also very interesting.
He presented &lt;a href=&#34;https://gizmo-vis.github.io/gizmo/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;GizMO&lt;/a&gt;, a methodology and tool to visualize ontologies.
It is based on an ontology which addresses the graphical appearance of OWL constructs in one layer, and visual properties for conceptual elements in another layer.
&lt;strong&gt;Great work! I am not aware of any ontology to describe the visual appearance of graph based models.
Additionally nested visualizations like UML are also mentioned in the paper and can be generated in the demo.
I&amp;rsquo;m pretty sure this work will be relevant for a lot of visualizations out there and can also serve as a common format to interchange visual information.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Vitalis Wiens explains the data model of GizMO&#34; srcset=&#34;
               /media/kcap-2019/2019-11-24-kcap-vitalis-gizmo_hud1cb86aa437beed06e065ca728a49573_1051279_04587360628c9f63ba54394af3767a96.webp 400w,
               /media/kcap-2019/2019-11-24-kcap-vitalis-gizmo_hud1cb86aa437beed06e065ca728a49573_1051279_f064feaeada1a76d673b1b066c6dbedb.webp 760w,
               /media/kcap-2019/2019-11-24-kcap-vitalis-gizmo_hud1cb86aa437beed06e065ca728a49573_1051279_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/kcap-2019/2019-11-24-kcap-vitalis-gizmo_hud1cb86aa437beed06e065ca728a49573_1051279_04587360628c9f63ba54394af3767a96.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;closing-event&#34;&gt;Closing event&lt;/h1&gt;
&lt;p&gt;Within the closing event, Ling Cai was awarded the best paper award for the paper &lt;em&gt;TransGCN:Coupling Transformation Assumptions with Graph Convolutional Networks for Link Prediction&lt;/em&gt;.
&lt;strong&gt;Congratulations!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;This was K-Cap 2019! I enjoyed the many conversations I had. Additionally I also had my first main track presentation.
I will definitely have a look into Wikidata and will also follow up on some of the presented papers.
I&amp;rsquo;m looking forward to submit something to next year&amp;rsquo;s EKAW and then of course to K-Cap again in two years!&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>ISWC 2019</title>
      <link>https://sven-lieber.org/en/2019/11/05/iswc-2019/</link>
      <pubDate>Tue, 05 Nov 2019 08:41:37 +0100</pubDate>
      <guid>https://sven-lieber.org/en/2019/11/05/iswc-2019/</guid>
      <description>&lt;p&gt;In the last week of October 2019 I attended the International Semantic Web Conference (ISWC) in Auckland, New Zealand.
I had the honor to present my PhD plan at the ISWC Doctoral Consortium.
In this blog post I talk a bit about the conference and focus on a few presentations I found particular interesting or relevant for my PhD regarding guidelines for the design and use of ontologies.&lt;/p&gt;
&lt;p&gt;First things first:
I will provide a bit of context and say what I liked/disliked.&lt;/p&gt;
&lt;p&gt;The International Semantic Web Conference (ISWC) is &lt;em&gt;THE&lt;/em&gt; conference of research regarding the Semantic Web also covering Linked Data, Knowledge Graphs and AI.
This year it was held in Auckland New Zealand which on the one hand is an amazing venue, but on the other hand resulted in super long flights and visa issues for a lot of attendees.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I liked the most&lt;/strong&gt;
The conference also had several sessions regarding ontologies and ontology design which is relevant for my PhD.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I didn&amp;rsquo;t like&lt;/strong&gt;
I missed a few really good workshops because I had to attend the Doctoral Consortium a full day. Fully attending it is a good thing, just a pity it overlaps with other workshops.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h1 id=&#34;ottr-tutorial&#34;&gt;OTTR tutorial&lt;/h1&gt;
&lt;p&gt;For the first workshop day I&amp;rsquo;ve chosen to attend the tutorial on &lt;a href=&#34;http://ottr.xyz/event/2019-10-267-iswc/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;scalable construction of sustainable knowledge bases&lt;/a&gt; which was all about the Reasonaable Ontology Templates (OTTR) framework.
It was given by &lt;a href=&#34;http://folk.uio.no/martige/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Martin G. Skjæveland&lt;/a&gt;, &lt;a href=&#34;https://www.mn.uio.no/ifi/personer/adm/leifhka/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Leif Harald Karlsen&lt;/a&gt;  and &lt;a href=&#34;https://www.mn.uio.no/ifi/english/people/aca/danielup/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Daniel Lupp&lt;/a&gt; from the University of Oslo,
and &lt;a href=&#34;https://www.uwa.edu.au/Profile/Melinda-Hodkiewicz&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Melinda Hodkiewicz&lt;/a&gt; from the University of Western Australia.&lt;/p&gt;
&lt;p&gt;In a nutshell OTTR is a language and a framework to describe knowledge on a conceptual level, whereas it abstracts from the actual knowledge representation in e.g. OWL.
Thus ontology engineers can communicate with domain experts in their language and talk about for example a Person without needing to know which concrete axioms are used to represent a person in OWL.
Templates contain parameters representing the variable part of a template which then gets extended to RDF together with its static part; working with templates feels like working with a constructor in languages such as Java.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;I was particular interested in this tutorial as I see OTTR as the missing building block to generate different ontology representations based on the same conceptual model which is relevant for my PhD.&lt;/strong&gt;
The tutorial introduced the basics, we could get our hands dirty on some examples and then two real-life use cases were introduced.&lt;/p&gt;
&lt;p&gt;Melinda Hodkiewicz stressed the need for ontologies in the Engineering sector, which she did as well in each session I have seen her during the main conference 😃.
Which is nice, we need more enthusiastic people such as her to promote our work and highlight its concrete added value for businesses or the end user.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Usage of menti.com to gather questions the audience would like to get answered by the tutorial&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-ottr-tutorial-feedback_hu6a225bdc1feb4ba013a91f30bc0689a1_603140_a44c1e40e674b9d351c8b4c391f5fa91.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-ottr-tutorial-feedback_hu6a225bdc1feb4ba013a91f30bc0689a1_603140_f3d68f1b7ef63f68e01dac07f4ac66c0.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-ottr-tutorial-feedback_hu6a225bdc1feb4ba013a91f30bc0689a1_603140_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-ottr-tutorial-feedback_hu6a225bdc1feb4ba013a91f30bc0689a1_603140_a44c1e40e674b9d351c8b4c391f5fa91.webp&#34;
               width=&#34;760&#34;
               height=&#34;413&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Besides the content I really liked the usage of &lt;a href=&#34;https://menti.com&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;https://menti.com&lt;/a&gt; to interactively get feedback from the audience such as demographics or incentives to join the tutorial.
And the end of the tutorial the organizers went back to the initially asked questions and incentives of the audience and it was discussed whether each point was addressed.
Thumbs up, please more of that in the future!&lt;/p&gt;
&lt;h1 id=&#34;doctoral-consortium&#34;&gt;Doctoral Consortium&lt;/h1&gt;
&lt;p&gt;A lot of workshops took also place at the second day but I didn&amp;rsquo;t attend as I has the great pleasure of participating in the ISWC Doctoral Consortium all day where I could present the plan for my PhD.&lt;/p&gt;
&lt;p&gt;The doctoral consortium started with a great keynote given by &lt;a href=&#34;https://twitter.com/vtamma&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Valentina Tamma&lt;/a&gt;.
She started by defining what a PhD means according to its definition, i.e. doctor of philosophiae in Greek &amp;ldquo;love of knowledge&amp;rdquo;, and continued with further details regarding the scientific process.
She recommended the book &lt;a href=&#34;https://en.wikipedia.org/wiki/Science_in_Action_%28book%29&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Science in Action from Bruno Latour (1987)&lt;/a&gt; and stressed that &lt;strong&gt;we should read a lot and talk about our research as much as possible, first to the broader community and then in lay terms also to the general public.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Different dimensions of scientific investigations, a slide from the presentation of Valentina Tamma&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-dc-intro_hu19463e6930a26ca5cf9bf1d6746a943d_1190477_0f42ce96293625d9b00b8093b9c9b787.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-dc-intro_hu19463e6930a26ca5cf9bf1d6746a943d_1190477_b9532c808334194f98def7a465e330cb.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-dc-intro_hu19463e6930a26ca5cf9bf1d6746a943d_1190477_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-dc-intro_hu19463e6930a26ca5cf9bf1d6746a943d_1190477_0f42ce96293625d9b00b8093b9c9b787.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Each presenting PhD student got upfront a mentor assigned who gave feedback to the presentation, but of course everyone in the audience was invited to ask questions and provide feedback.
The presentations varied in quality and depth, but I think no one was able to present in the given time of 12 minutes (me included).
It actually is a challenging task to present your whole PhD research within 12 minutes, especially if you are in an early-stage and it is just a plan.&lt;/p&gt;
&lt;p&gt;The presentations were supposed to be accompanied by a poster presented in poster sessions in the coffee breaks. However, due to the fact that no one finished in time with the presentation, the poster session was moved to the general poster session of the main conference on Monday.
Thanks to the chairs and local organizers to find some spare walls and arrange that we could present next to the main poster session.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Me and my ISWC Doctoral Consortium poster at the general poster session of ISWC 2019&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_ca54ba1c770e0164c0875be2c4ae89ff.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_726327b97a9c025dc3b6c69901e4f2f2.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-poster-sven_hu8ce16bb784ea6d31c3139b8f49646045_1055553_ca54ba1c770e0164c0875be2c4ae89ff.webp&#34;
               width=&#34;570&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;A panel session rounded up the doctoral consortium.
Several senior researchers discussed questions regarding the role of the PhD in industry, I will just briefly mention the questions and the main points:&lt;/p&gt;
&lt;ol&gt;
&lt;li&gt;Which is your opinion with respect to this migration from academia to industry?”
&lt;ul&gt;
&lt;li&gt;Higher salaries and more stability in industry&lt;/li&gt;
&lt;li&gt;Nowadays more opportunities in industry and a more research-influenced style of work&lt;/li&gt;
&lt;li&gt;On the negative side: whole research areas get confiscated by industry, e.g. information retrieval and NLP&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;From your experience, which is the actual need of industries in terms of innovation? Are they really believing in it or it is just a trend that they want to follow?
&lt;ul&gt;
&lt;li&gt;The industry wants simple things which work&lt;/li&gt;
&lt;li&gt;Companies need to stay ahead of competition&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;May this new age of AI affect the programs of your PhD students?
&lt;ul&gt;
&lt;li&gt;Putting PhDs on irrelevant topics makes also everyone unhappy&lt;/li&gt;
&lt;li&gt;We also benefit from it, e.g. knowledge embeddings, but one has to be careful because building a system by themselves is not research, how does it help our community?&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;li&gt;Do you think that a strong collaboration with companies would have a detrimental effect on the quality of research?
&lt;ul&gt;
&lt;li&gt;A lot of tools we use nowadays are built by industries, we wouldn&amp;rsquo;t have build them for ourselves&lt;/li&gt;
&lt;li&gt;Industry brings a lot of insights&lt;/li&gt;
&lt;li&gt;Of course also the risk in credibility if too much research is directly funded by industry&lt;/li&gt;
&lt;/ul&gt;
&lt;/li&gt;
&lt;/ol&gt;
&lt;h1 id=&#34;monday&#34;&gt;Monday&lt;/h1&gt;
&lt;p&gt;At the first official conference day I participated in the opening session, the keynote by &lt;a href=&#34;https://nz.linkedin.com/in/dougalwatt&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Dougal Watt&lt;/a&gt;, a session with industry talks and two sessions regarding data quality.&lt;/p&gt;
&lt;h2 id=&#34;opening-session&#34;&gt;Opening session&lt;/h2&gt;
&lt;p&gt;&lt;em&gt;Within the opening session&lt;/em&gt; it was stated that from 194 valid research track submissions 42 papers were selected and thus the acceptance rate of the research track was 21.6%.
According to the presented statistics of submissions per topic in the research track (topics with more than 10 submissions) almost 80 submissions concerned database, information retrieval, information extraction and NLP.
Whereas Ontology Engineering and ontology design patterns for the web are the topics with the least submissions in this more-than-10 category.
It was a pity to see that the privacy topic was ranked in the statistics of topics with less than 5 submissions.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The number of submissions per research track topic with more than 10 submissions&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-opening-stats_hu81b707aa0ae7fcc812a0d9029b882706_1370280_50adbdac0065564318b83bf07940dc65.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-opening-stats_hu81b707aa0ae7fcc812a0d9029b882706_1370280_9ee582f7ace23747ecb984aca00ecde2.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-opening-stats_hu81b707aa0ae7fcc812a0d9029b882706_1370280_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-opening-stats_hu81b707aa0ae7fcc812a0d9029b882706_1370280_50adbdac0065564318b83bf07940dc65.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Congratulations to &lt;a href=&#34;https://twitter.com/olafhartig&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Olaf Hartig&lt;/a&gt;, &lt;a href=&#34;https://www.uni-mannheim.de/dws/people/professors/prof-dr-christian-bizer/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Christian Bizer&lt;/a&gt;, &lt;a href=&#34;https://scholar.google.com/citations?user=yVGyEzIAAAAJ&amp;amp;hl=de&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Johann-Christoph Freytag&lt;/a&gt; to win the SWSA Ten-year award with their work &amp;ldquo;Executing SPARQL Queries over the Web of Linked Data&amp;rdquo;.
Also congratulations to &lt;a href=&#34;https://juliusv.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Julius Volz&lt;/a&gt;, Christian Bizer, &lt;a href=&#34;https://twitter.com/gaedke&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Martin Gaedke&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/gkob&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Georgi Kobilarov&lt;/a&gt; to also win the SWSA Ten-year award with their work &amp;ldquo;Discovering and Maintaining Links on the Web of Data&amp;rdquo;.&lt;/p&gt;
&lt;h2 id=&#34;keynote-by-dougal-watt&#34;&gt;Keynote by Dougal Watt&lt;/h2&gt;
&lt;p&gt;&lt;em&gt;Monday&amp;rsquo;s keynote&lt;/em&gt; was given by Dougal Watt and was titled &amp;ldquo;Semantics: the business technology disruptor of the future&amp;rdquo;.
He gave a small tour through the different eras of information processing: from pre-literate and literate via the printed word to modern computing, the web and context nowadays.
Furthermore he talked about application centric thinking and about the different problems of relational databases, enterprise resource planning, enterprise data modelling, SOA and data warehousing.
Summarized these problems mainly focus on limited semantics, massively complex models, interoperability and missing schema agreement leading to a 28 minute discussion about definitions within a 30 minutes meeting.
Of course semantic technologies and ontologies offer a possible solution for these problems.
&lt;strong&gt;But according to my opinion some of these problems still remain and are just shifted to how ontologies are modeled and how they are used, i.e. discussions of definitions and complex models.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The problems of application centric thinking in relational databases, enterprise resource planning, enterprise data modeling, SOA and data warehousing presented by Dougal Watt&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-keynote-dougal-application-centric_hu752df7cab006d09daf7f92153f70a2f8_1308645_e157e3d428563431d5abc933537048cf.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-keynote-dougal-application-centric_hu752df7cab006d09daf7f92153f70a2f8_1308645_f3e9507a7133e239b59230df1865e363.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-keynote-dougal-application-centric_hu752df7cab006d09daf7f92153f70a2f8_1308645_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-keynote-dougal-application-centric_hu752df7cab006d09daf7f92153f70a2f8_1308645_e157e3d428563431d5abc933537048cf.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Within the QA the topic data shapes and SHACL were brought up. Dougal said that SHACL shapes are fantastic things, as they can hide the complexity of ontologies which can then be used by businesses.
Personally I find this a very interesting point which is also relevant to my PhD:
&lt;strong&gt;Shapes become important components when working with ontologies and thus need also to be built systematically and/or considered when designing ontologies.&lt;/strong&gt;&lt;/p&gt;
&lt;h2 id=&#34;presentations&#34;&gt;Presentations&lt;/h2&gt;
&lt;p&gt;&lt;a href=&#34;https://www.mn.uio.no/ifi/english/people/aca/evgenykh/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Evgeny Kharlamov&lt;/a&gt; presented the work of colleagues regarding a knowledge-base for material science within the industry track.
Working in material science involves the creation of new materials for which knowledge of thousands of existing papers and patents need to be searched.
Queries such as which material X reacts with material Y and has its melting point below Z need to be asked.
An ontology was designed and a knowledge graph was built to facilitate the task of this search scenario.
Furthermore they are working on an end-to-end system including not only the ontology but also search interfaces for end-users.
I was interested in how the ontology was designed to deal with that task: apparently a lot of subclass axioms were used.
I like their use case in which they have to provide a user-friendly system powered by an ontology.
&lt;strong&gt;As Dougal pointed out in his keynote: enterprise data modeling tend to create massively complex models. I would say that this enterprise modeling mindset still exists in the mind of some ontology engineers, a use case such as presented here shows that ontologies also have to be used by systems which also comes with extra requirements regarding their use.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;Within one of the &lt;em&gt;data quality sessions&lt;/em&gt;, I listened to &lt;a href=&#34;https://twitter.com/albertmeronyo&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Albert Meroño-Peñuela&lt;/a&gt; who presented the work of him and his colleagues about RDF-lists.
MIDI applications (Musical Instrument Digital Interface) express songs as long lists of musical chords.
In the context of this, the authors surveyed common list modeling practices taking e.g. different ontology design patterns for lists into account.
The authors reported coherent results among triple stores when evaluating the triple stores&amp;rsquo; performance when processing the different ways in which lists are currently represented.
The rdf:List is the only expression which produces &amp;ldquo;closed&amp;rdquo; lists, but it is also the one with the poorest performance.
&lt;strong&gt;It is interesting to see that certain design decisions have an impact on performance. It actually shows that guidelines of how to implement an ontology for a certain use case are needed to provide an optimal solution, i.e. in a use case in which performance is important the usage of rdf:List might not be optimal, whereas in a use case demanding more strict semantics the closeness of lists might be required.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Conclusions about lists in the Semantic Web&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-rdf-lists_hud1c5caa700994de9e67861bb3eecf41a_1198619_0c1edbbe53d7573dd916422b1ac99b13.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-rdf-lists_hud1c5caa700994de9e67861bb3eecf41a_1198619_e10cf2ba4827d6810d1cc7d42a532b1d.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-rdf-lists_hud1c5caa700994de9e67861bb3eecf41a_1198619_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-rdf-lists_hud1c5caa700994de9e67861bb3eecf41a_1198619_0c1edbbe53d7573dd916422b1ac99b13.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;tuesday&#34;&gt;Tuesday&lt;/h1&gt;
&lt;p&gt;For Tuesday I will discuss the keynote, panel session, PhD lunch, town hall meeting, gala dinner but also actual research sessions :-)&lt;/p&gt;
&lt;h2 id=&#34;keynote-by-jérôme-euzenat&#34;&gt;Keynote by Jérôme Euzenat&lt;/h2&gt;
&lt;p&gt;On Tuesday, &lt;a href=&#34;https://moex.inria.fr/~euzenat/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Jérôme Euzenat&lt;/a&gt; presented the keynote &lt;em&gt;For Knowledge&lt;/em&gt;.
He quotes that data without knowledge is like a map without a legend and stresses that explicit knowledge is important.
Also, &lt;em&gt;a great idea remains a great idea&lt;/em&gt;.
Sometimes the time for an invention has not yet come and reinventing the wheel sometimes makes sense respectively invent different variants of a wheel.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Jerome Euzenat quotes that data without knowledge is like a map without a legend&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-keynote-Jerome-data-without-knowledge_hu802ddd28e8f4227b4268f09c96be912d_849361_6f379f1e166b02f91c73d694fb77a96f.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-keynote-Jerome-data-without-knowledge_hu802ddd28e8f4227b4268f09c96be912d_849361_05333aa55b576642831c88694c628d8b.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-keynote-Jerome-data-without-knowledge_hu802ddd28e8f4227b4268f09c96be912d_849361_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-keynote-Jerome-data-without-knowledge_hu802ddd28e8f4227b4268f09c96be912d_849361_6f379f1e166b02f91c73d694fb77a96f.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Another interesting idea he proposes concerns eScience.
Nowadays scientific experiments can be described semantically: why not submitting a detailed experiment description to conferences/journals to receive reviews purely on the scientific process and not the outcome?!&lt;/p&gt;
&lt;h2 id=&#34;panel&#34;&gt;Panel&lt;/h2&gt;
&lt;p&gt;After the keynote a panel session took place.
It was titled &amp;ldquo;How much semantics goes a long way&amp;rdquo;, emphasizing on the &lt;em&gt;how much&lt;/em&gt; with respect to the famous &amp;ldquo;a little semantics goes a long way&amp;rdquo; from &lt;a href=&#34;http://www.cs.rpi.edu/~hendler/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;James Hendler&lt;/a&gt;.
Valentina Tamma, Jérôme Euzenat and James Hendler were the panel and the room was packed with senior SemWeb researchers.
The audience lively participated in this panel discussion.
I will briefly list &lt;em&gt;some&lt;/em&gt; of the main comments.&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;The past 20 years demonstrated that different ontologies in the same domain can exist, people choose what they want and one ontology possibly becomes the lingua franca of that domain&lt;/li&gt;
&lt;li&gt;There are obstacles; if technologies and languages are used which are too complex for people to use they won’t use it. That’s why expressivity might be reduced&lt;/li&gt;
&lt;li&gt;A way to achieve consensus is maybe a more expressive language to express nuance and isolate which statement you agree with and to which not. So you might end up to agree&lt;/li&gt;
&lt;li&gt;The problem is not about expressivity per se, but about the incentives to put the expressivity. E.g. in the research of genes, researchers put the expressivity in. We should ask, how to incentive the people to put the expressivity&lt;/li&gt;
&lt;li&gt;How much semantics is perceived wrongly, choose either to define less or richer expressivity, but it depends! For certain tasks we don’t need much expressivity. Focus should be to make these different types to coexist with the same data&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id=&#34;phd-lunch&#34;&gt;PhD lunch&lt;/h2&gt;
&lt;p&gt;On Tuesday also the &lt;em&gt;PhD lunch&lt;/em&gt; took place.
This program point belongs to the Doctoral Consortium; PhD students and their assigned mentors should have a more informal environment to talk.
I liked the setup, I was on a table with a few other PhD students and a few mentors.
We could ask questions and a lively discussion was sparked.
There were certainly very helpful tips, however given the time constraint we couldn&amp;rsquo;t reach a certain depth in the discussion so everything stayed more high-level.&lt;/p&gt;
&lt;h2 id=&#34;town-hall-meeting&#34;&gt;Town hall meeting&lt;/h2&gt;
&lt;p&gt;The ISWC conference provides the &lt;em&gt;town hall meeting&lt;/em&gt; as general session in which all kind of feedback can be provided.&lt;/p&gt;
&lt;p&gt;One of the first points brought up was the review process.
It was the first time that the research track was double-blind.
Some people liked it and some not; there are certainly pros and cons but it was mentioned that better guidelines in how to review should be provided.&lt;/p&gt;
&lt;p&gt;Another point of discussion was the location of the conference and thus the visa problematic a lot of attendees faced.
The visa regulations of New Zealand were just recently changed so it couldn&amp;rsquo;t be taken into account back then when the conference venue was chosen.
Unfortunately the university also had zero power in influencing the government&amp;rsquo;s visa decisions.
Comments from the audience were among others that we should be more inclusive as a community and choose venues more wisely.&lt;/p&gt;
&lt;h2 id=&#34;session-on-ontology-design&#34;&gt;Session on ontology design&lt;/h2&gt;
&lt;p&gt;Two papers in this session gathered my attention, a paper regarding the detection of ontology design patterns in biomedical ontologies,
and a spotlight paper about the use of OWL and WebProtégé at Pinterest.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://www.research.manchester.ac.uk/portal/christian.kindermann.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Christian Kindermann&lt;/a&gt; presented how the influence of ontology design patterns in the biomedical domain can be measured.
The presented results indicate that there does not seem to be much use of ontology design patterns in the analyzed ontologies.
However, regardless of the results this is very interesting research as it provides a baseline which allows further investigation.
&lt;strong&gt;One introduced way to detect ontology design patterns is by checking first if a specific axiom used by a pattern is present in an ontology.
This indirect way to conclude if an ontology design patterns is present in an ontology can be facilitated by our &lt;a href=&#34;https://w3id.org/montolo&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;MontoloStats&lt;/a&gt; dataset/webservice on which we recently published &lt;a href=&#34;https://sven-lieber.org/en/publication/montolo-stats-ontology-modeling-statistics&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;a paper&lt;/a&gt;.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Christian Kindermann shows how many analyzed ontologies contained ontology design patterns grouped by how the patterns were detected&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-design-pattern-detection_hu48987cc9da07741e672127cc1047b4a7_933010_68c8537b0ab426c24a852bf1e4d26474.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-design-pattern-detection_hu48987cc9da07741e672127cc1047b4a7_933010_00dfb056a805467c18e5bf76c7d2b5f4.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-design-pattern-detection_hu48987cc9da07741e672127cc1047b4a7_933010_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-design-pattern-detection_hu48987cc9da07741e672127cc1047b4a7_933010_68c8537b0ab426c24a852bf1e4d26474.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://www.rsgoncalves.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Rafael S Gonçalves&lt;/a&gt; presented how an OWL ontology and the use of WebProtégé drastically shortened the development cycle of the Pinterest taxonomy.
The Pinterest taxonomy contained around 6,000 &amp;ldquo;interests&amp;rdquo; users can have which were arranged in 3 layers.
This taxonomy was increased to 11,000 interests represented as OWL classes organized in 12 layers.
It took less than 2 months to build the end-to-end system the Pinterest employees without any ontology experience can use.
Furthermore it took less than 1 month to build the final ontology.
&lt;strong&gt;I would say that OWL is often perceived as very complex, but this presentation shows that with the use of appropriate tooling such as WebProtégé, which offers collaborative features and complete provenance, even non Semantic Web experts can handle an ontology!&lt;/strong&gt;
A question asked by the audience further gathered my attention: why did you use OWL and not just SKOS?
Apparently there are future needs which can be addressed by OWL, especially the implementation of existential restrictions.
&lt;strong&gt;I find this particularly interesting as it shows that there exist specific needs for certain restriction types which is relevant for my PhD and the definition of guidelines in how to model restrictions.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The WebProtege interface showing parts of the Pinterest taxonomy&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-pinterest-ontology_hub1f5040c32543ae0e31cf0bd7f0e3ae9_1021067_3e88c1514fb953c272c803fa48bb9639.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-pinterest-ontology_hub1f5040c32543ae0e31cf0bd7f0e3ae9_1021067_6b0351616f2f6f1a369b7696321fda33.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-pinterest-ontology_hub1f5040c32543ae0e31cf0bd7f0e3ae9_1021067_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-pinterest-ontology_hub1f5040c32543ae0e31cf0bd7f0e3ae9_1021067_3e88c1514fb953c272c803fa48bb9639.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h2 id=&#34;session-on-domain-ontologies&#34;&gt;Session on domain ontologies&lt;/h2&gt;
&lt;p&gt;Within this session I liked the talk from &lt;a href=&#34;https://twitter.com/karlhammar&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Karl Hammar&lt;/a&gt; about a Real Estate ontology and the talk about the W3C recommended SSN ontology.&lt;/p&gt;
&lt;p&gt;Real estate data are not just about prices, there are also needs regarding energy efficiency, maintenance and sustainability.
Karl Hammer presented a real estate ontology which was built following an adapted version of the eXtreme Design methodology.
The ontology consists of a core and different modules and it deliberately renounces the reuse of foundational ontologies.
&lt;strong&gt;I like that he emphasized that an ontology is useful but not enough! There must be also a way to interact with it, i.e. the ontology has to be used.&lt;/strong&gt;
Short after the conference, Karl released a first version of &lt;a href=&#34;https://github.com/hammar/OWL2OAS&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;a tool&lt;/a&gt; to create a OpenAPI specification from an OWL ontology.
Cool stuff 👍&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Karl Hammer presents the tradeoffs and challenges of the presented real estate ontology&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-real-estate-ontology_huf698025bf02fc38878c9bed8d9f182cd_1022199_5d0ef11c359ff5e37641f28b246b5419.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-real-estate-ontology_huf698025bf02fc38878c9bed8d9f182cd_1022199_f19e285634a8a077b0105e4362304602.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-real-estate-ontology_huf698025bf02fc38878c9bed8d9f182cd_1022199_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-real-estate-ontology_huf698025bf02fc38878c9bed8d9f182cd_1022199_5d0ef11c359ff5e37641f28b246b5419.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;The SSN ontology is a W3C recommendation to model sensor data.
&lt;a href=&#34;https://researchers.anu.edu.au/researchers/taylor-kl&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Kerry Taylor&lt;/a&gt; presented the work around the SSN ontology.
Interestingly, not only the SSN ontology was built but several best practices were published under the umbrella of W3C.
I have read a journal paper related to this talk, it discusses design decisions to make the initially defined data model easier.
This was achieved by modularizing the ontology into a &amp;ldquo;lightweight&amp;rdquo; core and outsource alignments to other ontologies into specific modules.
The core ontology mainly defines concepts and no axioms, even to determine domain and range schema.org&amp;rsquo;s domainIncludes and rangeIncludes are used.
&lt;strong&gt;In my opinion the presented solution solves a lot of problems and demonstrates how a trade-off between strict semantics regarding interoperability to other standards and lightweight description of concepts can be achieved!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The main objectives for a new version of the SSN ontology&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-ssn-ontology_hu6596811e0baf3795041491147cc7790f_957703_202ef6ea25855e9ea5810b4c76e4dee6.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-ssn-ontology_hu6596811e0baf3795041491147cc7790f_957703_64beb8f0d58fadbce466d6cded543f18.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-ssn-ontology_hu6596811e0baf3795041491147cc7790f_957703_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-ssn-ontology_hu6596811e0baf3795041491147cc7790f_957703_202ef6ea25855e9ea5810b4c76e4dee6.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h2 id=&#34;gala-dinner&#34;&gt;Gala dinner&lt;/h2&gt;
&lt;p&gt;Besides all the research the conference of course also provided a lot of room for social gathering :-)
On Tuesday&amp;rsquo;s agenda was the gala dinner. A Māori ensemble performed several songs and traditional dances to welcome the ISWC participants from all over the world to New Zealand!
The Māori are the indigenous people of New Zealand which inhabited the islands roughly 700 years ago.
This is also interesting as it means that New Zealand was the last place on earth on which the Homo Sapiens settled down, just a few hundred years ago!
After that we enjoyed our food to which also the ice-cream kiwi from the blog posts header image belongs.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A maori ensemble performs dances to welcome the ISWC attendees from around the world to New Zealand&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-gala-dinner_hu5f64ea111531b9633c53dd310f20c72e_1286773_a1ef3c00e7cd06165845a01ea4d166d4.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-gala-dinner_hu5f64ea111531b9633c53dd310f20c72e_1286773_65d3a2f4b7e527f5f550a00589b343c1.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-gala-dinner_hu5f64ea111531b9633c53dd310f20c72e_1286773_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-gala-dinner_hu5f64ea111531b9633c53dd310f20c72e_1286773_a1ef3c00e7cd06165845a01ea4d166d4.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;wednesday&#34;&gt;Wednesday&lt;/h1&gt;
&lt;p&gt;The last day was shorter, my highlights are the keynote, a presentation about ontology ranking within a search scenario and the closing ceremony.&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://en.wikipedia.org/wiki/Melanie_Johnston-Hollitt&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Melanie Johnston-Hollitt&lt;/a&gt; gave the keynote &lt;em&gt;Extracting Knowledge from the Data Deluge to Reveal the Mysteries of the Universe&lt;/em&gt;.
This keynote was very interesting as it revealed a lot of really cool scientific applications which could profit from explicit knowledge.
I can only recommend to listen to the keynote which is available on &lt;a href=&#34;https://www.youtube.com/watch?v=GXEzJ2ko3oA&amp;amp;list=PLx7cxNxrqEEqRdWux0Xqzq9xaMVTcrbxV&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;YouTube&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;Short before Lunch I listened to some talks regarding ontologies for recipes and food. This seems very interesting and I&amp;rsquo;m happy some efforts regarding this direction already exist.&lt;/p&gt;
&lt;p&gt;In the &lt;em&gt;last talk&lt;/em&gt; I&amp;rsquo;ve heard at ISWC 2019, &lt;a href=&#34;https://wwwen.uni.lu/snt/people/niklas_kolbe&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Niklas Kolbe&lt;/a&gt; talked about the ranking of ontologies regarding relevancy when searching for concepts.
He pointed out that people have a certain trust into the first few search results and hence a ranking plays a crucial role.
In the already mentioned &lt;a href=&#34;https://sven-lieber.org/en/publication/montolo-stats-ontology-modeling-statistics/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;recently published paper&lt;/a&gt; we argue that for an ontology reuse scenario the search for ontologies containing certain axioms might be relevant.
&lt;strong&gt;It was interesting to hear that something like our solution could be integrated into the presented solution based on qualitative features.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Niklas Kolbe concludes on their work on predicting an ontology&amp;amp;rsquo;s relevance for ranking&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-ontology-ranking_huf0e82845438defbb70752d8b3c9cb00f_1205902_95d34b88c513487c769389d3944f61a9.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-ontology-ranking_huf0e82845438defbb70752d8b3c9cb00f_1205902_e0c88b12585d899a0ec372545d6c13a2.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-ontology-ranking_huf0e82845438defbb70752d8b3c9cb00f_1205902_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-ontology-ranking_huf0e82845438defbb70752d8b3c9cb00f_1205902_95d34b88c513487c769389d3944f61a9.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In &lt;em&gt;the closing event&lt;/em&gt; best papers in several categories were awarded, I will not list all the winners, just a general &amp;ldquo;congratulations&amp;rdquo;! 😃&lt;/p&gt;
&lt;p&gt;The next ISWC conference will take place in Athens, Greece.
In 2 years ISWC will take place in Albany, USA. The audience seemed not that enthusiastic when this venue was announced, .. maybe because of the already mentioned issues with visas for certain countries such as the US.&lt;/p&gt;
&lt;p&gt;That was the ISWC 2019 conference from my point of view, I hope you enjoyed reading it!
I got a lot of feedback for my PhD and also saw the relevancy for it.&lt;/p&gt;
&lt;p&gt;I was writing this blog post while traveling around in New Zealand,
I will now enjoy the remaining time in New Zealand, see you at the next conference!&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Goodbye ISWC, I&amp;amp;rsquo;m on holiday now&#34; srcset=&#34;
               /media/iswc-2019/2019-11-05-iswc-goodbye_hu64dfa71369f1ea54a6fe97db7aeaa3db_94496_da280338965f8420b4ea76ed0bbfee86.webp 400w,
               /media/iswc-2019/2019-11-05-iswc-goodbye_hu64dfa71369f1ea54a6fe97db7aeaa3db_94496_f818b1ea7aa1c3e1ca15c5190cd17853.webp 760w,
               /media/iswc-2019/2019-11-05-iswc-goodbye_hu64dfa71369f1ea54a6fe97db7aeaa3db_94496_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/iswc-2019/2019-11-05-iswc-goodbye_hu64dfa71369f1ea54a6fe97db7aeaa3db_94496_da280338965f8420b4ea76ed0bbfee86.webp&#34;
               width=&#34;760&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>ISWS 2018</title>
      <link>https://sven-lieber.org/en/2018/07/09/isws-2018/</link>
      <pubDate>Mon, 09 Jul 2018 14:27:27 +0200</pubDate>
      <guid>https://sven-lieber.org/en/2018/07/09/isws-2018/</guid>
      <description>&lt;p&gt;In the first week of July I participated the first edition of the International Semantic Web Research Summer School (ISWS) in Bertinoro, Italy.
It was an exciting event, where I had a lot of fun and learned also a lot!
Continue reading for the full travel report.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h1 id=&#34;about-the-summer-school&#34;&gt;About the summer school&lt;/h1&gt;
&lt;p&gt;This first edition of the International Semantic Web Research Summer School (&lt;a href=&#34;http://isws2018.semanticwebschool.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;ISWS&lt;/a&gt;) was held in July 1 until July 7 in the University Residency Center in Bertinoro, Italy.
The Residency Center is basically an old castle used for conferences and training courses and is a very nice venue. The center also provides hotel-ish bedrooms with bathrooms.
&lt;strong&gt;113 people applied&lt;/strong&gt; for the summer school, from which &lt;strong&gt;60 students were selected&lt;/strong&gt; for participation! 73% were PhD students, 17% master students and 10% postdocs and researchers from mostly France, Italy and Germany, but also a lot of other Countries.
Among other, I met people from a university in Chile (which are originally from Cuba), the United States, the UK and Denmark. The gender balance could be better, but we at least reached a 63% to 37% ratio.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Street sign to the castle&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-castle-sign_hud4fee02e674f3268966bc71831503b68_4900605_a32196e6c63a1b416ce50ad7f1d9d8b8.webp 400w,
               /media/isws/2018-07-09-isws-castle-sign_hud4fee02e674f3268966bc71831503b68_4900605_b2990321fce4a21335e0738803c077c9.webp 760w,
               /media/isws/2018-07-09-isws-castle-sign_hud4fee02e674f3268966bc71831503b68_4900605_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-castle-sign_hud4fee02e674f3268966bc71831503b68_4900605_a32196e6c63a1b416ce50ad7f1d9d8b8.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;The summer school aimed to teach basics of the semantic web in the form of &lt;strong&gt;tutorials&lt;/strong&gt;, &lt;strong&gt;keynotes&lt;/strong&gt; and so called &lt;strong&gt;in-depth-pills&lt;/strong&gt;.
Furthermore there was a bigger group work, in which each team worked on a research topic with a paper and presentation as outcome.
Overall the summer school aimed to model and represent a researcher&amp;rsquo;s life in one week, including team work, socializing and deadlines!&lt;/p&gt;
&lt;h1 id=&#34;overview&#34;&gt;Overview&lt;/h1&gt;
&lt;p&gt;I&amp;rsquo;m gonna give now a short description of each program point with further details later on.
After an introduction of the school and the tutors, two &lt;a href=&#34;#poster-session&#34;&gt;poster session&lt;/a&gt; were held, where each participant presented his/her research.
In the course of the week different &lt;a href=&#34;#talks&#34;&gt;presentations&lt;/a&gt; (keynotes, tutorials, in-depth-pills) regarding the Semantic Web were given.
There was also a bit of room in the schedule to work in so called &lt;a href=&#34;#research-task-force-work&#34;&gt;Research Task Forces&lt;/a&gt;, small teams led by an tutor, to investigate in one perspective on the theme &amp;ldquo;Linked Open Data validity&amp;rdquo;. Besides socializing in the form of team work and coffee-break conversations, there was also a gala dinner with small party, an excursion to Rimini and things like slide karaoke 😉&lt;/p&gt;
&lt;p&gt;For the lazy reader I can already conclude, that it was a lot of fun and that I learned a lot!
You can make yourself a picture if you have a look at the Twitter hashtag &lt;a href=&#34;https://twitter.com/search?src=typd&amp;amp;q=%23isws2018&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;#isws2018&lt;/a&gt; to see what happened during the week and see some conclusions from the participants.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The coffee machine at the cafeteria&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-coffee-machine_hu1d31dc570fc4854783dd998a0a4325c7_2588934_88f6f982e8fe5a3c910875e1c13e10bb.webp 400w,
               /media/isws/2018-07-09-isws-coffee-machine_hu1d31dc570fc4854783dd998a0a4325c7_2588934_cbf041e0df36e13da29e2b27cad60e6f.webp 760w,
               /media/isws/2018-07-09-isws-coffee-machine_hu1d31dc570fc4854783dd998a0a4325c7_2588934_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-coffee-machine_hu1d31dc570fc4854783dd998a0a4325c7_2588934_88f6f982e8fe5a3c910875e1c13e10bb.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;poster-session&#34;&gt;Poster Session&lt;/h1&gt;
&lt;p&gt;Well, it was a poster session, not much to say about that. I saw a lot of interesting work. Integrating uncertainty in the Semantic Web Stack, a Web-based heritage platform built on Linked Data, Using machine-learning for Ontology matching and much, much more.
An award for best posters went to Maximilian Zocholl, &lt;a href=&#34;https://twitter.com/phparis2&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Piere-Henri Paris&lt;/a&gt; and &lt;a href=&#34;https://twitter.com/vale_carriero&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Valentina Carriero&lt;/a&gt;, Congratulations!
I was lucky to receive an honorably mention for creativity for &lt;a href=&#34;https://sven-lieber.org/en/2018/07/02/isws-2018-poster/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;my Poster&lt;/a&gt;, thank you!&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;The poster session&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-poster-session_hua0cec5696a8f957a94c95c4da4962915_2469472_12ade79050e1486951245c9f9399fa6d.webp 400w,
               /media/isws/2018-07-09-isws-poster-session_hua0cec5696a8f957a94c95c4da4962915_2469472_806a6be0ad7bcfdce10334b862528c72.webp 760w,
               /media/isws/2018-07-09-isws-poster-session_hua0cec5696a8f957a94c95c4da4962915_2469472_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-poster-session_hua0cec5696a8f957a94c95c4da4962915_2469472_12ade79050e1486951245c9f9399fa6d.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;talks&#34;&gt;Talks&lt;/h1&gt;
&lt;p&gt;Keynotes were given by &lt;a href=&#34;http://people.kmi.open.ac.uk/motta/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Enrico Motta&lt;/a&gt; about Data Analytics,
&lt;a href=&#34;https://twitter.com/martasabou&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marta Sabou&lt;/a&gt; about Rigour and Relevance with Design Science,
and Sebastian Rudolph about a Logician&amp;rsquo;s view of the Semantic Web.
The keynotes covered more high level concepts and lessons learned. They were very inspiring and gave a good perspective on a researcher&amp;rsquo;s life.
Additionally three tutorials were given, which focused more on concrete techniques.
In the tutorial given by &lt;a href=&#34;https://twitter.com/Maria11576561&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Maria-Esther Vidal&lt;/a&gt; and Sebastian Rudolph, we learned about the basics of Reasoning and SPARQL query execution.
&lt;a href=&#34;https://twitter.com/cldamat&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Claudia d&amp;rsquo;Amato&lt;/a&gt; and &lt;a href=&#34;http://users.jyu.fi/~miselico/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Michael Cochez&lt;/a&gt; talked about Machine Learning.
&lt;a href=&#34;https://twitter.com/johndmk&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;John Domingue&lt;/a&gt; talked about Blockchain and decentralization (&lt;a href=&#34;https://twitter.com/rubenverborgh&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Ruben Verborgh&lt;/a&gt; was supposed to give the talk as well, but unfortunately he couldn&amp;rsquo;t make it to the summer school).&lt;/p&gt;
&lt;p&gt;As a third kind-of talk, the summer school offered so called in-depth-pills. &lt;a href=&#34;https://twitter.com/aldogangemi&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Aldo Gangemi&lt;/a&gt; talked about ontology design patterns, &lt;a href=&#34;https://twitter.com/merpeltje&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Marieke van Erp&lt;/a&gt; about Natural Language Processing and Marta Sabou about crowdsourcing.
Overall, each talk gave a lot of insights and the numerous questions from the audience were answered.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;A slide from Marta Sabou&amp;amp;rsquo;s talk about crowdsourcing&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-marta-sabou-talk_hu758f60e8ab11f1158d24ca381bfa3a8b_2936756_00551dc08a68b64a8e3d4a8c452696e9.webp 400w,
               /media/isws/2018-07-09-isws-marta-sabou-talk_hu758f60e8ab11f1158d24ca381bfa3a8b_2936756_7b8178770d7ebd43117a79af4ceae279.webp 760w,
               /media/isws/2018-07-09-isws-marta-sabou-talk_hu758f60e8ab11f1158d24ca381bfa3a8b_2936756_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-marta-sabou-talk_hu758f60e8ab11f1158d24ca381bfa3a8b_2936756_00551dc08a68b64a8e3d4a8c452696e9.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;research-task-force-work&#34;&gt;Research Task Force work&lt;/h1&gt;
&lt;p&gt;This was probably the most interesting part of the summer school.
Based on a bigger subject, which was Linked Open Data Validity, each tutor presented research questions covering his/her point of view on the subject.
Participants could indicate preferences regarding tutors and then Research Task Forces (RTF&amp;rsquo;s) were built.
The assignment of participants to tutors took place after dinner at the evening of the second day. Each RTF was assigned a name, following a scientifically sound methodology, a speaking hat!&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Assigning Research Task Force names with the magical hat&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-research-task-force-selection_hu7b497301a85c33326637318a9b3c081a_158479_750f22e9cc554243dd8626dfc6149c9c.webp 400w,
               /media/isws/2018-07-09-isws-research-task-force-selection_hu7b497301a85c33326637318a9b3c081a_158479_a39314f9077c4fc7f6828ee0dab6626e.webp 760w,
               /media/isws/2018-07-09-isws-research-task-force-selection_hu7b497301a85c33326637318a9b3c081a_158479_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-research-task-force-selection_hu7b497301a85c33326637318a9b3c081a_158479_750f22e9cc554243dd8626dfc6149c9c.webp&#34;
               width=&#34;760&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;

Thus we ended up with the Teams&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;Hufflepuff (John Domingue), validity in a decentralized disintermediated world&lt;/li&gt;
&lt;li&gt;Ravenclaw (&lt;a href=&#34;https://twitter.com/lysander07&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Harald Sack&lt;/a&gt;), how does validity relate to context? (&lt;strong&gt;my task force&lt;/strong&gt;)&lt;/li&gt;
&lt;li&gt;Gryffindor (Aldo Gangemi), validity patterns and how to establish metrics&lt;/li&gt;
&lt;li&gt;Mordor (Claudia d&amp;rsquo;Amato), can validity be formalized? learning validity&lt;/li&gt;
&lt;li&gt;The 42&amp;rsquo;s (Marieke Van Erp), validity for incomplete information or fluid definitions&lt;/li&gt;
&lt;li&gt;The Deloreans (Valentina Presutti), validity and common-sense&lt;/li&gt;
&lt;li&gt;the Jedis (Maria-Esther Vidal), validity in SPARQL federated querying&lt;/li&gt;
&lt;li&gt;Dragons (Sebastian Rudolph), logical LOD validity&lt;/li&gt;
&lt;li&gt;Hobbits (Michael Cochez), detect anomalies with e.g. Machine Learning&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Each RTF had to prepare an 8-page research report, a 10 minutes presentation and a one minute funny video.
The best presentation was awarded to the Hufflepuffs and the Deloreans!
Again the Hufflepuffs were awarded with the best research report, this time together with the Jedis.&lt;/p&gt;
&lt;h1 id=&#34;social-events&#34;&gt;Social Events&lt;/h1&gt;
&lt;p&gt;An important part of the summer school was fun!
To achieve that we had multiple events like slide karaoke (40 seconds time to present a random power point slide), dinner in Bertinoro, an excursion to Rimini with subsequent dinner at the beach and a Gala dinner with DJ and dancefloor.
Usually one could also find multiple groups and/or tutors in a close-by gelateria or bar in the evening.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Sight Seing in Rimini - Excursion&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-excursion-rimini_hu0294b174b4fb6ee948358a8c7fc33e6e_1601359_e737b57e286369467fc5461d3d7371b6.webp 400w,
               /media/isws/2018-07-09-isws-excursion-rimini_hu0294b174b4fb6ee948358a8c7fc33e6e_1601359_8594e57bd7b10f7a164c729c0d931659.webp 760w,
               /media/isws/2018-07-09-isws-excursion-rimini_hu0294b174b4fb6ee948358a8c7fc33e6e_1601359_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-excursion-rimini_hu0294b174b4fb6ee948358a8c7fc33e6e_1601359_e737b57e286369467fc5461d3d7371b6.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;my-conclusion&#34;&gt;My Conclusion&lt;/h1&gt;
&lt;p&gt;Go there! I haven&amp;rsquo;t been to other summer schools, but I&amp;rsquo;ve heard that they can be boring. But this one is an exception!
The Semantic Web Community is quite small, here you will meet your peers, people you will work together with, see on conferences and so on.
As stated above, I really enjoyed it and learned a lot.
But notice, if you go to Italy, bring a &lt;a href=&#34;https://twitter.com/SvenLieber/status/1013457307282309121&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;power adapter&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Evening view from the balcony&#34; srcset=&#34;
               /media/isws/2018-07-09-isws-last-day-sunset_hue9c2d15c8a3b2c61def1fad51dd6fc81_3623218_d201657554d88b7f407961474e215716.webp 400w,
               /media/isws/2018-07-09-isws-last-day-sunset_hue9c2d15c8a3b2c61def1fad51dd6fc81_3623218_3038d21c334d125cd1a51b009cd5a131.webp 760w,
               /media/isws/2018-07-09-isws-last-day-sunset_hue9c2d15c8a3b2c61def1fad51dd6fc81_3623218_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws/2018-07-09-isws-last-day-sunset_hue9c2d15c8a3b2c61def1fad51dd6fc81_3623218_d201657554d88b7f407961474e215716.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>ISWS 2018 Poster</title>
      <link>https://sven-lieber.org/en/2018/07/02/isws-2018-poster/</link>
      <pubDate>Mon, 02 Jul 2018 10:20:27 +0200</pubDate>
      <guid>https://sven-lieber.org/en/2018/07/02/isws-2018-poster/</guid>
      <description>&lt;p&gt;For the ISWS summer school I prepared an interactive poster. Read in his post what it is about and see how it was made.&lt;/p&gt;
&lt;p&gt;My summer this year will start with the attendance of the International Semantic Web Research Summer School (&lt;a href=&#34;http://stlab.istc.cnr.it/isws/wordpress/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;ISWS&lt;/a&gt;) in Bertinoro, Italy.
For that we were invited to bring a poster, describing our PhD research. My current research tasks involve so called shapes, describing data within a graph of information.
&lt;strong&gt;Since graphs and shapes are very tangible concepts&lt;/strong&gt;, and one of the main goals of the poster is to gather feedback, we had the idea of making the poster a bit interactive.
My supervisor suggested post-its, but as I have a lot of creative energy (and probably procrastinate writing papers), I decided to go for a small Knowledge Jigsaw Puzzle.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;h1 id=&#34;what-is-the-whole-context&#34;&gt;What is the whole context?&lt;/h1&gt;
&lt;p&gt;Within the Semantic Web, information are modeled using vocabularies of terms, relationships among these terms and also constraints, things which aren&amp;rsquo;t allowed.
The combination of the mentioned vocabularies, relationships and constraints is usually called an ontology and can be expressed in a formal language.
That formalization allows computer programs then to &amp;ldquo;understand&amp;rdquo; the data, but also to infer new knowledge. For example, if I model persons and say that there is a relationship called &amp;ldquo;knows&amp;rdquo; which means that both people know each other, and I have the Information that Sven knows you, then the information that you know Sven can be inferred.&lt;/p&gt;
&lt;h1 id=&#34;what-has-this-now-to-do-with-shapes&#34;&gt;What has this now to do with shapes?&lt;/h1&gt;
&lt;p&gt;Well, constraints modeled in an ontology are used to infer new knowledge, they are not meant to use for validating data, e.g. that a person must have at least one last name. There are various reasons for that, which are out of scope for this blog post. Important is now, that recently the W3C committee released a standard to express constraints as shapes on data used i.a. for validation.
It&amp;rsquo;s all quite new, so there are no clear methodologies yet, when to use shapes of data, when to use constraints in the ontology or when to use a formal rule language (yes these exist as well).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Puzzle preparation&#34; srcset=&#34;
               /media/isws-poster/2018-07-02-puzzle-preparation_hu0e795173ff2ab0512fcaf9a92aec5416_140452_8e1409f2dd4380bbdc078de66ab985b0.webp 400w,
               /media/isws-poster/2018-07-02-puzzle-preparation_hu0e795173ff2ab0512fcaf9a92aec5416_140452_19448aae5801f75c111f7512aeb0949c.webp 760w,
               /media/isws-poster/2018-07-02-puzzle-preparation_hu0e795173ff2ab0512fcaf9a92aec5416_140452_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws-poster/2018-07-02-puzzle-preparation_hu0e795173ff2ab0512fcaf9a92aec5416_140452_8e1409f2dd4380bbdc078de66ab985b0.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;okay-okay-enough-background-whats-the-interactive-poster-now&#34;&gt;Okay okay, enough background, what&amp;rsquo;s the interactive poster now?&lt;/h1&gt;
&lt;p&gt;I am currently working on a project, which is about improving the customer journey when using public online services.
An example: You want to move, so you have to do certain tasks at the city you are currently living in, like e.g. sign-off and provide your new address.
At the city you are moving to, you have to provide certain information as well, e.g. if you&amp;rsquo;d like to request a parking spot for your car.
These are all steps you should do, and they can be understood as a workflow.
The aim of the project is to build customized user interfaces across cities. So that it seems you are only interacting with one system, whereas in the background multiple forms of different websites are filled in.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Puzzle pieces&#34; srcset=&#34;
               /media/isws-poster/2018-07-02-poster-pieces_hu85079b3d972520c01e478ee1de44ca4f_113761_8e7fbbb56445ebff892b853f1dc208e5.webp 400w,
               /media/isws-poster/2018-07-02-poster-pieces_hu85079b3d972520c01e478ee1de44ca4f_113761_b59a257aaf70b5bd4e93ea0d93e4e54c.webp 760w,
               /media/isws-poster/2018-07-02-poster-pieces_hu85079b3d972520c01e478ee1de44ca4f_113761_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws-poster/2018-07-02-poster-pieces_hu85079b3d972520c01e478ee1de44ca4f_113761_8e7fbbb56445ebff892b853f1dc208e5.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;The customization here is the key point!&lt;/strong&gt;
Beside simple optimization like asking you only once for your old address (or even reuse it if the government has that information already), we can think of more personalization.&lt;/p&gt;
&lt;p&gt;That&amp;rsquo;s were shapes come in.
Given one or multiple workflows, containing different steps, your preferences can be modeled as shape(s) and be used to select the optimal steps for you.
Similarly, context information like the device you are using or a certain life situation you are currently in, are somewhat constrains which can be used for optimal steps selection.&lt;/p&gt;
&lt;p&gt;What I did now for the &lt;a href=&#34;https://www.slideshare.net/SvenLieber/knowledge-jigsaw-puzzle&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;poster&lt;/a&gt; is to model the steps to do (blue), conditions need to be met for that step (grey) a transaction (green) with a certain system (orange).
In the simple puzzle setup there are shapes for secure transactions (https) and conditions.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Puzzle pieces on the poster&#34; srcset=&#34;
               /media/isws-poster/2018-07-02-puzzle-on-poster_hu6f4b2651c32fe786506abb5801109070_122558_65e5013f0f6c546f68bd970dddcd166c.webp 400w,
               /media/isws-poster/2018-07-02-puzzle-on-poster_hu6f4b2651c32fe786506abb5801109070_122558_653698935d1a3bd9f118b75bb6830612.webp 760w,
               /media/isws-poster/2018-07-02-puzzle-on-poster_hu6f4b2651c32fe786506abb5801109070_122558_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/isws-poster/2018-07-02-puzzle-on-poster_hu6f4b2651c32fe786506abb5801109070_122558_65e5013f0f6c546f68bd970dddcd166c.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;That is of course only a toy example of one possible way of modeling such steps.&lt;/p&gt;
&lt;p&gt;Research-wise it is interesting to investigate in the use of shapes, finding out when they are useful and when it is better to use other means of representing constraints.
If shapes are used it is also interesting to investigate if shapes shall be only applied one after each other or if they can be combined somehow. Can we derive new shapes?
Also Reasoning becomes interesting, given different context constraints and other user constraints, can we infer new shapes?&lt;/p&gt;
&lt;p&gt;A lot of research to be done!
&lt;strong&gt;I am looking forward to get a lot of interesting feedback in the poster session.&lt;/strong&gt;&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>WebSci 2018</title>
      <link>https://sven-lieber.org/en/2018/06/23/websci-2018/</link>
      <pubDate>Sat, 23 Jun 2018 11:27:27 +0200</pubDate>
      <guid>https://sven-lieber.org/en/2018/06/23/websci-2018/</guid>
      <description>&lt;p&gt;At the end of May I had the chance to attend the WebScience conference. A venue where computer scientists and social scientists meet!
ACM Turing Lecture by Sir Tim Berners-Lee, adaptive learning analytics, roots of radicalisation on Twitter, internet regulation in Russia.
These are only some of the very interesting topics. Continue reading for the full travel report.&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;p&gt;Rather than chronologically list all the talks, I will focus on one or two talks per day which I liked the most and the social activities.
Let&amp;rsquo;s first start with a bit of context, what is the WebSci conference?&lt;/p&gt;
&lt;p&gt;The &lt;a href=&#34;https://websci18.webscience.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;WebScience&lt;/a&gt;  conference is all about the Web, the largest socio-technical network in human history, and aims to bring together researchers from various disciplines like computer science, sociology, economics or psychology.
This year&amp;rsquo;s WebSci was already the 10th edition and one special agenda point was the ACM Turing Lecture held by Sir Tim Berners-Lee the inventor of the Web (beside the 150 conference attendees, 1,500 guests were expected).&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I liked the most&lt;/strong&gt;
The interdisciplinary approach of the conference. Algorithms and data meet sociological and psychological use-cases.&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;What I didn&amp;rsquo;t like&lt;/strong&gt;
Each session had a scheduled Q&amp;amp;A panel, but none of the sessions I attended made use of it and additionally some talks were out of sync across sessions which made it difficulty to switch talks within a session.&lt;/p&gt;
&lt;h1 id=&#34;lile-workshop&#34;&gt;LILE workshop&lt;/h1&gt;
&lt;p&gt;I already attended the program on Sunday,
as I had to present our work about &lt;a href=&#34;https://sven-lieber.org/en/publication/linked-data-generation-for-adaptive-learning-analytics-systems/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Linked Data Generation for Adaptive Learning Analytics Systems&lt;/a&gt; in the Linked Learning workshop (&lt;a href=&#34;https://lile2018.wordpress.com/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;LILE&lt;/a&gt;).
Past editions were held together with semantic web conferences (ESWC or ISWC) and the WWW conference.
This year focused more on interdisciplinary approaches (hence together with WebSci), which was also reflected in the scoring for the best paper award which was won by &lt;a href=&#34;https://at.linkedin.com/in/simonekopeinik&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Simone Kopeinik&lt;/a&gt; and colleagues and their work about the adaption of an open source social bookmarking system to observe critical information behaviour. Congratulations!&lt;/p&gt;
&lt;p&gt;A keynote was given by &lt;a href=&#34;https://twitter.com/inge_molenaar&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Inge Molenaar&lt;/a&gt; about Self regulated Learning, who early in the beginning of the talk also raised the question: &amp;ldquo;Which processes we want computers to take over?&amp;rdquo;. Later in the presentation I learned that self report is not a really good method to measure self regulation.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Keynote about self regulated learning given by Inge Molenaar&#34; srcset=&#34;
               /media/websci-2018/_hu5489f3685fb866e59a982c10505ddadc_2515781_ff83d85efed64c9a3651e72fe0007657.webp 400w,
               /media/websci-2018/_hu5489f3685fb866e59a982c10505ddadc_2515781_f1b3e635cddbd419487256b2f21dbc69.webp 760w,
               /media/websci-2018/_hu5489f3685fb866e59a982c10505ddadc_2515781_45bc36b31206b2f464d27aea0f8da325.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/_hu5489f3685fb866e59a982c10505ddadc_2515781_ff83d85efed64c9a3651e72fe0007657.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;There was also a Keynote by &lt;a href=&#34;https://twitter.com/johndmk&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;John Domingue&lt;/a&gt; about Using Blockchains to Open Higher Education. I learned among other things about Uber style universities, which are basically virtual universities where a Blockchain is used to record and verify students &amp;lsquo;attendance&amp;rsquo; in online courses, grades and diplomas.&lt;/p&gt;
&lt;h1 id=&#34;monday&#34;&gt;Monday&lt;/h1&gt;
&lt;p&gt;The first official conference day started directly with the &lt;strong&gt;&amp;ldquo;Best of Web Science&amp;rdquo; session&lt;/strong&gt;, where &lt;a href=&#34;https://twitter.com/miriam_fs&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Miriam Fernandez&lt;/a&gt; talked about understanding the roots of radicalisation on Twitter.
Their approach: &amp;ldquo;Don&amp;rsquo;t try to detect whether someone is radical, but measure the influence on the micro, meso and macro level&amp;rdquo;.
Interesting fact is, that the terror organization ISIS has a weekly magazine which is now used by researchers to learn word embeddings for further research.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Best paper award presentation about understanding the roots of radicalisation on twitter, given by Miriam Fernandez&#34; srcset=&#34;
               /media/websci-2018/_hu831d642798263e96e7b40cbac998e9ee_2309737_d754ecb276448ea64ba1ea6033149f4e.webp 400w,
               /media/websci-2018/_hu831d642798263e96e7b40cbac998e9ee_2309737_8e48faa5acd5354dfffbb36c1ff5b34e.webp 760w,
               /media/websci-2018/_hu831d642798263e96e7b40cbac998e9ee_2309737_a0d53f100653978b5dda0dc8561ea944.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/_hu831d642798263e96e7b40cbac998e9ee_2309737_d754ecb276448ea64ba1ea6033149f4e.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Another talk from that session focused on third party tracking in the mobile ecosystem.
&lt;a href=&#34;https://twitter.com/rdbinns&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Reuben Binns&lt;/a&gt; pointed out that parent companies relationships often are data sharing relationships and that one app usual has 5 trackers installed.
Their x-ray transparency tool is available at &lt;a href=&#34;https://github.com/sociam/xray&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;GitHub&lt;/a&gt; proudly presented to you by Mircosoft 😛 (which by the way was also mentioned in the talk as one most prevalent tracker company).&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Slide from Reuben Binns about cross border data transfer&#34; srcset=&#34;
               /media/websci-2018/2018-06-23-websci-reuben-binns-cross-border-transfer_hu390888f64ea53ee231f3f2accfc985cf_2339420_fb184d8dc378cf5d76889f54515c60de.webp 400w,
               /media/websci-2018/2018-06-23-websci-reuben-binns-cross-border-transfer_hu390888f64ea53ee231f3f2accfc985cf_2339420_ac0130e03973ebc733690e5cdcfb7353.webp 760w,
               /media/websci-2018/2018-06-23-websci-reuben-binns-cross-border-transfer_hu390888f64ea53ee231f3f2accfc985cf_2339420_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/2018-06-23-websci-reuben-binns-cross-border-transfer_hu390888f64ea53ee231f3f2accfc985cf_2339420_fb184d8dc378cf5d76889f54515c60de.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Later that day I mostly attended talks in the &lt;strong&gt;&amp;ldquo;Flow, Information and News&amp;rdquo; session&lt;/strong&gt;.
Interesting there, was the investigation of effects of Google&amp;rsquo;s result pages for evaluating the credibility of online news sources given by &lt;a href=&#34;https://twitter.com/exploredeeper&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Emma Lurie&lt;/a&gt;.
I learned that snopes.com performs fact-checking and that some study participants value social media signals like facebook ranking or tweet recency.
More information can be found &lt;a href=&#34;http://cs.wellesley.edu/~credlab/websci18/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;here&lt;/a&gt;.&lt;/p&gt;
&lt;p&gt;Monday was also the poster session day.
On the Web Science spectrum everything from the &lt;strong&gt;online media discourse in authoritarian elections in the case of 2016 St. Petersburg electoral campaign&lt;/strong&gt; to &lt;strong&gt;Linkflows: enabling a web of linked semantic publishing workflows&lt;/strong&gt;, everything was covered.
The best poster award was given to &lt;a href=&#34;https://twitter.com/pip__t?lang=en&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Pip Thornton&lt;/a&gt; and her poster about &lt;a href=&#34;https://linguisticgeographies.com/2016/06/12/poem-py-a-critique-of-linguistic-capitalism/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;{poem}.py&lt;/a&gt;. Famous poems or texts printed on a receipt next to their current  monetary value determined by GoogleAdWords.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Best poster award ballot bin&#34; srcset=&#34;
               /media/websci-2018/2018-06-23-websci-best-poster-bin_hu3644d4c772b2e238cddce9e853793d6e_2887306_07a2c008f47ec967017a96cabc11296d.webp 400w,
               /media/websci-2018/2018-06-23-websci-best-poster-bin_hu3644d4c772b2e238cddce9e853793d6e_2887306_00d8cd50628fd446dbd805a64e7e535f.webp 760w,
               /media/websci-2018/2018-06-23-websci-best-poster-bin_hu3644d4c772b2e238cddce9e853793d6e_2887306_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/2018-06-23-websci-best-poster-bin_hu3644d4c772b2e238cddce9e853793d6e_2887306_07a2c008f47ec967017a96cabc11296d.webp&#34;
               width=&#34;570&#34;
               height=&#34;760&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;The day was rounded out with a social gathering and dinner for PhD students sponsored by &lt;a href=&#34;http://www.webscience.org/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Web Science Trust&lt;/a&gt;.
It was tasty and fun, thank you!&lt;/p&gt;
&lt;h1 id=&#34;tuesday&#34;&gt;Tuesday&lt;/h1&gt;
&lt;p&gt;Tuesday was in the shadow of the ACM Turing Lecture held by &lt;a href=&#34;https://twitter.com/timberners_lee&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Sir Tim Berners-Lee&lt;/a&gt; (the inventor of the web).
Next to the 150 conference participants around 1,500 more guests were expected.
The difference between the Internet and the Web, according to Tim: When retrieving an server error while requesting a website, one can clearly see that the Internet works, but the web part has issues.
An important take-home message for me was the call to make use of our rights and protest against initiatives which are threaten the Web, like cancellation of the net neutrality.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Tim Berners-Lee about the web&#34; srcset=&#34;
               /media/websci-2018/2018-06-23-websci-TBL_hu3be8957e2b3fd764e68fcd89890fe1f8_1992084_843dd747be6b8ae4581e3c8c3e89f20e.webp 400w,
               /media/websci-2018/2018-06-23-websci-TBL_hu3be8957e2b3fd764e68fcd89890fe1f8_1992084_dbcc92121d856def6f03afafe80ecffe.webp 760w,
               /media/websci-2018/2018-06-23-websci-TBL_hu3be8957e2b3fd764e68fcd89890fe1f8_1992084_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/2018-06-23-websci-TBL_hu3be8957e2b3fd764e68fcd89890fe1f8_1992084_843dd747be6b8ae4581e3c8c3e89f20e.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;In the &lt;strong&gt;session &amp;ldquo;Methods and Practice&amp;rdquo;&lt;/strong&gt; &lt;a href=&#34;https://www.cs.ox.ac.uk/people/alex.darer/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Alexander Darer&lt;/a&gt; was talking about the automatic discovery of internet censorship by web crawling.
His colleagues and him tried to make transparent how countries block the Internet.
One recommendation from the audience regarding the performed research was to take the initial seed list of blocked websites which was taken from Wikipedia and
try to find similar sites using techniques like cluster analysis.
Also a deeper analysis of the distribution of types of blocked websites was suggested (e.g. gambling, news, entertainment).
&lt;strong&gt;This social science viewpoint on the data, which can improve the performed research, is in my opinion one of the main benefits of events like the WebSci!&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;Tuesday was also the day of the official social event of the conference.
We had a nice boat tour (unfortunately with a bit of rain) through the canals of Amsterdam and ended up having dinner in the restaurant Ij-kantine, where also the awarding of best papers happened.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Selfie made during the boat tour through the canals of Amsterdam&#34; srcset=&#34;
               /media/websci-2018/2018-06-23-websci-boat-tour_hu0b7f80efde6504e4c704856b5328c8a0_1280659_f88241a00d2d6ac054f70f293ca63007.webp 400w,
               /media/websci-2018/2018-06-23-websci-boat-tour_hu0b7f80efde6504e4c704856b5328c8a0_1280659_951680320c955a7b4cb50596a83de0eb.webp 760w,
               /media/websci-2018/2018-06-23-websci-boat-tour_hu0b7f80efde6504e4c704856b5328c8a0_1280659_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/2018-06-23-websci-boat-tour_hu0b7f80efde6504e4c704856b5328c8a0_1280659_f88241a00d2d6ac054f70f293ca63007.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;h1 id=&#34;wednesday&#34;&gt;Wednesday&lt;/h1&gt;
&lt;p&gt;The last keynote of the conference was held by &lt;a href=&#34;https://twitter.com/johndmk&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;John Domingue&lt;/a&gt; and was about the future of semantics on the web.
After going through some history of semantics, he talked about &lt;strong&gt;real-world issues regarding provenance and centralization&lt;/strong&gt;.
The windrush generation in the UK are British citizens who came to the UK from the Commonwealth.
Their rights were guaranteed in the Immigration Act of 1971.
But now according to a new law, these people have to prove continuous residence in the UK since 1973.
A task which is hard to achieve if you don&amp;rsquo;t have detailed provenance of what you did in those years.
The Home Office also destroyed the landing cards (which were centralized stored), which makes it even harder.
The keynote was also about machine learning algorithms and how they can use RDF triples in a vector space model.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Conclusion slide of John Domingue&amp;amp;rsquo;s keynote about the future of semantics on the Web&#34; srcset=&#34;
               /media/websci-2018/2018-06-23-websci-john-domingue-future-semantics_huc587adc9f8f73ca2d6c4b5666dae491b_1963752_e4ea64b86ff5d7d6864a0d7f9b262ae6.webp 400w,
               /media/websci-2018/2018-06-23-websci-john-domingue-future-semantics_huc587adc9f8f73ca2d6c4b5666dae491b_1963752_592abda19a4eb3213efd06dd55aa1fd5.webp 760w,
               /media/websci-2018/2018-06-23-websci-john-domingue-future-semantics_huc587adc9f8f73ca2d6c4b5666dae491b_1963752_1200x1200_fit_q75_h2_lanczos.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/2018-06-23-websci-john-domingue-future-semantics_huc587adc9f8f73ca2d6c4b5666dae491b_1963752_e4ea64b86ff5d7d6864a0d7f9b262ae6.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Beside the mentioned facts, my take home message from the talk was the position of semantics to sit in between machines and humans.
Also that even though - as pointed out by Miriam Fernandez - semantics are subjective, the bias is at least visible and provenance can be expressed to see where the bias comes from.&lt;/p&gt;
&lt;p&gt;Another interesting talk that day was &lt;strong&gt;Internet Regulation Media Coverage in Russia: Topics and Countries&lt;/strong&gt; given by &lt;a href=&#34;https://www.hse.ru/en/staff/shirokanova&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Anna Shirokanova&lt;/a&gt;, where I learned that a small blogger sphere went to a huge politics sphere in Russia.
In the beginning of the 2000s a lot of Blogs existed which were also critically talking about the Government
and that later on multiple laws were introduced to e.g. make bloggers recognize themselves as media companies (with regulations and obligations involved) or laws regarding the storage of data of Russian citizens and an anti VPN law.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;Anna Shirokanova presents a timeline of internet regulation in Russia&#34; srcset=&#34;
               /media/websci-2018/_hu9791a7c3e0d6ed21e47e91194c9ffced_1523748_53130ffb40df764c2ba733a637cc48e3.webp 400w,
               /media/websci-2018/_hu9791a7c3e0d6ed21e47e91194c9ffced_1523748_e2310e2ff073fa477e2359a15f4e759d.webp 760w,
               /media/websci-2018/_hu9791a7c3e0d6ed21e47e91194c9ffced_1523748_185678c733eb03d09c62ff04ac2447ed.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/websci-2018/_hu9791a7c3e0d6ed21e47e91194c9ffced_1523748_53130ffb40df764c2ba733a637cc48e3.webp&#34;
               width=&#34;760&#34;
               height=&#34;570&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;&lt;a href=&#34;https://twitter.com/gefiont&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Gefion Thuermer&lt;/a&gt; was talking about online participation within the Green Party of Germany.
According to a survey they made, young people think that online participation is a good thing for elderly, as they are not that mobile anymore.
Similarly elderly think that online participation is a good thing for younger people as their are digital natives anyway.
Future research will be a panel survey before and after the online participation tool gets introduced.&lt;/p&gt;
&lt;p&gt;That was the WebScience conference from my perspective. As already pointed out, a great venue where Computer Science and Social Sciences meet.
I met a lot of interesting people and hope to attend next year as well!&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>Transparency in Legislation, a must-have</title>
      <link>https://sven-lieber.org/en/2017/07/13/transparency-in-legislation-a-must-have/</link>
      <pubDate>Thu, 13 Jul 2017 00:00:00 +0000</pubDate>
      <guid>https://sven-lieber.org/en/2017/07/13/transparency-in-legislation-a-must-have/</guid>
      <description>&lt;p&gt;The cabinet of Germany has the plan to publish the drafts of laws and associated position papers of the lobbies. This should be done for more than 600 laws of the past 4 years. This transparency can reveal a lot. For instance, which Lobby influenced which law actively. Nearly 17,000 documents are planned to be published. But how can one handle this huge amount?&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;p&gt;4 years, over 600 laws, nearly 17,000 documents. These numbers are currently everywhere in the German media. The law &amp;ldquo;Informationsfreiheitsgesetz&amp;rdquo;, enables all citizens to
request administrative data of the government. This include documents of the origin of law, like the first draft, the position papers of the lobbies regarding this draft or
the later draft of the government, which considers the position papers.
Over 1,600 citizen requests were just received within a week, after the portal &lt;a href=&#34;https://fragdenstaat.de/gesetze/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;fragdenstaat.de&lt;/a&gt; (ask the state), launched the campaign #GläserneGesetze (transparent laws).
That are more requests, than received in the whole past year.
The administrative overhead seems overwhelming, according to &lt;a href=&#34;https://netzpolitik.org/2017/glaesernegesetze-erfolgreich-bundesregierung-will-tausende-lobby-dokumente-veroeffentlichen/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Netzpolitik.org&lt;/a&gt;, the German cabinet plans to gradually publish all law related documents of the past legislative period, which is 4 years.&lt;/p&gt;
&lt;h2 id=&#34;why-transparency-is-a-must-have&#34;&gt;Why transparency is a must-have?&lt;/h2&gt;
&lt;p&gt;In a democracy, the citizens are voting politicians to office. These politicians have to make decisions on a daily basis, which directly affects the citizens due to the created laws.
But a politician is also just a human and cannot be expert in anything. After a first draft of a law is created, various interest groups, lobbyists, give their input.
The respective groups are usually closer to the matter of the law. For instance the group BITKOM (the federal association for information technology) is active for subjects around IT.
Besides represent direct members, these groups also represent companies. And as the name interest group suggests, these groups have interests.
This is where transparency comes into play. Was a law improved, due to the technical competence of the lobby, or were financial interests the main reason and the citizens have to suffer?&lt;/p&gt;
&lt;h2 id=&#34;how-to-handle-17000-documents&#34;&gt;How to handle 17,000 documents?&lt;/h2&gt;
&lt;p&gt;Who, did when, how participate in the development phase of a law. These meta information, which is called provenance, provide a real added value to assess a law. &lt;a href=&#34;https://sven-lieber.org/en/2017/04/07/what-is-provenance/&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Provenance&lt;/a&gt; is the focus of my research.
If the documents of the law development process are published in a digital format, they can be further processed. With techniques of the semantic web, raw data can be semantically enriched.&lt;/p&gt;
&lt;p&gt;One with specific questioning is then able to search the data efficiently with a computer program.
The sentence &amp;ldquo;The apple is red&amp;rdquo; for instance, can be semantically enriched, by saying an apple is a fruit and red is a color.
Assumed the whole dataset (all sentences) is semantically enriched with the same technique, specific questions can be answered. E.g.&amp;quot; &lt;em&gt;Give me all sentences, in which red fruits occur&lt;/em&gt;&amp;quot; or &amp;ldquo;&lt;em&gt;Give me all sentences which are about fruits&lt;/em&gt;&amp;rdquo;.&lt;/p&gt;
&lt;p&gt;With these techniques, simpler or more complex questions regarding legislation and the involved organizations can be answered.&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>Preparing Travels Abroad</title>
      <link>https://sven-lieber.org/en/2017/06/28/preparing-travels-abroad/</link>
      <pubDate>Wed, 28 Jun 2017 00:00:00 +0000</pubDate>
      <guid>https://sven-lieber.org/en/2017/06/28/preparing-travels-abroad/</guid>
      <description>&lt;p&gt;You are traveling abroad to a conference and you think you are prepared? Think again! There is always something you probably will forget about and if some unfavorable circumstances following each other you may find yourself in real trouble. Check out my list, learn from my mistakes and extend it, in case I missed something!&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;p&gt;In most of the cases a good preparation seems time consuming and who need all this, nothing will happen, right?! On a recent trip to Seattle in the United states the worst case happened,
I lost my passport. A lot of steps needed to be done, but outside of the usual environment obvious, self-evident things become challenging.
Let’s go through my use-case to check what is needed. A small remark, I heed some of the mentioned points, but for some I didn’t even saw the problem arising.&lt;/p&gt;
&lt;h2 id=&#34;carry-copies-of-important-documents&#34;&gt;Carry copies of important documents&lt;/h2&gt;
&lt;p&gt;Having copies of all important documents. Depending on what disappeared, copies can save your life. Always make copies of your passport, your ID, driver license, Visa or other travel authorization documents like ESTA and carry it with you in printed form or digital.
In my case I still had my ID to identify myself and a printed version of the approved ESTA application. This document included my passport ID and allowed me to log-in online,
in order to find out other important passport information like its expiry date.
This information can help speed up the process of requesting a temporary passport from your country’s diplomatic mission.&lt;/p&gt;
&lt;p&gt;A short introduction to diplomatic missions. Usually your country has an embassy in the capitol of other countries. But you don’t have to travel to the capitol in order to get a temporary passport.
Your country may operate a hand full of consulates distributed across the foreign country. In my case the next consulate was in San Francisco which is still around 1.300 km away!
There could still be some honorary consulates nearby, which act in behalf of your country, google for that! In my case there was one in Seattle, but unfortunately they were on vacation!
That’s why I had to travel to Portland, which gave me time to write this blog post.
Ask also your Airline, for example the German Lufthansa used to take German citizens back to Germany just with the ID.&lt;/p&gt;
&lt;h2 id=&#34;phone-calls&#34;&gt;Phone calls&lt;/h2&gt;
&lt;p&gt;In any case, you have to make some calls in order to arrange everything, e.g. the consulate or the airline.
We are living in the 21st century, nearly everybody has some sort of a phone. Conversely, there is no need for phone booths to make calls anymore.
But abroad you may have to make calls, but your provider may charge significant fees! The amount you are usually paying for your whole monthly subscription will may be due for a single minute!
(For me it was around 6 Euro/min) Load money on your Skype account, WIFI is nearly everywhere.
Of course you can ask around if you can use someone else’s phone, at your hotel’s desk or at random stores, but being independent is much more helpful, especially if you are around in the city!&lt;/p&gt;
&lt;h2 id=&#34;activate-your-debitcredit-cards-fully-functionality&#34;&gt;Activate your debit/credit card&amp;rsquo;s fully functionality&lt;/h2&gt;
&lt;p&gt;Money rules the world! In general, having cash is very useful. Your bank might deactivate your debit card’s functioning in foreign countries. So, activate your debit card in advance for the use in the foreign country! I fortunately did this.&lt;/p&gt;
&lt;p&gt;In case you your consulate is not around the corner, you might have to undertake a travel there. Trains or especially intercity buses are booked online.
Sometimes online is the only way of purchase, or if not, you might have to pay a much higher price at the station or from the driver.
Also sometimes, credit cards are the only possible payment method. Both, Visa and MasterCard offer fraud detection techniques like “Verified by Visa” or “MasterCard SecureCode”.
If you aware of it, it can be your life vest. If you are not aware of it, this life vest might strangle you!
If these functionality isn’t activated beforehand, you cannot use the cards in online shopping environment, like your ticket purchase.
The activation can take multiple days, so activate your credit card’s online shopping functionality beforehand. Luckily I had friends and colleagues willing to helping me out!&lt;/p&gt;
&lt;h2 id=&#34;summarized-preparation&#34;&gt;Summarized preparation&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;Make copies of important documents beforehand&lt;/li&gt;
&lt;li&gt;Load money on your Skype account&lt;/li&gt;
&lt;li&gt;Activate debit card for the duration of your stay for the foreign country&lt;/li&gt;
&lt;li&gt;Activate your credit card’s online shopping capabilities&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;Please comment if you can think of something I maybe missed.&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>What is Provenance?</title>
      <link>https://sven-lieber.org/en/2017/04/07/what-is-provenance/</link>
      <pubDate>Fri, 07 Apr 2017 00:00:00 +0000</pubDate>
      <guid>https://sven-lieber.org/en/2017/04/07/what-is-provenance/</guid>
      <description>&lt;p&gt;Having Provenance or not having it, that could be worth than several hundred million dollar. Where does something come from, how was it changed since it exists, that is provenance. These informations help human beings to asses things. Which questions need to be answered and what is my role as computer scientists in this research field?!&lt;/p&gt;
&lt;div class=&#34;alert alert-note&#34;&gt;
  &lt;div&gt;
    If you want to get notified about more content or want to receive bi-weekly updates on FAIR and Linked Data,
please consider subscribing to my bi-weekly Newsletter.
  &lt;/div&gt;
&lt;/div&gt;
&lt;iframe src=&#34;https://fairdata.substack.com/embed&#34; width=&#34;100%&#34; style=&#34;border:1px solid #EEE; background:white;&#34; frameborder=&#34;0&#34; scrolling=&#34;no&#34;&gt;&lt;/iframe&gt;
&lt;p&gt;In 2015 the painting Les Femmes d’Alger from Picasso was sold for
unbelievable &lt;a href=&#34;http://www.independent.co.uk/arts-entertainment/art/news/pablo-picasso-les-femmes-dalger-version-o-sells-for-179m-and-sets-new-world-record-10243056.html&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;179 Mio Dollars&lt;/a&gt;.
The painting La Bella Principessa from Leonardo da Vinci was sold for surprisingly cheap 20,000 Dollars (see &lt;a href=&#34;https://link.springer.com/chapter/10.1007%2F978-3-319-40226-0_7&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Ludäscher&lt;/a&gt;).
But what is the main difference, why did the first painting obtained such a high price compared to the second?
The answer is, that there is a peace of paper which documented every owner and former places of custody of the first painting.
The same information for the second painting was &lt;a href=&#34;https://en.wikipedia.org/wiki/La_Bella_Principessa#Provenance&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;incomplete&lt;/a&gt;. This documented history of ownership is called Provenance.&lt;/p&gt;
&lt;h2 id=&#34;why-is-this-provenance-so-important-to-us&#34;&gt;Why is this Provenance so important to us?&lt;/h2&gt;
&lt;p&gt;Well, let’s say I sell a car and you want to buy it. If I can provide you a service record which documents how often the car was checked in a garage the last years you have more trust in its functionality. In the same manner is it also from interest how the car was changed, which parts are newer because they were replaced for example.
Of course there is no direct trust involved, but the provenance helps to form an opinion and allows one to draw conclusions about the object. Other properties than quality can be important as well, the value of the car might also be higher if the former owner was a celebrity.&lt;/p&gt;
&lt;p&gt;The car example shows the importance of provenance and the millions of dollars price difference of the paintings that it is definitely worth to capture!
There are plenty of use cases where provenance is helpful, just think about the food you are eating, where does it come from? How was it produced? Was the cooling chain never interrupted?
All these information influence your buying decision. Providing these information is therefore also in the interest of the producer.
There is actually a start-up company called &lt;a href=&#34;https://www.provenance.org&#34; target=&#34;_blank&#34; rel=&#34;noopener&#34;&gt;Provenance&lt;/a&gt;, which tries to make the supply chain of products transparent.
Information if the food comes from regional producers or information if the employees were treated fair justify a higher price.
Whereas this kind of provenance help to address a more selective client base other provenance like the consistent cooling chain is enforced by law and you have to proof it in case of an audit.&lt;/p&gt;
&lt;h2 id=&#34;provenance-and-computer-science&#34;&gt;Provenance and Computer Science&lt;/h2&gt;
&lt;p&gt;I am a computer scientist and not a quality assurance officer, where is the link to Provenance?
We are living in the 21 century, no one is using pen and paper to record the Provenance. Machines of the supply chain, Smart Sensors of the Internet of Things (IoT) are measuring and transmitting the provenance information. The provenance needs then to be stored in a queryable way . Why? Continue reading …
Imagine you are a pharmaceutical company, you keep track of the bottling of the medicine and to which drugstore each bottle was sold. Let’s say you have all the provenance just split into hundred of thousands of files somewhere on an external hard drive in an custom data format, nobody knows about.
After encountering a possible contamination of some of the produced drugs, you need to know where they end up so you can take selective only the affected bottles from the market. You don’t want to waste time and money to instruct one employee to carefully search for the information within the files, which can take multiple days. Having the Provenance stored in an machine understandable format which is queryable, will give you the answer within seconds!
Using an appropriate data storage instead of an external hard drive ensures availability and prevents data loss. Having the provenance in an standardized data format allows the usage of standard software instead of investing time and money to reinvent the wheel.
o
















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;woman code&#34; srcset=&#34;
               /media/what-is-provenance/_hu98f28285dee8e8ed77297b31e519e137_579371_d356e343851855a3ff63352f8136b8be.webp 400w,
               /media/what-is-provenance/_hu98f28285dee8e8ed77297b31e519e137_579371_325aeed69579f16c8eaebdee21ad43cd.webp 760w,
               /media/what-is-provenance/_hu98f28285dee8e8ed77297b31e519e137_579371_0254216d300af0e0e0c499eb586cbd98.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/what-is-provenance/_hu98f28285dee8e8ed77297b31e519e137_579371_d356e343851855a3ff63352f8136b8be.webp&#34;
               width=&#34;760&#34;
               height=&#34;509&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;Overall, standards are important! Imagine your smartphone runs out of battery, you are at a friends place, but your charger doesn’t fit into the electrical socket because every house has another socket type.
Thankfully there is a standard and every electrical socket within a country or larger region is built upon a standard. The same applies for digital standards, you can probably open PDF documents on every of your devices.
For Provenance the World Wide Web Consortium (W3C) released a standard called PROV.&lt;/p&gt;
&lt;h2 id=&#34;provenance-and-social-media&#34;&gt;Provenance and Social Media&lt;/h2&gt;
&lt;p&gt;Speaking of the world wide web, everybody of us produces and/or consumes so much information. Just think about all the blogs and news sites producing content every minute.
Journalists normally investigate carefully and check the reliability of all their sources. But classical media are nowadays only one possible source from which we get information from.
Social media like Twitter or Facebook allows everybody to create content and offer it to a broad variety of people. How do you trust if it is real what is written there?
Political oriented groups can spread so called fake news to influence the opinion of a lot of people. Who actually takes the time to really read an article and check its sources?!
A dramatic headline together with some emotional pictures will get much more response than an article based on facts from boring statistics.
Additionally, such articles also generate a lot of ‘clicks’, meaning that people click on them, visit the site and are seeing some special ads  and the site operator gets a lot of money.
Reason enough for some people to just make up stories, only interested in the money they will get.&lt;/p&gt;
&lt;p&gt;















&lt;figure  &gt;
  &lt;div class=&#34;d-flex justify-content-center&#34;&gt;
    &lt;div class=&#34;w-100&#34; &gt;&lt;img alt=&#34;fake-news&#34; srcset=&#34;
               /media/what-is-provenance/_hu97378ac9bb004fe8e2da724d2e0662a3_200215_8f196b7a86511e27b7b44ce6cc55edb8.webp 400w,
               /media/what-is-provenance/_hu97378ac9bb004fe8e2da724d2e0662a3_200215_a898046016bdc243bf5e8974030b480e.webp 760w,
               /media/what-is-provenance/_hu97378ac9bb004fe8e2da724d2e0662a3_200215_83145448f10a336b475396e24224c1b0.webp 1200w&#34;
               src=&#34;https://sven-lieber.org/media/what-is-provenance/_hu97378ac9bb004fe8e2da724d2e0662a3_200215_8f196b7a86511e27b7b44ce6cc55edb8.webp&#34;
               width=&#34;760&#34;
               height=&#34;519&#34;
               loading=&#34;lazy&#34; data-zoomable /&gt;&lt;/div&gt;
  &lt;/div&gt;&lt;/figure&gt;
&lt;/p&gt;
&lt;p&gt;This whole process of information spreading only partially produces provenance, but the provenance is somehow there. It is implicitly hidden in the timing, the textual similarity or the social connections between the participants of the spreading.
The provenance can partially be reconstructed. That brings us directly to the last part, the research about provenance. As a computer scientists I do actual science.
Research about what kind of provenance can be reconstructed from social media and to which extend is this useful? This is only one example, focusing on social media.
Also general aspects about provenance are not fully explored right now. How to handle provenance, how to secure the provenance records itself against forgery.
Provenance sometimes means total transparency, to which extend is this desirable, at which point the privacy of individuals gets affected?
How can new technologies help to answer some of those questions?&lt;/p&gt;
&lt;p&gt;You hopefully have now an idea what provenance is. What do you think, where else could provenance be useful? Leave a comment!&lt;/p&gt;</description>
    </item>
    
  </channel>
</rss>
