Semantic Modelling of Organizational Knowledge as a Basis for Enterprise Data Governance 4.0 -- Application to a Unified Clinical Data Model
Miguel AP Oliveira, Stephane Manara, Bruno Mol\'e, Thomas Muller,, Aur\'elien Guillouche, Lysann Hesske, Bruce Jordan, Gilles Hubert, Chinmay, Kulkarni, Pralipta Jagdev, Cedric R. Berger

TL;DR
This paper presents a semantic web-based, metadata-driven framework for agile, semi-automated enterprise data governance, demonstrated through integrating 25 years of clinical data to improve data quality and management.
Contribution
It introduces a novel, cost-efficient Data Governance 4.0 framework utilizing knowledge graphs and ontologies for scalable, flexible clinical data management.
Findings
Successful integration of 25 years of clinical data at enterprise scale
Implementation of a semantic web-based metadata model for governance
Enhanced agility and automation in data management processes
Abstract
Individuals and organizations cope with an always-growing amount of data, which is heterogeneous in its contents and formats. An adequate data management process yielding data quality and control over its lifecycle is a prerequisite to getting value out of this data and minimizing inherent risks related to multiple usages. Common data governance frameworks rely on people, policies, and processes that fall short of the overwhelming complexity of data. Yet, harnessing this complexity is necessary to achieve high-quality standards. The latter will condition any downstream data usage outcome, including generative artificial intelligence trained on this data. In this paper, we report our concrete experience establishing a simple, cost-efficient framework that enables metadata-driven, agile and (semi-)automated data governance (i.e. Data Governance 4.0). We explain how we implement and use…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsBiomedical Text Mining and Ontologies · Research Data Management Practices · Scientific Computing and Data Management
MethodsOntology
