Inicio  /  Applied Sciences  /  Vol: 12 Par: 18 (2022)  /  Artículo
ARTÍCULO
TITULO

A Linked Data Application for Harmonizing Heterogeneous Biomedical Information

Nicola Capuano    
Pasquale Foggia    
Luca Greco and Pierluigi Ritrovato    

Resumen

In the biomedical field, there is an ever-increasing number of large, fragmented, and isolated data sources stored in databases and ontologies that use heterogeneous formats and poorly integrated schemes. Researchers and healthcare professionals find it extremely difficult to master this huge amount of data and extract relevant information. In this work, we propose a linked data approach, based on multilayer networks and semantic Web standards, capable of integrating and harmonizing several biomedical datasets with different schemas and semi-structured data through a multi-model database providing polyglot persistence. The domain chosen concerns the analysis and aggregation of available data on neuroendocrine neoplasms (NENs), a relatively rare type of neoplasm. Integrated information includes twelve public datasets available in heterogeneous schemas and formats including RDF, CSV, TSV, SQL, OWL, and OBO. The proposed integrated model consists of six interconnected layers representing, respectively, information on the disease, the related phenotypic alterations, the affected genes, the related biological processes, molecular functions, the involved human tissues, and drugs and compounds that show documented interactions with them. The defined scheme extends an existing three-layer model covering a subset of the mentioned aspects. A client?server application was also developed to browse and search for information on the integrated model. The main challenges of this work concern the complexity of the biomedical domain, the syntactic and semantic heterogeneity of the datasets, and the organization of the integrated model. Unlike related works, multilayer networks have been adopted to organize the model in a manageable and stratified structure, without the need to change the original datasets but by transforming their data ?on the fly? to respond to user requests.

 Artículos similares

       
 
J. Javier Samper-Zapater, Julián Gutiérrez-Moret, Jose Macario Rocha, Juan José Martinez-Durá and Vicente R. Tomás    
The significance of Linked Open Data datasets for traffic information extends beyond just including open traffic data. It incorporates links to other relevant thematic datasets available on the web. This enables federated queries across different data pl... ver más
Revista: Information

 
Matharit Namsai, Butsawan Bidorn, Ruetaitip Mama and Warit Charoenlerkthawin    
The construction of large dams in the upper tributary basin of the Chao Phraya River (CPR) has been linked to a significant decrease in sediment load in the CPR system, estimated between 75?85%. This study, utilizing historical and recent river flow and ... ver más
Revista: Water

 
Rui Yuan, Ruiyang Xu, Hezhenjia Zhang, Yutao Hua, Hongsheng Zhang, Xiaojing Zhong and Shenliang Chen    
This study presents an in-depth analysis of the dynamic beach landscapes of Hainan Island, which is located at the southernmost tip of China. Home to over a hundred natural and predominantly sandy beaches, Hainan Island confronts significant challenges p... ver más
Revista: Water

 
Moritz Müller, Ambre Dupuis, Tobias Zeulner, Ignacio Vazquez, Johann Hagerer and Peter A. Gloor    
Well-being is one of the pillars of positive psychology, which is known to have positive effects not only on the personal and professional lives of individuals but also on teams and organizations. Understanding and promoting individual well-being is esse... ver más
Revista: Applied Sciences

 
Giovanni Briguglio and Vincenzo Crupi    
The increasingly stringent requirements?in terms of limiting pollutants and the constant need to make maritime transport safer?generated the necessity to foresee different solutions that are original. According to the European Maritime Safety Agency, the... ver más