<script type="text/javascript">
<!--
document.write('<div id="oa_widget"></div>');
document.write('<script type="text/javascript" src="https://www.openaire.eu/index.php?option=com_openaire&view=widget&format=raw&projectId=undefined&type=result"></script>');
-->
</script>

COPY SCRIPT

For further information contact us at helpdesk@openaire.eu

Large language models for intelligent RDF knowledge graph construction: results from medical ontology mapping

Name: Large language models for intelligent RDF knowledge graph construction: results from medical ontology mapping
Keywords: LLM, knowledge graph, Artificial Intelligence, Electronic computers. Computer science, health data, ontology, QA75.5-76.95, SNOMED CT, RDF

descriptionPublicationkeyboard_double_arrow_right Article , Other literature type 25 Apr 2025Publisher:Frontiers Media SAJournal:Frontiers in Artificial Intelligence, volume 8 (eissn: 2624-8212,

Authors: Mavridis, Apostolos; Tegos, Stergios; Anastasiou, Christos; Papoutsoglou, Maria; Meditskos, Georgios;

doi: 10.3389/frai.2025.1546179

pmid: 40352975

pmc: PMC12061982

Large language models for intelligent RDF knowledge graph construction: results from medical ontology mapping

- Summary
- Subjects
- Metrics

Abstract

The exponential growth of digital data, particularly in specialized domains like healthcare, necessitates advanced knowledge representation and integration techniques. RDF knowledge graphs offer a powerful solution, yet their creation and maintenance, especially for complex medical ontologies like Systematized Nomenclature of Medicine - Clinical Terms (SNOMED CT), remain challenging. Traditional methods often struggle with the scale, heterogeneity, and semantic complexity of medical data. This paper introduces a methodology leveraging the contextual understanding and reasoning capabilities of Large Language Models (LLMs) to automate and enhance medical ontology mapping for Resource Description Framework (RDF) knowledge graph construction. We conduct a comprehensive comparative analysis of six systems–GPT-4o, Claude 3.5 Sonnet v2, Gemini 1.5 Pro, Llama 3.3 70B, DeepSeek R1, and BERTMap—using a novel evaluation framework that combines quantitative metrics (precision, recall, and F1-score) with qualitative assessments of semantic accuracy. Our approach integrates a data preprocessing pipeline with an LLM-powered semantic mapping engine, utilizing BioBERT embeddings and ChromaDB vector database for efficient concept retrieval. Experimental results on a dataset of 108 medical terms demonstrate the superior performance of modern LLMs, particularly GPT-4o, achieving a precision of 93.75% and an F1-score of 96.26%. These findings highlight the potential of LLMs in bridging the gap between structured medical data and semantic knowledge representation, toward more accurate and interoperable medical knowledge graphs.

Related Organizations

Aristotle University of Thessaloniki
Greece

Keywords

LLM, knowledge graph, Artificial Intelligence, Electronic computers. Computer science, health data, ontology, QA75.5-76.95, SNOMED CT, RDF

Impact byBIP!

	citations This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

Average

Green

gold

Funded by

EC| ENCRYPT

Related to Research communities

Knowmad Institut