
This repository contains the code and data used in our study on the impact of Large Language Model (LLM) scale on ontology learning performance. We present a controlled evaluation of 13 models, including dense and Mixture-of-Experts variants from the Qwen3.5 and Qwen3.6 families as well as proprietary GPT release variants, using the OntoLearner retrieval-augmented generation pipeline. All models are evaluated under the same embedding model, retrieval settings, prompts, decoding parameters, datasets, and metrics across term typing, taxonomy discovery, and non-taxonomic relation extraction tasks on four ontologies from biomedical and materials science and engineering domains. The repository includes evaluation scripts, configuration files, cached ontology data, and result outputs supporting reproducible LLM-assisted ontology engineering experiments.
Large Language Models, OntoLearner, Ontology Learning, LLMs4OL
Large Language Models, OntoLearner, Ontology Learning, LLMs4OL
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
