Downloads provided by UsageCounts
This dataset provides three tables which evaluate the capabilities of GPT 1, GPT 2, GPT 3, and GPT 4 regarding the extraction of tasks from https://doi.org/10.5281/zenodo.7783492 dataset. The performance of the LLMs is measured by calculating a range of similarity metrics: Extracted number of tasks from text vs. extracted number of tasks from Model Semantic Text Similarity: Contextual and Non-Contextual between extracted sets of tasks Semantic Text Similarity: Contextual and Non-Contextual between extracted individual tasks Similarities and Prevalence for length restricted extracted labels Similarities and Prevalence for augmented texts (each text has been paraphrased by 9 different paraphrasing methods)
BPMN, Process Models, ChatGPT, GPT, LLM, Evaluation
BPMN, Process Models, ChatGPT, GPT, LLM, Evaluation
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 1 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
| views | 44 | |
| downloads | 8 |

Views provided by UsageCounts
Downloads provided by UsageCounts