A Survey Study on the State of the Art of Programming Exercise Generation Using Large Language Models

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Conference object 29 Jul 2024Embargo end date: 01 Jan 2024Publisher:IEEEJournal:2024 36th International Conference on Software Engineering Education and Training (CSEE&amp;T)

Authors: Frankford, Eduard; Höhn, Ingo; Sauerwein, Clemens; Breu, Ruth;

doi: 10.1109/cseet62301.2024.10662990 , 10.48550/arxiv.2405.20183

arXiv: 2405.20183

A Survey Study on the State of the Art of Programming Exercise Generation Using Large Language Models

- Summary
- Subjects
- Metrics

Abstract

This paper analyzes Large Language Models (LLMs) with regard to their programming exercise generation capabilities. Through a survey study, we defined the state of the art, extracted their strengths and weaknesses and finally proposed an evaluation matrix, helping researchers and educators to decide which LLM is the best fitting for the programming exercise generation use case. We also found that multiple LLMs are capable of producing useful programming exercises. Nevertheless, there exist challenges like the ease with which LLMs might solve exercises generated by LLMs. This paper contributes to the ongoing discourse on the integration of LLMs in education.

5 pages, 0 figures, CSEE&T 2024

Related Organizations

University of Innsbruck
Austria

Keywords

Software Engineering (cs.SE), FOS: Computer and information sciences, Computer Science - Software Engineering, Artificial Intelligence (cs.AI), Computer Science - Artificial Intelligence

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	1
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

1

Average

Green

Beta

SDGs Suggest

4. Education

Beta

SDGs:

4. Education,

Related to Research communities

Aurora Universities Network

Knowmad Institut