Name: Assessing LLMs for Front-end Software Architecture Knowledge
Keywords: I.2, Software Engineering (cs.SE), FOS: Computer and information sciences, Computer Science - Software Engineering, Artificial Intelligence (cs.AI), Computer Science - Artificial Intelligence, D.2; I.2, D.2

descriptionPublicationkeyboard_double_arrow_right Article , Preprint 27 Apr 2025Embargo end date: 01 Jan 2025Publisher:IEEEJournal:2025 IEEE/ACM International Workshop on Designing Software (Designing)

Authors: Guerra, L. P. Franciscatto; Ernst, N.;

doi: 10.1109/designing66910.2025.00007 , 10.48550/arxiv.2502.19518

arXiv: 2502.19518

Assessing LLMs for Front-end Software Architecture Knowledge

- Summary
- Subjects
- Metrics

Abstract

Large Language Models (LLMs) have demonstrated significant promise in automating software development tasks, yet their capabilities with respect to software design tasks remains largely unclear. This study investigates the capabilities of an LLM in understanding, reproducing, and generating structures within the complex VIPER architecture, a design pattern for iOS applications. We leverage Bloom's taxonomy to develop a comprehensive evaluation framework to assess the LLM's performance across different cognitive domains such as remembering, understanding, applying, analyzing, evaluating, and creating. Experimental results, using ChatGPT 4 Turbo 2024-04-09, reveal that the LLM excelled in higher-order tasks like evaluating and creating, but faced challenges with lower-order tasks requiring precise retrieval of architectural details. These findings highlight both the potential of LLMs to reduce development costs and the barriers to their effective application in real-world software design scenarios. This study proposes a benchmark format for assessing LLM capabilities in software architecture, aiming to contribute toward more robust and accessible AI-driven development tools.

4 pages, 1 figure, to appear in the International Workshop on Designing Software at ICSE 2025

Related Organizations

University of Victoria
Canada

Keywords

I.2, Software Engineering (cs.SE), FOS: Computer and information sciences, Computer Science - Software Engineering, Artificial Intelligence (cs.AI), Computer Science - Artificial Intelligence, D.2; I.2, D.2

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

Average

Green