Generating UML class diagrams as conceptual models through large language models

Abstract The breakthrough of Large Language Models (LLMs) has changed how different kinds of complex tasks are approached, including the ones that require a higher level of abstraction and critical thinking together with advanced domain-specific knowledge, like Conceptual Modeling. Several experiments on testing the modeling capabilities of LLMs have already been conducted, but the literature still lacks a structured analysis of how different LLMs and prompting techniques impact the extraction of conceptual models, as UML class diagrams, from textual specifications. In this paper, we present a comprehensive comparison of open-source and closed-source LLMs used in conjunction with the most effective and accessible prompting techniques, on a newly crafted high-quality dataset of case specifications, implementing an automated evaluation on generated UML class diagrams. Finally, we assess how factors like model size or case complexity impact the quality of the generated models and what LLM and what prompting technique to choose for which task. The dataset and the experimental source code are made available through GitHub( https://github.com/IlKaiser/text2uml ).

Authors

Institutions

Publication Details

Journal
Software & Systems Modeling
Published
2026-09-30
DOI
https://doi.org/10.1007/s10270-026-01427-0
Primary Topic
Topic Modeling
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Generating UML class diagrams as conceptual models through large language models

Massimo Mecella, Marco Calamo, Monique Snoeck
Software & Systems Modeling
Topic Modeling
article

Generating UML class diagrams as conceptual models through large language models

Massimo Mecella, Marco Calamo, Monique Snoeck
article en

Abstract

Abstract The breakthrough of Large Language Models (LLMs) has changed how different kinds of complex tasks are approached, including the ones that require a higher level of abstraction and critical thinking together with advanced domain-specific knowledge, like Conceptual Modeling. Several experiments on testing the modeling capabilities of LLMs have already been conducted, but the literature still lacks a structured analysis of how different LLMs and prompting techniques impact the extraction of conceptual models, as UML class diagrams, from textual specifications. In this paper, we present a comprehensive comparison of open-source and closed-source LLMs used in conjunction with the most effective and accessible prompting techniques, on a newly crafted high-quality dataset of case specifications, implementing an automated evaluation on generated UML class diagrams. Finally, we assess how factors like model size or case complexity impact the quality of the generated models and what LLM and what prompting technique to choose for which task. The dataset and the experimental source code are made available through GitHub( https://github.com/IlKaiser/text2uml ).

Software & Systems Modeling
Sapienza University of Rome (IT), KU Leuven (BE)
Quality Education
Openalex Percentile: Top 9%
Topic Modeling
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Generating UML class diagrams as conceptual models through large language models — Massimo Mecella, Marco Calamo, et al. · Software & Systems Modeling (2026) | TGRS Research Map | TGRS