paper-with-me

홈 › Papers

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

2026-05-13 · Mohammad Reza Mousavi arxiv

Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether these tasks are performed consistently with respect to the semantics of the artefacts being handled. This question is particularly under-researched concerning architectural design specification. In this paper, we address this question for High-Level Message Sequence Charts (HMSCs). These are visual models with a rigorous formal semantics that have been used for various purposes, including as a foundation for Sequence Diagrams in the Unified Modelling Language (UML). We examine whether LLMs "understand" the semantics of HMSCs by examining three LLMs (Gemini-3, GPT-5.4, and Qwen-3.6) on how they perform 129 semantic tasks ranging from querying basic semantic constructs in HMSCs (i.e., events and their ordering) to semantic-preserving abstractions and compositions, and calculating the set of traces and trace-equivalent labelled transition systems. The results show that LLMs only have a modest understanding of the formal semantics of HMSCs (ca. 52% overall accuracy), with great variability across different semantic concepts: while LLMs seem to understand the basic semantic concepts of MSCs (ca. 88% accuracy), they struggle with semantic reasoning in tasks involving abstraction and composition (ca. 36% accuracy) and traces and LTSs (ca. 42% accuracy). In particular, all three LLMs struggle with the notions of co-region and explicit causal dependencies and never employed them in semantic-preserving transformations.

📄 PDF Abstract BibTeX arXiv:2605.13773

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MeLT: Message-Level Transformer with Masked Document Representations as Pre-Training for Stance Detection

2021-09-16 · Findings (EMNLP) 2021 11 · Matthew Matero, Nikita Soni, Niranjan Balasubramanian, H. Andrew Schwartz

Much of natural language processing is focused on leveraging large capacity language models, typically trained over single messages with a task of predicting one or more tokens. However, modeling human language at higher…

AttributeLanguage ModelingLanguage ModellingMasked Language Modeling+1

Using Large Language Models to Enhance Programming Error Messages

2022-10-20 · Juho Leinonen, Arto Hellas, Sami Sarsa, Brent Reeves 외

A key part of learning to program is learning to understand programming error messages. They can be hard to interpret and identifying the cause of errors can be time-consuming. One factor in this challenge is that the me…

Accuracy of a Large Language Model in Distinguishing Anti- And Pro-vaccination Messages on Social Media: The Case of Human Papillomavirus Vaccination

2024-04-10 · Soojong Kim, Kwanho Kim, Claire Wonjeong Jo

Objective. Vaccination has engendered a spectrum of public opinions, with social media acting as a crucial platform for health-related discussions. The emergence of artificial intelligence technologies, such as large lan…

Language ModelingLanguage ModellingLarge Language ModelSentiment Analysis

LogLLaMA: Transformer-based log anomaly detection with LLaMA

2025-03-19 · Zhuoyi Yang, Ian G. Harris

Log anomaly detection refers to the task that distinguishes the anomalous log messages from normal log messages. Transformer-based large language models (LLMs) are becoming popular for log anomaly detection because of th…

Anomaly DetectionReinforcement Learning (RL)

Multi-resolution Interpretation and Diagnostics Tool for Natural Language Classifiers

2023-03-06 · Peyman Jalali, Nengfeng Zhou, Yufei Yu

Developing explainability methods for Natural Language Processing (NLP) models is a challenging task, for two main reasons. First, the high dimensionality of the data (large number of tokens) results in low coverage and …