paper-with-me

홈 › Papers

Large Language Models (LLMs) for Source Code Analysis: applications, models and datasets

2025-03-21 · Hamed Jelodar, Mohammad Meymani, Roozbeh Razavi-Far

Large language models (LLMs) and transformer-based architectures are increasingly utilized for source code analysis. As software systems grow in complexity, integrating LLMs into code analysis workflows becomes essential for enhancing efficiency, accuracy, and automation. This paper explores the role of LLMs for different code analysis tasks, focusing on three key aspects: 1) what they can analyze and their applications, 2) what models are used and 3) what datasets are used, and the challenges they face. Regarding the goal of this research, we investigate scholarly articles that explore the use of LLMs for source code analysis to uncover research developments, current trends, and the intellectual structure of this emerging field. Additionally, we summarize limitations and highlight essential tools, datasets, and key challenges, which could be valuable for future work.

📄 PDF Abstract BibTeX arXiv:2503.17502

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Similar Papers 제목 키워드 기반

Analysis on LLMs Performance for Code Summarization

2024-12-22 · Md. Ahnaf Akib, Md. Muktadir Mazumder, Salman Ahsan

Code summarization aims to generate concise natural language descriptions for source code. Deep learning has been used more and more recently in software engineering, particularly for tasks like code creation and summari…

Code Summarization

Semantic Source Code Segmentation using Small and Large Language Models

2025-07-11 · Abdelhalim Dahou, Ansgar Scherp, Sebastian Kurten, Brigitte Mathiak 외 arxiv

Source code segmentation, dividing code into functionally coherent segments, is crucial for knowledge retrieval and maintenance in software development. While enabling efficient navigation and comprehension of large code…

Harnessing the Power of LLMs in Source Code Vulnerability Detection

2024-08-07 · Andrew A Mahyari

Software vulnerabilities, caused by unintentional flaws in source code, are a primary root cause of cyberattacks. Static analysis of source code has been widely used to detect these unintentional defects introduced by so…

Vulnerability Detection

Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis

2023-04-10 · Wenhao Zhu, Hongyi Liu, Qingxiu Dong, Jingjing Xu 외

Large language models (LLMs) have demonstrated remarkable potential in handling multilingual machine translation (MMT). In this paper, we systematically investigate the advantages and challenges of LLMs for MMT by answer…

Machine TranslationTranslation

Analysis of ChatGPT on Source Code

2023-06-01 · Ahmed R. Sadik, Antonello Ceravola, Frank Joublin, Jibesh Patra

This paper explores the use of Large Language Models (LLMs) and in particular ChatGPT in programming, source code analysis, and code generation. LLMs and ChatGPT are built using machine learning and artificial intelligen…

Code Generation