paper-with-me

홈 › Papers

Private GPTs for LLM-driven testing in software development and machine learning

2025-06-06 · Jakub Jagielski, Markus Abel

In this contribution, we examine the capability of private GPTs to automatically generate executable test code based on requirements. More specifically, we use acceptance criteria as input, formulated as part of epics, or stories, which are typically used in modern development processes. This gives product owners, or business intelligence, respectively, a way to directly produce testable criteria through the use of LLMs. We explore the quality of the so-produced tests in two ways: i) directly by letting the LLM generate code from requirements, ii) through an intermediate step using Gherkin syntax. As a result, it turns out that the two-step procedure yields better results -where we define better in terms of human readability and best coding practices, i.e. lines of code and use of additional libraries typically used in testing. Concretely, we evaluate prompt effectiveness across two scenarios: a simple "Hello World" program and a digit classification model, showing that structured prompts lead to higher-quality test outputs.

📄 PDF Abstract BibTeX arXiv:2506.06509

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large Language Models

2023-03-17 · Tyna Eloundou, Sam Manning, Pamela Mishkin, Daniel Rock

We investigate the potential implications of large language models (LLMs), such as Generative Pre-trained Transformers (GPTs), on the U.S. labor market, focusing on the increased capabilities arising from LLM-powered sof…

LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning

2026-04-17 · Mahir Labib Dihan, Abir Muhtasim arxiv

The rapid proliferation of Large Language Models (LLMs) in software development has made distinguishing AI-generated code from human-written code a critical challenge with implications for academic integrity, code qualit…

Contrastive Learning

The importance of visual modelling languages in generative software engineering

2024-11-27 · Roberto Rossi

Multimodal GPTs represent a watershed in the interplay between Software Engineering and Generative Artificial Intelligence. GPT-4 accepts image and text inputs, rather than simply natural language. We investigate relevan…

From Code Generation to Software Testing: AI Copilot with Context-Based RAG

2025-04-02 · Yuchen Wang, Shangxin Guo, Chee Wei Tan

The rapid pace of large-scale software development places increasing demands on traditional testing methodologies, often leading to bottlenecks in efficiency, accuracy, and coverage. We propose a novel perspective on sof…

ChatbotCode GenerationRAGRetrieval-augmented Generation+1

Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead

2025-03-02 · Yiran Cheng, Hong Jin Kang, Lwin Khin Shar, Chaopeng Dong 외

Fuzz testing is a crucial component of software security assessment, yet its effectiveness heavily relies on valid fuzz drivers and diverse seed inputs. Recent advancements in Large Language Models (LLMs) offer transform…

software testingvalid