paper-with-me

홈 › Papers

AstraAI: LLMs, Retrieval, and AST-Guided Assistance for HPC Codebases

2026-03-28 · Mahesh Natarajan, Xiaoye Li, Weiqun Zhang arxiv

We present AstraAI, a command-line interface (CLI) coding framework for high-performance computing (HPC) software development. AstraAI operates directly within a Linux terminal and integrates large language models (LLMs) with Retrieval-Augmented Generation (RAG) and Abstract Syntax Tree (AST)-based structural analysis to enable context-aware code generation for complex scientific codebases. The central idea is to construct a high-fidelity prompt that is passed to the LLM for inference. This prompt augments the user request with relevant code snippets retrieved from the underlying framework codebase via RAG and structural context extracted from AST analysis, providing the model with precise information about relevant functions, data structures, and overall code organization. The framework is designed to perform scoped modifications to source code while preserving structural consistency with the surrounding code. AstraAI supports both locally hosted models from Hugging Face and API-based frontier models accessible via the American Science Cloud, enabling flexible deployment across HPC environments. The system generates code that aligns with existing project structures and programming patterns. We demonstrate AstraAI on representative HPC code generation tasks within AMReX, a DOE-supported HPC software infrastructure for exascale applications.

📄 PDF Abstract BibTeX arXiv:2603.27423

Code (0)

등록된 구현이 없습니다.

Tasks

Code Generation

Similar Papers 제목 키워드 기반

Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation

2025-10-30 · Musfiqur Rahman, SayedHassan Khatoonabadi, Emad Shihab arxiv

Large language models (LLMs) have demonstrated strong performance on function-level code generation benchmarks, yet real-world software development increasingly demands class-level implementations that integrate multiple…

Class-level Code Generation

Knowledge Graph Based Repository-Level Code Generation

2025-05-20 · Mihir Athale, Vishal Vaddina

Recent advancements in Large Language Models (LLMs) have transformed code generation from natural language queries. However, despite their extensive knowledge and ability to produce high-quality code, LLMs often struggle…

Code GenerationCode SearchNatural Language QueriesRetrieval

HCAG: Hierarchical Abstraction and Retrieval-Augmented Generation on Theoretical Repositories with LLMs

2026-03-19 · Yusen Wu, Xiaotie Deng arxiv

Existing Retrieval-Augmented Generation (RAG) methods for code struggle to capture the high-level architectural patterns and cross-file dependencies inherent in complex, theory-driven codebases, such as those in algorith…

Code Generation

CodeAssistBench (CAB): Dataset & Benchmarking for Multi-turn Chat-Based Code Assistance

2025-07-14 · Myeongsoo Kim, Shweta Garg, Baishakhi Ray, Varun Kumar 외

Programming assistants powered by large language models have transformed software development, yet most benchmarks focus narrowly on code generation tasks. Recent efforts like InfiBench and StackEval attempt to address t…

BenchmarkingCode Generation

Meta-RAG on Large Codebases Using Code Summarization

2025-08-04 · Vali Tawosi, Salwa Alamir, Xiaomo Liu, Manuela Veloso arxiv

Large Language Model (LLM) systems have been at the forefront of applied Artificial Intelligence (AI) research in a multitude of domains. One such domain is software development, where researchers have pushed the automat…

Information Retrieval