paper-with-me

Code Completion

6개 벤치마크 · 논문 279편 · 이 태스크의 논문 보기 →

Benchmarks

SAFIM

결과 15개

CodeXGLUE - PY150

결과 3개

DotPrompts

결과 3개

Defects4J

결과 2개

Rambo Benchmark

결과 2개

Most implemented

Papers

CS-Guard: Benchmarking LLM Guardrails for Code Generation Security

2026-09-09 · Jinyang Li, Mingyu Guo, Hung X. Nguyen arxiv

Large language models (LLMs) have been ex- ploited to generate malware, but the effective- ness of guardrails for code generation secu- rity remains unclear. We introduce CS-Guard, the first benchmark to systematically e…

Text-to-Code GenerationCode TranslationCode Completion

Large Language Models at the Intersection of Software Engineering and Software Security:An Evidence-Centered Structured Survey and Research Agenda

2026-08-21 · Wei Lin, Tao Zhou, Zhaofei Xie, Changgui Hong arxiv

Large Language Models (LLMs) are moving from code completion toward repository-scale agents that retrieve context, edit files, execute tools, and participate in security-sensitive workflows. The evidence for these system…

Vulnerability DetectionCode Completion

Beware What You Autocomplete: Forensic Attribution of Backdoored Code Completions

2026-07-09 · Anjun Gao, Yueyang Quan, Zhuqing Liu, Minghong Fang arxiv

Large language models have enabled powerful code completion systems that assist developers by predicting subsequent lines of code. However, these models remain vulnerable to backdoor attacks, where malicious fine-tuning …

Code Completion

Teaching LLMs a Low-Resource Language: Enhancing Code Completion in Pharo

2026-07-06 · Kilian Kier, Alessandro Giagnorio, Omar AbedelKader, Oleksandr Zaitsev 외 hf

Large Language Models (LLMs) unlocked new possibilities in automated code writing, becoming the backbone of most code completion tools. While LLMs excel in mainstream languages, they often lack support for the so-called …

Code Completion

To Tab or Not to Tab: Measuring Critical Engagement in AI Code Completion Tools Using Behavioral Signals and Attention Checks

2026-06-29 · Jessica Hutchison, Ian Tyler Applebaum, Kenneth Angelikas, Kush Rakesh Patel 외 arxiv

AI code completion tools, such as Github Copilot, provide students with code suggestions to help them write programs. However, recent qualitative studies suggest that students fail to critically evaluate these suggestion…

Code Completion

JAMER: Project-Level Code Framework Dataset and Benchmark on Professional Game Engines

2026-06-18 · Jianwen Sun, Chuanhao Li, Zizhen Li, Yukang Feng 외 arxiv

Current AI-driven game development has made substantial progress in asset generation, gameplay design, and web-based game coding, yet project-level code engineering on professional game engines remains largely unexplored…

Code Completion

전체 279편 보기 →