paper-with-me

홈 › Papers

AGACCI : Affiliated Grading Agents for Criteria-Centric Interface in Educational Coding Contexts

2025-07-07 · Kwangsuk Park, Jiwoong Yang arxiv

Recent advances in AI-assisted education have encouraged the integration of vision-language models (VLMs) into academic assessment, particularly for tasks that require both quantitative and qualitative evaluation. However, existing VLM based approaches struggle with complex educational artifacts, such as programming tasks with executable components and measurable outputs, that require structured reasoning and alignment with clearly defined evaluation criteria. We introduce AGACCI, a multi-agent system that distributes specialized evaluation roles across collaborative agents to improve accuracy, interpretability, and consistency in code-oriented assessment. To evaluate the framework, we collected 360 graduate-level code-based assignments from 60 participants, each annotated by domain experts with binary rubric scores and qualitative feedback. Experimental results demonstrate that AGACCI outperforms a single GPT-based baseline in terms of rubric and feedback accuracy, relevance, consistency, and coherence, while preserving the instructional intent and evaluative depth of expert assessments. Although performance varies across task types, AGACCI highlights the potential of multi-agent systems for scalable and context-aware educational evaluation.

📄 PDF Abstract BibTeX arXiv:2507.05321

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PathMem: Toward Cognition-Aligned Memory Transformation for Pathology MLLMs

2026-03-10 · Jinyue Li, Yuci Liang, Qiankun Li, Xinheng Lyu 외 arxiv

Computational pathology demands both visual pattern recognition and dynamic integration of structured domain knowledge, including taxonomy, grading criteria, and clinical evidence. In practice, diagnostic reasoning requi…

KIEGLFN: A unified acne grading framework on face images

2022-06-01 · Computer Methods and Programs in Biomedicine 2022 6 · Yi Lin, Jingchi Jiang, Zhaoyang Ma, Dongxin Chen 외

Grading the severity level is an extremely important procedure for correct diagnoses and personalized treatment schemes for acne. However, the acne grading criteria are not unified in the medical field. This work aims to…

Acne Severity Grading

An Acne Grading Framework on Face Images via Skin Attention and SFNet

2022-01-14 · IEEE International Conference on Bioinformatics and Biomedicine (BIBM) 2022 1 · Yi Lin, Yi Guan, Zhaoyang Ma, Haiyan You 외

Severity level grading is a vitally important step to make correct diagnoses and personalized treatment schemes for acne, which is mainly carried out in two ways: criterion-based lesion counting and experience-based glob…

Acne Severity GradingDiagnostic

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

2026-02-01 · Xueyi Li, Zhuoneng Zhou, Zitao Liu, Yongdong Wu arxiv

Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world educational environments, significantly boosting assessment efficiency and scalability.…

Adversarial Attack

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

2026-05-27 · Melih Catal, Alex Wolf, Tiago Ferreiro Matos, Pooja Rani 외 arxiv

Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many frontier LLMs are affiliated with specific providers. This raises the quest…

Code Generation