paper-with-me

홈 › Papers

Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation

2025-12-05 · Ju-Young Kim, Ji-Hong Park, Myeongjun Kim, Gun-Woo Kim arxiv

Smart farming has emerged as a key technology for advancing modern agriculture through automation and intelligent control. However, systems relying on RGB cameras for perception and robotic manipulators for control, common in smart farming, are vulnerable to photometric perturbations such as hue, illumination, and noise changes, which can cause malfunction under adversarial attacks. To address this issue, we propose an explainable adversarial-robust Vision-Language-Action model based on the OpenVLA-OFT framework. The model integrates an Evidence-3 module that detects photometric perturbations and generates natural language explanations of their causes and effects. Experiments show that the proposed model reduces Current Action L1 loss by 21.7% and Next Actions L1 loss by 18.4% compared to the baseline, demonstrating improved action prediction accuracy and explainability under adversarial conditions.

📄 PDF Abstract BibTeX arXiv:2512.11865

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TactEx: An Explainable Multimodal Robotic Interaction Framework for Human-Like Touch and Hardness Estimation

2026-02-21 · Felix Verstraete, Lan Wei, Wen Fan, Dandan Zhang arxiv

Accurate perception of object hardness is essential for safe and dexterous contact-rich robotic manipulation. Here, we present TactEx, an explainable multimodal robotic interaction framework that unifies vision, touch, a…

Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics

2024-11-18 · Taowen Wang, Cheng Han, James Chenhao Liang, Wenhao Yang 외

Recently in robotics, Vision-Language-Action (VLA) models have emerged as a transformative approach, enabling robots to execute complex tasks by integrating visual and linguistic inputs within an end-to-end learning fram…

Vision-Language-Action

Partially Observable Adversarial Patch Attacks on Vision-Language-Action Models in Robotics

2026-06-02 · Xiaofei Wang, Mingliang Han, Tianyu Hao, Yi Yang 외 arxiv

Vision-language-action (VLA) models are gaining attention in robotics, yet their robustness to adversarial attacks remains largely unexplored. Existing work shows that adversarial patches can mislead VLA-based robots but…

Adversarial Attacks on Robotic Vision Language Action Models

2025-06-03 · Eliot Krzysztof Jones, Alexander Robey, Andy Zou, Zachary Ravichandran 외

The emergence of vision-language-action models (VLAs) for end-to-end control is reshaping the field of robotics by enabling the fusion of multimodal sensory inputs at the billion-parameter scale. The capabilities of VLAs…

Vision-Language-Action

FreezeVLA: Action-Freezing Attacks against Vision-Language-Action Models

2025-09-24 · Xin Wang, Jie Li, Zejia Weng, Yixu Wang 외 arxiv

Vision-Language-Action (VLA) models are driving rapid progress in robotics by enabling agents to interpret multimodal inputs and execute complex, long-horizon tasks. However, their safety and robustness against adversari…