paper-with-me

Papers

Light as Deception: GPT-driven Natural Relighting Against Vision-Language Pre-training Models

2025-05-30 · Ying Yang, Jie Zhang, Xiao Lv, Di Lin, Tao Xiang, Qing Guo

While adversarial attacks on vision-and-language pretraining (VLP) models have been explored, generating natural adversarial samples crafted through realistic and semantically meaningful perturbations remains an open challenge. Existing methods, primarily designed for classification tasks, struggle when adapted to VLP models due to their restricted optimization spaces, leading to ineffective attacks or unnatural artifacts. To address this, we propose \textbf{LightD}, a novel framework that generates natural adversarial samples for VLP models via semantically guided relighting. Specifically, LightD leverages ChatGPT to propose context-aware initial lighting parameters and integrates a pretrained relighting model (IC-light) to enable diverse lighting adjustments. LightD expands the optimization space while ensuring perturbations align with scene semantics. Additionally, gradient-based optimization is applied to the reference lighting image to further enhance attack effectiveness while maintaining visual naturalness. The effectiveness and superiority of the proposed LightD have been demonstrated across various VLP models in tasks such as image captioning and visual question answering.

📄 PDF Abstract BibTeX arXiv:2505.24227

Code (0)

등록된 구현이 없습니다.

Tasks

Image CaptioningQuestion AnsweringVisual Question Answering

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Adversarial Relighting Against Face Recognition

2021-08-18 · Qian Zhang, Qing Guo, Ruijun Gao, Felix Juefei-Xu 외

Deep face recognition (FR) has achieved significantly high accuracy on several challenging datasets and fosters successful real-world applications, even showing high robustness to the illumination variation that is usual…

Adversarial AttackFace Recognition

FashionPose: Text to Pose to Relight Image Generation for Personalized Fashion Visualization

2025-07-17 · Chuancheng Shi, Yixiang Chen, Burong Lei, Jichao Chen

Realistic and controllable garment visualization is critical for fashion e-commerce, where users expect personalized previews under diverse poses and lighting conditions. Existing methods often rely on predefined poses, …

Image Generation

Light-A-Video: Training-free Video Relighting via Progressive Light Fusion

2025-02-12 · Yujie Zhou, Jiazi Bu, Pengyang Ling, Pan Zhang 외

Recent advancements in image relighting models, driven by large-scale datasets and pre-trained diffusion models, have enabled the imposition of consistent lighting. However, video relighting still lags, primarily due to …

Image Relighting

SwitchLight: Co-design of Physics-driven Architecture and Pre-training Framework for Human Portrait Relighting

2024-02-29 · CVPR 2024 1 · Hoon Kim, Minje Jang, Wonjun Yoon, Jisoo Lee 외

We introduce a co-designed approach for human portrait relighting that combines a physics-guided architecture with a pre-training framework. Drawing on the Cook-Torrance reflectance model, we have meticulously configured…

DreamLight: Towards Harmonious and Consistent Image Relighting

2025-06-17 · Yong liu, Wenpeng Xiao, Qianqian Wang, Junlin Chen 외

We introduce a model named DreamLight for universal image relighting in this work, which can seamlessly composite subjects into a new background while maintaining aesthetic uniformity in terms of lighting and color tone.…

DisentanglementImage Relighting