paper-with-me

홈 › Papers

Calling a Spade a Heart: Gaslighting Multimodal Large Language Models via Negation

2025-01-31 · Bin Zhu, Hui yan Qi, Yinxuan Gui, Jingjing Chen, Chong-Wah Ngo, Ee Peng Lim

Multimodal Large Language Models (MLLMs) have exhibited remarkable advancements in integrating different modalities, excelling in complex understanding and generation tasks. Despite their success, MLLMs remain vulnerable to conversational adversarial inputs, particularly negation arguments. This paper systematically evaluates state-of-the-art MLLMs across diverse benchmarks, revealing significant performance drops when negation arguments are introduced to initially correct responses. We show critical vulnerabilities in the reasoning and alignment mechanisms of these models. Proprietary models such as GPT-4o and Claude-3.5-Sonnet demonstrate better resilience compared to open-source counterparts like Qwen2-VL and LLaVA. However, all evaluated MLLMs struggle to maintain logical consistency under negation arguments during conversation. This paper aims to offer valuable insights for improving the robustness of MLLMs against adversarial inputs, contributing to the development of more reliable and trustworthy multimodal AI systems.

📄 PDF Abstract BibTeX arXiv:2501.19017

Code (0)

등록된 구현이 없습니다.

Tasks

Negation

Similar Papers 제목 키워드 기반

Can a large language model be a gaslighter?

2024-10-11 · Wei Li, Luyao Zhu, Yang song, Ruixi Lin 외

Large language models (LLMs) have gained human trust due to their capabilities and helpfulness. However, this in turn may allow LLMs to affect users' mindsets by manipulating language. It is termed as gaslighting, a psyc…

Language ModelingLanguage ModellingLarge Language Modelmodel+1

Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models

2026-04-20 · Ziyao Tang, Pengkun Jiao, Bin Zhu, Huiyan Qi 외 arxiv

Video Large Language Models (Vid-LLMs) have demonstrated remarkable performance in video understanding tasks, yet their robustness under conversational interaction remains largely underexplored. In this paper, we identif…

Visual Grounding

Keep Your Friends Close & Enemies Farther: Debiasing Contrastive Learning with Spatial Priors in 3D Radiology Images

2022-11-16 · Yejia Zhang, Nishchal Sapkota, Pengfei Gu, Yaopeng Peng 외

Understanding of spatial attributes is central to effective 3D radiology image analysis where crop-based learning is the de facto standard. Given an image patch, its core spatial properties (e.g., position & orientation)…

Contrastive LearningRepresentation Learning

Benchmarking Gaslighting Attacks Against Speech Large Language Models

2025-09-24 · Jinyang Wu, Bin Zhu, Xiandong Zou, Qiquan Zhang 외 arxiv

As Speech Large Language Models (Speech LLMs) become increasingly integrated into voice-based applications, ensuring their robustness against manipulative or adversarial input becomes critical. Although prior work has st…

4D Semantic Cardiac Magnetic Resonance Image Synthesis on XCAT Anatomical Model

2020-02-17 · MIDL 2019 7 · Samaneh Abbasi-Sureshjani, Sina Amirrajab, Cristian Lorenz, Juergen Weese 외

We propose a hybrid controllable image generation method to synthesize anatomically meaningful 3D+t labeled Cardiac Magnetic Resonance (CMR) images. Our hybrid method takes the mechanistic 4D eXtended CArdiac Torso (XCAT…

AnatomyGenerative Adversarial NetworkImage GenerationMedical Image Analysis+1