paper-with-me

Papers

MultiTrust: A Comprehensive Benchmark Towards Trustworthy Multimodal Large Language Models

2024-06-11 · Yichi Zhang, Yao Huang, Yitong Sun, Chang Liu, Zhe Zhao, Zhengwei Fang, Yifan Wang, Huanran Chen, Xiao Yang, Xingxing Wei, Hang Su, Yinpeng Dong, Jun Zhu

Despite the superior capabilities of Multimodal Large Language Models (MLLMs) across diverse tasks, they still face significant trustworthiness challenges. Yet, current literature on the assessment of trustworthy MLLMs remains limited, lacking a holistic evaluation to offer thorough insights into future improvements. In this work, we establish MultiTrust, the first comprehensive and unified benchmark on the trustworthiness of MLLMs across five primary aspects: truthfulness, safety, robustness, fairness, and privacy. Our benchmark employs a rigorous evaluation strategy that addresses both multimodal risks and cross-modal impacts, encompassing 32 diverse tasks with self-curated datasets. Extensive experiments with 21 modern MLLMs reveal some previously unexplored trustworthiness issues and risks, highlighting the complexities introduced by the multimodality and underscoring the necessity for advanced methodologies to enhance their reliability. For instance, typical proprietary models still struggle with the perception of visually confusing images and are vulnerable to multimodal jailbreaking and adversarial attacks; MLLMs are more inclined to disclose privacy in text and reveal ideological and cultural biases even when paired with irrelevant images in inference, indicating that the multimodality amplifies the internal risks from base LLMs. Additionally, we release a scalable toolbox for standardized trustworthiness research, aiming to facilitate future advancements in this important field. Code and resources are publicly available at: https://multi-trust.github.io/.

📄 PDF Abstract BibTeX arXiv:2406.07057

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingFairness

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Unveiling Trust in Multimodal Large Language Models: Evaluation, Analysis, and Mitigation

2025-08-21 · Yichi Zhang, Yao Huang, Yifan Wang, Yitong Sun 외 arxiv

The trustworthiness of Multimodal Large Language Models (MLLMs) remains an intense concern despite the significant progress in their capabilities. Existing evaluation and mitigation approaches often focus on narrow aspec…

LUMA: A Benchmark Dataset for Learning from Uncertain and Multimodal Data

2024-06-14 · Grigor Bezirganyan, Sana Sellami, Laure Berti-ÉQuille, Sébastien Fournier

Multimodal Deep Learning enhances decision-making by integrating diverse information sources, such as texts, images, audio, and videos. To develop trustworthy multimodal approaches, it is essential to understand how unce…

BenchmarkingDecision MakingDiversityLanguage Modelling+4

MoHoBench: Assessing Honesty of Multimodal Large Language Models via Unanswerable Visual Questions

2025-07-29 · Yanxu Zhu, Shitong Duan, Xiangxu Zhang, Jitao Sang 외 arxiv

Recently Multimodal Large Language Models (MLLMs) have achieved considerable advancements in vision-language tasks, yet produce potentially harmful or untrustworthy content. Despite substantial work investigating the tru…

Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models

2024-08-18 · Kening Zheng, Junkai Chen, Yibo Yan, Xin Zou 외

Hallucination issues continue to affect multimodal large language models (MLLMs), with existing research mainly addressing object-level or attribute-level hallucinations, neglecting the more complex relation hallucinatio…

AttributeHallucinationHallucination EvaluationRelation

MedMKEB: A Comprehensive Knowledge Editing Benchmark for Medical Multimodal Large Language Models

2025-08-07 · Dexuan Xu, Jieyi Wang, Zhongyan Chai, Yongzhi Cao 외 arxiv

Recent advances in multimodal large language models (MLLMs) have significantly improved medical AI, enabling it to unify the understanding of visual and textual information. However, as medical knowledge continues to evo…

Adversarial Robustnessknowledge editing