paper-with-me

Papers

VMDT: Decoding the Trustworthiness of Video Foundation Models

2025-11-07 · Yujin Potter, Zhun Wang, Nicholas Crispino, Kyle Montgomery, Alexander Xiong, Ethan Y. Chang, Francesco Pinto, Yuqi Chen, Rahul Gupta, Morteza Ziyadi, Christos Christodoulopoulos, Bo Li, Chenguang Wang, Dawn Song arxiv

As foundation models become more sophisticated, ensuring their trustworthiness becomes increasingly critical; yet, unlike text and image, the video modality still lacks comprehensive trustworthiness benchmarks. We introduce VMDT (Video-Modal DecodingTrust), the first unified platform for evaluating text-to-video (T2V) and video-to-text (V2T) models across five key trustworthiness dimensions: safety, hallucination, fairness, privacy, and adversarial robustness. Through our extensive evaluation of 7 T2V models and 19 V2T models using VMDT, we uncover several significant insights. For instance, all open-source T2V models evaluated fail to recognize harmful queries and often generate harmful videos, while exhibiting higher levels of unfairness compared to image modality models. In V2T models, unfairness and privacy risks rise with scale, whereas hallucination and adversarial robustness improve -- though overall performance remains low. Uniquely, safety shows no correlation with model size, implying that factors other than scale govern current safety levels. Our findings highlight the urgent need for developing more robust and trustworthy video foundation models, and VMDT provides a systematic framework for measuring and tracking progress toward this goal. The code is available at https://sunblaze-ucb.github.io/VMDT-page/.

📄 PDF Abstract BibTeX arXiv:2511.05682

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial Robustness

Similar Papers 제목 키워드 기반

MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models

2025-03-19 · Chejian Xu, Jiawei Zhang, Zhaorun Chen, Chulin Xie 외

Multimodal foundation models (MMFMs) play a crucial role in various applications, including autonomous driving, healthcare, and virtual assistants. However, several studies have revealed vulnerabilities in these models, …

Adversarial RobustnessAutonomous DrivingFairnessHallucination+1

TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models

2026-04-15 · Yichuan Mo, Yukun Jiang, Yanbo Shi, Mingjie Li 외 arxiv

The rapid development of Language Diffusion Models (LDMs) challenges the dominant position of auto-regressive competitors in language processing. However, their flexible, any-order decoding strategies not only enable fas…

Developing trustworthy AI applications with foundation models

2024-05-08 · Michael Mock, Sebastian Schmidt, Felix Müller, Rebekka Görge 외

The trustworthiness of AI applications has been the subject of recent research and is also addressed in the EU's recently adopted AI Regulation. The currently emerging foundation models in the field of text, speech and i…

Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression

2024-03-18 · Junyuan Hong, Jinhao Duan, Chenhui Zhang, Zhangheng Li 외

Compressing high-capability Large Language Models (LLMs) has emerged as a favored strategy for resource-efficient inferences. While state-of-the-art (SoTA) compression methods boast impressive advancements in preserving …

EthicsFairnessQuantization

A Survey on Trustworthiness in Foundation Models for Medical Image Analysis

2024-07-03 · Congzhen Shi, Ryan Rezai, Jiaxi Yang, Qi Dou 외

The rapid advancement of foundation models in medical imaging represents a significant leap toward enhancing diagnostic accuracy and personalized treatment. However, the deployment of foundation models in healthcare nece…

DiagnosticFairnessMedical Image AnalysisMedical Report Generation+1