paper-with-me

홈 › Papers

Evaluation of Audio Language Models for Fairness, Safety, and Security

2026-02-25 · Ranya Aloufi, Srishti Gupta, Soumya Shaw, Battista Biggio, Lea Schönherr arxiv

Audio large language models (ALLMs) have recently advanced spoken interaction by integrating speech processing with large language models. However, existing evaluations of fairness, safety, and security (FSS) remain fragmented, largely because ALLMs differ fundamentally in how acoustic information is represented and where semantic reasoning occurs. Differences that are rarely made explicit. As a result, evaluations often conflate structurally distinct systems, obscuring the relationship between model design and observed FSS behavior. In this work, we introduce a structural taxonomy (system-level and representational) of ALLMs that categorizes systems along two axes: the form of audio input representation (e.g., discrete vs. continuous) and the locus of semantic reasoning (e.g., cascaded, multimodal, or audio-native). Building on the taxonomy, we propose a unified evaluation framework that assesses semantic invariance under paralinguistic variation, refusal and toxicity behavior under unsafe prompts, and robustness to adversarial audio perturbations. We apply this framework to two representative systems and observe systematic differences in refusal rates, attack success, and toxicity between audio and text inputs. Our findings demonstrate that FSS behavior is tightly coupled to how acoustic information is integrated into semantic reasoning, underscoring the need for structure-aware evaluation of audio language models.

📄 PDF Abstract BibTeX arXiv:2603.13262

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

aiXamine: Simplified LLM Safety and Security

2025-04-21 · Fatih Deniz, Dorde Popovic, Yazan Boshmaf, Euisuh Jeong 외

Evaluating Large Language Models (LLMs) for safety and security remains a complex task, often requiring users to navigate a fragmented landscape of ad hoc benchmarks, datasets, metrics, and reporting formats. To address …

2kAdversarial RobustnessFairnessHallucination+2

AHELM: A Holistic Evaluation of Audio-Language Models

2025-08-29 · Tony Lee, Haoqin Tu, Chi Heem Wong, Zijun Wang 외 arxiv

Evaluations of audio-language models (ALMs) -- multimodal models that take interleaved audio and text as input and output text -- are hindered by the lack of standardized benchmarks; most benchmarks measure only one or t…

Question Answering

SEA: Low-Resource Safety Alignment for Multimodal Large Language Models via Synthetic Embeddings

2025-02-18 · Weikai Lu, Hao Peng, Huiping Zhuang, Cen Chen 외

Multimodal Large Language Models (MLLMs) have serious security vulnerabilities.While safety alignment using multimodal datasets consisting of text and data of additional modalities can effectively enhance MLLM's security…

GPUSafety Alignment

AudioTrust: Benchmarking the Multifaceted Trustworthiness of Audio Large Language Models

2025-05-22 · Kai Li, Can Shen, Yile Liu, Jirui Han 외

The rapid advancement and expanding applications of Audio Large Language Models (ALLMs) demand a rigorous understanding of their trustworthiness. However, systematic research on evaluating these models, particularly conc…

BenchmarkingFairnessHallucination

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

2026-06-25 · Beatrice Savoldi, Sara Papi, Wafa Aissa, Matteo Negri 외 hf

Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and under naturalistic conditions remain understudied. We survey safety r…