paper-with-me

홈 › Papers

Same Words, Different Judgments: How Preferences Vary Across Modalities

2026-02-26 · Aaron Broukhim, Nadir Weibel, Eshin Jolly arxiv

Preference-based reinforcement learning (PbRL) is the dominant framework for aligning AI systems to human preferences. However, evaluation protocols for such data were designed for text and have not been validated for speech. We present the first ICC-based, controlled cross-modal study of human and synthetic preference annotations, comparing text and audio evaluations of identical semantic content across 100 prompts. We show that achieving $\textit{good}$ agreement within either modality (ICC(2,$k$) $\approx$ .80) requires $\sim$9 raters. At the same time, modalities show marked differences in how people report preferences: audio raters exhibit narrower decision thresholds, reduced length bias, and more user-oriented evaluation criteria, with near-chance cross-modality agreement. We demonstrate that synthetic ratings can be used to effectively predict inter-rater agreement, thus serving as an early signal for stimulus selection and proxy for human annotations. Together, these findings argue that evaluation protocols for audio preference data require modality-specific design rather than direct adaptation from text.

📄 PDF Abstract BibTeX arXiv:2602.22710

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods

2024-08-05 · Kyle Boerstler, Vijay Keswani, Lok Chan, Jana Schaich Borg 외

Preference elicitation frameworks feature heavily in the research on participatory ethical AI tools and provide a viable mechanism to enquire and incorporate the moral values of various stakeholders. As part of the elici…

Seeing Through Words, Speaking Through Pixels: Deep Representational Alignment Between Vision and Language Models

2025-09-25 · Zoe Wanying He, Sean Trott, Meenakshi Khosla arxiv

Recent studies show that deep vision-only and language-only models--trained on disjoint modalities--nonetheless project their inputs into a partially aligned representational space. Yet we still lack a clear picture of w…

Who Laughs with Whom? Disentangling Influential Factors in Humor Preferences across User Clusters and LLMs

2026-01-06 · Soichiro Murakami, Hidetaka Kamigaito, Hiroya Takamura, Manabu Okumura arxiv

Humor preferences vary widely across individuals and cultures, complicating the evaluation of humor using large language models (LLMs). In this study, we model heterogeneity in humor preferences in Oogiri, a Japanese cre…

Real Images, Worse Judgments: Evaluating Vision-Language Models on Concreteness and Imagery

2026-05-26 · Yifan Jiang, Ruoxi Ning, Sheng Yao, Freda Shi arxiv

Visual inputs are often assumed to improve language understanding in multimodal models. We examine this assumption by asking whether vision-language models (VLMs) can distinguish useful visual evidence from incidental im…

Lexicon Creation for Interpretable NLP Models

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Lexica--words and associated scores--are widely used as simple, interpretable, generalizable language features to predict sentiment, emotions, mental health, and personality traits. Applying different feature importance…

Feature Importance