paper-with-me

Papers

Measuring IIA Violations in Similarity Choices with Bayesian Models

2025-08-20 · Hugo Sales Corrêa, Suryanarayana Sankagiri, Daniel Ratton Figueiredo, Matthias Grossglauser arxiv

Similarity choice data occur when humans make choices among alternatives based on their similarity to a target, e.g., in the context of information retrieval and in embedding learning settings. Classical metric-based models of similarity choice assume independence of irrelevant alternatives (IIA), a property that allows for a simpler formulation. While IIA violations have been detected in many discrete choice settings, the similarity choice setting has received scant attention. This is because the target-dependent nature of the choice complicates IIA testing. We propose two statistical methods to test for IIA: a classical goodness-of-fit test and a Bayesian counterpart based on the framework of Posterior Predictive Checks (PPC). This Bayesian approach, our main technical contribution, quantifies the degree of IIA violation beyond its mere significance. We curate two datasets: one with choice sets designed to elicit IIA violations, and another with randomly generated choice sets from the same item universe. Our tests confirmed significant IIA violations on both datasets, and notably, we find a comparable degree of violation between them. Further, we devise a new PPC test for population homogeneity. Results show that the population is indeed homogenous, suggesting that the IIA violations are driven by context effects -- specifically, interactions within the choice sets. These results highlight the need for new similarity choice models that account for such context effects.

📄 PDF Abstract BibTeX arXiv:2508.14615

Code (0)

등록된 구현이 없습니다.

Tasks

Information Retrieval

Similar Papers 제목 키워드 기반

Metrics for Inter-Dataset Similarity with Example Applications in Synthetic Data and Feature Selection Evaluation -- Extended Version

2025-01-16 · Muhammad Rajabinasab, Anton D. Lautrup, Arthur Zimek

Measuring inter-dataset similarity is an important task in machine learning and data mining with various use cases and applications. Existing methods for measuring inter-dataset similarity are computationally expensive, …

feature selection

Latte-Mix: Measuring Sentence Semantic Similarity with Latent Categorical Mixtures

2020-10-21 · H. Bai, L. Tan, K. Xiong, M. Li 외

Measuring sentence semantic similarity using pre-trained language models such as BERT generally yields unsatisfactory zero-shot performance, and one main reason is ineffective token aggregation methods such as mean pooli…

Semantic SimilaritySemantic Textual SimilaritySentenceSTS+1

Measuring Stochastic Rationality

2023-03-14 · Efe A. Ok, Gerelt Tserenjigmid

Our goal is to develop a partial ordering method for comparing stochastic choice functions on the basis of their individual rationality. To this end, we assign to any stochastic choice function a one-parameter class of d…

Fast robustness quantification with variational Bayes

2016-06-23 · Ryan Giordano, Tamara Broderick, Rachael Meager, Jonathan Huggins 외

Bayesian hierarchical models are increasing popular in economics. When using hierarchical models, it is useful not only to calculate posterior expectations, but also to measure the robustness of these expectations to rea…

An unsupervised learning approach to evaluate questionnaire data -- what one can learn from violations of measurement invariance

2023-12-11 · Max Hahn-Klimroth, Paul W. Dierkes, Matthias W. Kleespies

In several branches of the social sciences and humanities, surveys based on standardized questionnaires are a prominent research tool. While there are a variety of ways to analyze the data, some standard procedures have …

Clustering