paper-with-me

홈 › Papers

LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users

2024-06-25 · Elinor Poole-Dayan, Deb Roy, Jad Kabbara

While state-of-the-art Large Language Models (LLMs) have shown impressive performance on many tasks, there has been extensive research on undesirable model behavior such as hallucinations and bias. In this work, we investigate how the quality of LLM responses changes in terms of information accuracy, truthfulness, and refusals depending on three user traits: English proficiency, education level, and country of origin. We present extensive experimentation on three state-of-the-art LLMs and two different datasets targeting truthfulness and factuality. Our findings suggest that undesirable behaviors in state-of-the-art LLMs occur disproportionately more for users with lower English proficiency, of lower education status, and originating from outside the US, rendering these models unreliable sources of information towards their most vulnerable users.

📄 PDF Abstract BibTeX arXiv:2406.17737

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Novel Bifurcation Method for Observation Perturbation Attacks on Reinforcement Learning Agents: Load Altering Attacks on a Cyber Physical Power System

2024-07-06 · Kiernan Broda-Milian, Ranwa Al-Mallah, Hanane Dagdougui

Components of cyber physical systems, which affect real-world processes, are often exposed to the internet. Replacing conventional control methods with Deep Reinforcement Learning (DRL) in energy systems is an active are…

continuous-controlContinuous ControlDeep Reinforcement LearningTime Series Analysis

On Configurable Defense against Adversarial Example Attacks

2018-12-06 · Bo Luo, Min Li, Yu Li, Qiang Xu

Machine learning systems based on deep neural networks (DNNs) have gained mainstream adoption in many applications. Recently, however, DNNs are shown to be vulnerable to adversarial example attacks with slight perturbati…

Equity Impacts of Public Transit Network Redesign with Shared Autonomous Mobility Services

2025-01-03 · Max T. M. Ng, Meredith Raymer, Hani S. Mahmassani, Omer Verbas 외

This study examines the equity impacts of integrating shared autonomous mobility services (SAMS) into transit system redesign. Using the Greater Chicago area as a case study, we compare two optimization objectives in mul…

Disparate Impact on Group Accuracy of Linearization for Private Inference

2024-02-06 · Saswat Das, Marco Romanelli, Ferdinando Fioretto

Ensuring privacy-preserving inference on cryptographically secure data is a well-known computational challenge. To alleviate the bottleneck of costly cryptographic computations in non-linear activations, recent methods h…

FairnessPrivacy Preserving

MICM: Rethinking Unsupervised Pretraining for Enhanced Few-shot Learning

2024-08-23 · Zhenyu Zhang, Guangyao Chen, Yixiong Zou, Zhimeng Huang 외

Humans exhibit a remarkable ability to learn quickly from a limited number of labeled samples, a capability that starkly contrasts with that of current machine learning systems. Unsupervised Few-Shot Learning (U-FSL) see…

Contrastive LearningFew-Shot LearningUnsupervised Few-Shot Learning