paper-with-me

Papers

The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation

2025-01-14 · XiaoYu Zhang, Juan Zhai, Shiqing Ma, Qingshuang Bao, Weipeng Jiang, Qian Wang, Chao Shen, Yang Liu

Large Language Models (LLMs) have emerged as the new recommendation engines, surpassing traditional methods in both capability and scope, particularly in code generation. In this paper, we reveal a novel provider bias in LLMs: without explicit directives, these models show systematic preferences for services from specific providers in their recommendations (e.g., favoring Google Cloud over Microsoft Azure). To systematically investigate this bias, we develop an automated pipeline to construct the dataset, incorporating 6 distinct coding task categories and 30 real-world application scenarios. Leveraging this dataset, we conduct the first comprehensive empirical study of provider bias in LLM code generation across seven state-of-the-art LLMs, utilizing approximately 500 million tokens (equivalent to $5,000+ in computational costs). Our findings reveal that LLMs exhibit significant provider preferences, predominantly favoring services from Google and Amazon, and can autonomously modify input code to incorporate their preferred providers without users' requests. Such a bias holds far-reaching implications for market dynamics and societal equilibrium, potentially contributing to digital monopolies. It may also deceive users and violate their expectations, leading to various consequences. We call on the academic community to recognize this emerging issue and develop effective evaluation and mitigation methods to uphold AI security and fairness.

📄 PDF Abstract BibTeX arXiv:2501.07849

Code (0)

등록된 구현이 없습니다.

Tasks

Code GenerationDataset GenerationFairness

Methods 이 논문이 사용한 방법론

Uphold 설명 없음

Similar Papers 제목 키워드 기반

Reverse CAPTCHA: Evaluating LLM Susceptibility to Invisible Unicode Instruction Injection

2026-02-26 · Marcus Graves arxiv

We introduce Reverse CAPTCHA, an evaluation framework that tests whether large language models follow invisible Unicode-encoded instructions embedded in otherwise normal-looking text. Unlike traditional CAPTCHAs that dis…

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

2026-04-17 · Nicolò Pagan, Christopher Barrie, Chris Andrew Bail, Petter Törnberg arxiv

Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these tasks remains poorly understood: which biases are robust across provi…

WMCopier: Forging Invisible Image Watermarks on Arbitrary Images

2025-03-28 · Ziping Dong, Chao Shuai, Zhongjie Ba, Peng Cheng 외

Invisible Image Watermarking is crucial for ensuring content provenance and accountability in generative AI. While Gen-AI providers are increasingly integrating invisible watermarking systems, the robustness of these sch…

Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated Images

2023-11-23 · Shicheng Xu, Danyang Hou, Liang Pang, Jingcheng Deng 외

With the advancement of generation models, AI-generated content (AIGC) is becoming more realistic, flooding the Internet. A recent study suggests that this phenomenon causes source bias in text retrieval for web search. …

Cross-Modal RetrievalImage RetrievalRetrievalText Retrieval

Unveiling the Invisible: Reasoning Complex Occlusions Amodally with AURA

2025-03-13 · Zhixuan Li, Hyunse Yoon, SangHoon Lee, Weisi Lin

Amodal segmentation aims to infer the complete shape of occluded objects, even when the occluded region's appearance is unavailable. However, current amodal segmentation methods lack the capability to interact with users…

Dataset GenerationReasoning SegmentationSegmentation