paper-with-me

홈 › Papers

Truthful Aggregation of LLMs with an Application to Online Advertising

2024-05-09 · Ermis Soumalias, Michael J. Curry, Sven Seuken

The next frontier of online advertising is revenue generation from LLM-generated content. We consider a setting where advertisers aim to influence the responses of an LLM to align with their interests, while platforms seek to maximize advertiser value and ensure user satisfaction. The challenge is that advertisers' preferences generally conflict with those of the user, and advertisers may misreport their preferences. To address this, we introduce MOSAIC, an auction mechanism that ensures that truthful reporting is a dominant strategy for advertisers and that aligns the utility of each advertiser with their contribution to social welfare. Importantly, the mechanism operates without LLM fine-tuning or access to model weights and provably converges to the output of the optimally fine-tuned LLM as computational resources increase. Additionally, it can incorporate contextual information about advertisers, which significantly improves social welfare. Through experiments with a publicly available LLM, we show that MOSAIC leads to high advertiser value and platform revenue with low computational overhead. While our motivating application is online advertising, our mechanism can be applied in any setting with monetary transfers, making it a general-purpose solution for truthfully aggregating the preferences of self-interested agents over LLM-generated replies.

📄 PDF Abstract BibTeX arXiv:2405.05905

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Truthful Online Preference Aggregation for LLM Fine-Tuning in Mobile Crowdsourcing

2026-05-22 · Shugang Hao, Lingjie Duan arxiv

To better serve users' demands in mobile applications (e.g., navigation), mobile crowdsourcing platforms can iteratively align large language model (LLM)-generated content (e.g., AI-generated traffic condition prediction…

An Analysis of Selection Bias Issue for Online Advertising

2022-06-07 · Shinya Suzumura, Hitoshi Abe

In online advertising, a set of potential advertisements can be ranked by a certain auction system where usually the top-1 advertisement would be selected and displayed at an advertising space. In this paper, we show a s…

Multi-Task LearningSelection bias

Online Learning from Strategic Human Feedback in LLM Fine-Tuning

2024-12-22 · Shugang Hao, Lingjie Duan

Reinforcement learning from human feedback (RLHF) has become an essential step in fine-tuning large language models (LLMs) to align them with human preferences. However, human labelers are selfish and have diverse prefer…

Online Deception Detection Refueled by Real World Data Collection

2017-07-28 · RANLP 2017 9 · Wenlin Yao, Zeyu Dai, Ruihong Huang, James Caverlee

The lack of large realistic datasets presents a bottleneck in online deception detection studies. In this paper, we apply a data collection method based on social network analysis to quickly identify high-quality decepti…

Deception Detection

GRATH: Gradual Self-Truthifying for Large Language Models

2024-01-22 · Weixin Chen, Dawn Song, Bo Li

Truthfulness is paramount for large language models (LLMs) as they are increasingly deployed in real-world applications. However, existing LLMs still struggle with generating truthful content, as evidenced by their modes…

TruthfulQA