paper-with-me

Papers

Fairshare Data Pricing via Data Valuation for Large Language Models

2025-01-31 · Luyang Zhang, Cathy Jiao, Beibei Li, Chenyan Xiong

Training data is the backbone of large language models (LLMs), yet today's data markets often operate under exploitative pricing -- sourcing data from marginalized groups with little pay or recognition. This paper introduces a theoretical framework for LLM data markets, modeling the strategic interactions between buyers (LLM builders) and sellers (human annotators). We begin with theoretical and empirical analysis showing how exploitative pricing drives high-quality sellers out of the market, degrading data quality and long-term model performance. Then we introduce fairshare, a pricing mechanism grounded in data valuation that quantifies each data's contribution. It aligns incentives by sustaining seller participation and optimizing utility for both buyers and sellers. Theoretically, we show that fairshare yields mutually optimal outcomes: maximizing long-term buyer utility and seller profit while sustaining market participation. Empirically when training open-source LLMs on complex NLP tasks, including math problems, medical diagnosis, and physical reasoning, fairshare boosts seller earnings and ensures a stable supply of high-quality data, while improving buyers' performance-per-dollar and long-term welfare. Our findings offer a concrete path toward fair, transparent, and economically sustainable data markets for LLM. Our code will be open sourced.

📄 PDF Abstract BibTeX arXiv:2502.00198

Code (0)

등록된 구현이 없습니다.

Tasks

Data ValuationMathMedical Diagnosis

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Balanced Off-Policy Evaluation for Personalized Pricing

2023-02-24 · Adam N. Elmachtoub, Vishal Gupta, Yunfan Zhao

We consider a personalized pricing problem in which we have data consisting of feature information, historical pricing decisions, and binary realized demand. The goal is to perform off-policy evaluation for a new persona…

Off-policy evaluation

PricingLogic: Evaluating LLMs Reasoning on Complex Tourism Pricing Tasks

2025-10-14 · Yunuo Liu, Dawei Zhu, Zena Al-Khalili, Dai Cheng 외 arxiv

We present PricingLogic, the first benchmark that probes whether Large Language Models(LLMs) can reliably automate tourism-related prices when multiple, overlapping fare rules apply. Travel agencies are eager to offload …

Arithmetic ReasoningDomain Adaptation

Dynamic Asset Pricing Theory for Life Contingent Risks

2025-03-27 · Patrick Ling

Although the valuation of life contingent assets has been thoroughly investigated under the framework of mathematical statistics, little financial economics research pays attention to the pricing of these assets in a non…

AIGP: An LLM-Based Framework for Long-Term Value Alignment in E-Commerce Pricing

2026-06-25 · Chennan Ma, Yanning Zhang, Siqi Hong, Xiuchong Wang 외 arxiv

Traditional dynamic pricing models in large-scale e-commerce suffer from limited interpretability, poor utilization of unstructured information, and misalignment with long-term business objectives such as cumulative Gros…

Knowledge DistillationReinforcement Learning

Convex Surrogate Loss Functions for Contextual Pricing with Transaction Data

2022-02-16 · Max Biggs

We study an off-policy contextual pricing problem where the seller has access to samples of prices that customers were previously offered, whether they purchased at that price, and auxiliary features describing the custo…

Generalization Bounds