paper-with-me

Papers

A Variational Framework for LLM Generator-Regulator Games

2026-06-16 · Quanyan Zhu arxiv

This paper develops a variational framework for regulated language generation. Starting from autoregressive token sampling, we derive the induced distribution over complete messages and relate it to an entropy-regularized Gibbs law. Regulation is modeled as an optimal discriminator whose convex-dual value is an f-divergence, and the generator-regulator interaction is formulated as a saddle-point problem. The framework applies to moderation, censorship, AI deception detection, compliance auditing, phishing defense, and manipulation control, where regulation concerns a distribution over possible messages rather than a single output. The equilibrium clarifies the tradeoff among utility, entropy, regulatory alignment, and finite-length detectability. Two finite-vocabulary case studies, censorship filtering and phishing defense, illustrate how the theory can be evaluated through utility, entropy, divergence, receiver-side scores, and detection probability.

📄 PDF Abstract BibTeX arXiv:2606.18424

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Principal agent mean field games in REC markets

2021-12-22 · Dena Firoozi, Arvind V Shrivats, Sebastian Jaimungal

Principal agent games are a growing area of research which focuses on the optimal behaviour of a principal and an agent, with the former contracting work from the latter, in return for providing a monetary award. While t…

Navigate

Variational Positive-incentive Noise: How Noise Benefits Models

2023-06-13 · Hongyuan Zhang, Sida Huang, Xuelong Li

A large number of works aim to alleviate the impact of noise due to an underlying conventional assumption of the negative role of noise. However, some existing works show that the assumption does not always hold. In this…

Variational Inference

Expected Variational Inequalities

2025-02-25 · Brian Hu Zhang, Ioannis Anagnostides, Emanuel Tewolde, Ratip Emin Berker 외

Variational inequalities (VIs) encompass many fundamental problems in diverse areas ranging from engineering to economics and machine learning. However, their considerable expressivity comes at the cost of computational …

Regulation Games for Trustworthy Machine Learning

2024-02-05 · Mohammad Yaghini, Patty Liu, Franziska Boenisch, Nicolas Papernot

Existing work on trustworthy machine learning (ML) often concentrates on individual aspects of trust, such as fairness or privacy. Additionally, many techniques overlook the distinction between those who train ML models …

FairnessGender Classification

Learning in Games with Lossy Feedback

2018-12-01 · NeurIPS 2018 12 · Zhengyuan Zhou, Panayotis Mertikopoulos, Susan Athey, Nicholas Bambos 외

We consider a game-theoretical multi-agent learning problem where the feedback information can be lost during the learning process and rewards are given by a broad class of games known as variationally stable games. We p…