paper-with-me

Papers

Regularized Conventions: Equilibrium Computation as a Model of Pragmatic Reasoning

2023-11-16 · Athul Paul Jacob, Gabriele Farina, Jacob Andreas

We present a model of pragmatic language understanding, where utterances are produced and understood by searching for regularized equilibria of signaling games. In this model (which we call ReCo, for Regularized Conventions), speakers and listeners search for contextually appropriate utterance--meaning mappings that are both close to game-theoretically optimal conventions and close to a shared, ''default'' semantics. By characterizing pragmatic communication as equilibrium search, we obtain principled sampling algorithms and formal guarantees about the trade-off between communicative success and naturalness. Across several datasets capturing real and idealized human judgments about pragmatic implicatures, ReCo matches or improves upon predictions made by best response and rational speech act models of language understanding.

📄 PDF Abstract BibTeX arXiv:2311.09712

Code (0)

등록된 구현이 없습니다.

Tasks

Implicatures

Similar Papers 제목 키워드 기반

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination

2026-06-06 · Yi Xie, Zhanke Zhou, Chentao Cao, Bo Liu 외 arxiv

Multi-agent large language model (LLM) systems often fail to reliably outperform a single strong model equipped with best-of-N sampling. We argue that a core source of this instability is ill-posed equilibrium selection:…

Pragmatics in Language Grounding: Phenomena, Tasks, and Modeling Approaches

2022-11-15 · Daniel Fried, Nicholas Tomlin, Jennifer Hu, Roma Patel 외

People rely heavily on context to enrich meaning beyond what is literally said, enabling concise but effective communication. To interact successfully and naturally with people, user-facing artificial intelligence system…

Grounded language learning

CROP: Token-Efficient Reasoning in Large Language Models via Regularized Prompt Optimization

2026-04-08 · Deep Shah, Sanket Badhe, Nehal Kathrotia, Priyanka Tiwari arxiv

Large Language Models utilizing reasoning techniques improve task performance but incur significant latency and token costs due to verbose generation. Existing automatic prompt optimization(APO) frameworks target task ac…

A Rate-Distortion view of human pragmatic reasoning

2020-05-13 · Noga Zaslavsky, Jennifer Hu, Roger P. Levy

What computational principles underlie human pragmatic reasoning? A prominent approach to pragmatics is the Rational Speech Act (RSA) framework, which formulates pragmatic reasoning as probabilistic speakers and listener…

Beyond Pessimism: Offline Learning in KL-regularized Games

2026-04-08 · Yuheng Zhang, Claire Chen, Nan Jiang arxiv

We study offline learning in KL-regularized two-player zero-sum games, where policies are optimized with respect to a fixed reference policy through KL regularization. Prior work relies on pessimistic value estimation to…