paper-with-me

Papers

Characterizing Tradeoffs in Language Model Decoding with Informational Interpretations

2023-11-16 · Chung-Ching Chang, William W. Cohen, Yun-Hsuan Sung

We propose a theoretical framework for formulating language model decoder algorithms with dynamic programming and information theory. With dynamic programming, we lift the design of decoder algorithms from the logit space to the action-state value function space, and show that the decoding algorithms are consequences of optimizing the action-state value functions. Each component in the action-state value function space has an information theoretical interpretation. With the lifting and interpretation, it becomes evident what the decoder algorithm is optimized for, and hence facilitating the arbitration of the tradeoffs in sensibleness, diversity, and attribution.

📄 PDF Abstract BibTeX arXiv:2311.10083

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDiversityLanguage ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Optimal Multi-bit Generative Watermarking Schemes Under Worst-Case False-Alarm Constraints

2026-04-09 · Yu-Shin Huang, Chao Tian, Krishna Narayanan arxiv

This paper considers the problem of multi-bit generative watermarking for large language models under a worst-case false-alarm constraint. Prior work established a lower bound on the achievable miss-detection probability…

A Study in Contradiction: Data and Annotation for AIDA Focusing on Informational Conflict in Russia-Ukraine Relations

2022-06-01 · LREC 2022 6 · Jennifer Tracey, Ann Bies, Jeremy Getman, Kira Griffitt 외

This paper describes data resources created for Phase 1 of the DARPA Active Interpretation of Disparate Alternatives (AIDA) program, which aims to develop language technology that can help humans manage large volumes of …

Embracing Contradiction: Theoretical Inconsistency Will Not Impede the Road of Building Responsible AI Systems

2025-05-23 · Gordon Dai, Yunze Xiao

This position paper argues that the theoretical inconsistency often observed among Responsible AI (RAI) metrics, such as differing fairness definitions or tradeoffs between accuracy and privacy, should be embraced as a v…

Fairness

A Theoretical Perspective for Speculative Decoding Algorithm

2024-10-30 · Ming Yin, Minshuo Chen, Kaixuan Huang, Mengdi Wang

Transformer-based autoregressive sampling has been the major bottleneck for slowing down large language model inferences. One effective way to accelerate inference is \emph{Speculative Decoding}, which employs a small mo…

Language ModelingLanguage ModellingLarge Language Model

BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Serving

2026-05-26 · Tianyao Shi, Yi Ding arxiv

Large language model (LLM) serving creates environmental impacts beyond carbon and water, including ecosystem damage through biodiversity-related pathways. We present BIRDS, a framework for Biodiversity Impact of Request…