paper-with-me

홈 › Papers

AI Alignment through a Game-theoretic Lens: A Survey

2026-08-28 · Yanan Cai, Zhongrui Zhao, Zhigang Lu, Ickjai Lee, Wei Emma Zhang, Minhui Xue, Yihong Zhang, Shuchao Pang, Wei Xiang arxiv

As large language models and increasingly capable AI agents are deployed in high-risk settings, aligning them with complex human values has become a central challenge. Existing alignment methods, while effective in improving helpfulness, harmlessness, and controllability, often struggle to capture real-world preferences that are context-dependent, non-transitive, and shaped by dynamic multi-party interactions. This survey reviews AI alignment through a game-theoretic lens. Specifically, it organizes recent progress around key game-theoretic elements and synthesizes the literature along three challenges: preference diversity, alignment priority, and temporal dynamics. This perspective clarifies where current alignment methods genuinely benefit from game-theoretic analysis, where the framework is looser, and what challenges remain in building robust, adaptive, and verifiable AI systems.

📄 PDF Abstract BibTeX arXiv:2608.27910

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Large Vision-Language Model Alignment and Misalignment: A Survey Through the Lens of Explainability

2025-01-02 · Dong Shu, Haiyan Zhao, Jingyu Hu, Weiru Liu 외

Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in processing both visual and textual information. However, the critical challenge of alignment between visual and linguistic representations…

AttributeLanguage ModelingLanguage Modelling

Affective Game Computing: A Survey

2023-09-25 · Georgios N. Yannakakis, David Melhart

This paper surveys the current state of the art in affective computing principles, methods and tools as applied to games. We review this emerging field, namely affective game computing, through the lens of the four core …

Survey

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

2026-02-12 · Pepijn Cobben, Xuanqiang Angelo Huang, Thao Amelia Pham, Isabel Dahlgren 외 arxiv

Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely evaluate single agents, leaving multi-agent risks such as coordination fai…

'That Darned Sandstorm': A Study of Procedural Generation through Archaeological Storytelling

2023-04-17 · Florence Smith Nicholls, Michael Cook

Procedural content generation has been applied to many domains, especially level design, but the narrative affordances of generated game environments are comparatively understudied. In this paper we present our first att…

A Survey of Automatic Prompt Engineering: An Optimization Perspective

2025-02-17 · Wenwu Li, Xiangfeng Wang, Wenhao Li, Bo Jin

The rise of foundation models has shifted focus from resource-intensive fine-tuning to prompt engineering, a paradigm that steers model behavior through input design rather than weight updates. While manual prompt engine…

cross-modal alignmentPrompt EngineeringSurvey