paper-with-me

Papers

Learning to Learn Group Alignment: A Self-Tuning Credo Framework with Multiagent Teams

2023-04-14 · David Radke, Kyle Tilbury

Mixed incentives among a population with multiagent teams has been shown to have advantages over a fully cooperative system; however, discovering the best mixture of incentives or team structure is a difficult and dynamic problem. We propose a framework where individual learning agents self-regulate their configuration of incentives through various parts of their reward function. This work extends previous work by giving agents the ability to dynamically update their group alignment during learning and by allowing teammates to have different group alignment. Our model builds on ideas from hierarchical reinforcement learning and meta-learning to learn the configuration of a reward function that supports the development of a behavioral policy. We provide preliminary results in a commonly studied multiagent environment and find that agents can achieve better global outcomes by self-tuning their respective group alignment parameters.

📄 PDF Abstract BibTeX arXiv:2304.07337

Code (0)

등록된 구현이 없습니다.

Tasks

Hierarchical Reinforcement LearningMeta-Learning

Similar Papers 제목 키워드 기반

The Importance of Credo in Multiagent Learning

2022-04-15 · David Radke, Kate Larson, Tim Brecht

We propose a model for multi-objective optimization, a credo, for agents in a system that are configured into multiple groups (i.e., teams). Our model of credo regulates how agents optimize their behavior for the groups …

reinforcement-learningReinforcement Learning (RL)

CREDO: Epistemic-Aware Conformalized Credal Envelopes for Regression

2026-03-06 · Luben M. C. Cabezas, Sabina J. Sloman, Bruno M. Resende, Fanyi Wu 외 arxiv

Conformal prediction delivers prediction intervals with distribution-free coverage, but its intervals can look overconfident in regions where the model is extrapolating, because standard conformal scores do not explicitl…

Neural Network Architecture for Credibility Assessment of Textual Claims

2018-03-28 · Nurendra Choudhary, Rajat Singh, Ishita Bindlish, Manish Shrivastava

Text articles with false claims, especially news, have recently become aggravating for the Internet users. These articles are in wide circulation and readers face difficulty discerning fact from fiction. Previous work on…

ArticlesSemantic SimilaritySemantic Textual Similarity

Anchored Alignment for Self-Explanations Enhancement

2024-10-17 · Luis Felipe Villa-Arenas, Ata Nizamoglu, Qianli Wang, Sebastian Möller 외

In this work, we introduce a methodology for alignment designed to enhance the ability of large language models (LLMs) to articulate their reasoning (self-explanation) even in the absence of annotated rationale explanati…

Dataset Generation

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning

2025-06-04 · Liang Chen, Xueting Han, Li Shen, Jing Bai 외

Harmful fine-tuning (HFT), performed directly on open-source LLMs or through Fine-tuning-as-a-Service, breaks safety alignment and poses significant threats. Existing methods aim to mitigate HFT risks by learning robust …

Safety Alignment