paper-with-me

홈 › Papers

Kindness in Multi-Agent Reinforcement Learning

2023-11-06 · Farinaz Alamiyan-Harandi, Mersad Hassanjani, Pouria Ramazi

In human societies, people often incorporate fairness in their decisions and treat reciprocally by being kind to those who act kindly. They evaluate the kindness of others' actions not only by monitoring the outcomes but also by considering the intentions. This behavioral concept can be adapted to train cooperative agents in Multi-Agent Reinforcement Learning (MARL). We propose the KindMARL method, where agents' intentions are measured by counterfactual reasoning over the environmental impact of the actions that were available to the agents. More specifically, the current environment state is compared with the estimation of the current environment state provided that the agent had chosen another action. The difference between each agent's reward, as the outcome of its action, with that of its fellow, multiplied by the intention of the fellow is then taken as the fellow's "kindness". If the result of each reward-comparison confirms the agent's superiority, it perceives the fellow's kindness and reduces its own reward. Experimental results in the Cleanup and Harvest environments show that training based on the KindMARL method enabled the agents to earn 89\% (resp. 37\%) and 44% (resp. 43\%) more total rewards than training based on the Inequity Aversion and Social Influence methods. The effectiveness of KindMARL is further supported by experiments in a traffic light control problem.

📄 PDF Abstract BibTeX arXiv:2311.04239

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualCounterfactual ReasoningFairnessMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

We Urgently Need Intrinsically Kind Machines

2024-10-21 · Joshua T. S. Hewson

Artificial Intelligence systems are rapidly evolving, integrating extrinsic and intrinsic motivations. While these frameworks offer benefits, they risk misalignment at the algorithmic level while appearing superficially …

The Hunger Game Debate: On the Emergence of Over-Competition in Multi-Agent Systems

2025-09-30 · Xinbei Ma, Ruotian Ma, Xingyu Chen, Zhengliang Shi 외 arxiv

LLM-based multi-agent systems demonstrate great potential for tackling complex problems, but how competition shapes their behavior remains underexplored. This paper investigates the over-competition in multi-agent debate…

Novel Artificial Human Optimization Field Algorithms - The Beginning

2019-03-26 · Satish Gajawada, Hassan Mustafa

New Artificial Human Optimization (AHO) Field Algorithms can be created from scratch or by adding the concept of Artificial Humans into other existing Optimization Algorithms. Particle Swarm Optimization (PSO) has been v…

Articles

Learning Through AI-Clones: Enhancing Self-Perception and Presentation Performance

2023-10-23 · Qingxiao Zheng, Zhuoer Chen, Yun Huang

This study examines the impact of AI-generated digital clones with self-images on enhancing perceptions and skills in online presentations. A mixed-design experiment with 44 international students compared self-recording…

Face SwappingVoice Cloning

Combining Theory of Mind and Kindness for Self-Supervised Human-AI Alignment

2024-10-21 · Joshua T. S. Hewson

As artificial intelligence (AI) becomes deeply integrated into critical infrastructures and everyday life, ensuring its safe deployment is one of humanity's most urgent challenges. Current AI models prioritize task optim…