paper-with-me

홈 › Papers

Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning

2024-10-14 · Aakanksha, Arash Ahmadian, Seraphina Goldfarb-Tarrant, Beyza Ermis, Marzieh Fadaee, Sara Hooker

Large Language Models (LLMs) have been adopted and deployed worldwide for a broad variety of applications. However, ensuring their safe use remains a significant challenge. Preference training and safety measures often overfit to harms prevalent in Western-centric datasets, and safety protocols frequently fail to extend to multilingual settings. In this work, we explore model merging in a diverse multi-task setting, combining safety and general-purpose tasks within a multilingual context. Each language introduces unique and varied learning challenges across tasks. We find that objective-based merging is more effective than mixing data, with improvements of up to 8% and 10% in general performance and safety respectively. We also find that language-based merging is highly effective -- by merging monolingually fine-tuned models, we achieve a 4% increase in general performance and 7% reduction in harm across all languages on top of the data mixtures method using the same available data. Overall, our comprehensive study of merging approaches provides a useful framework for building strong and safe multilingual models.

📄 PDF Abstract BibTeX arXiv:2410.10801

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-Task Learning

Similar Papers 제목 키워드 기반

Emergent Hand Morphology and Control from Optimizing Robust Grasps of Diverse Objects

2020-12-22 · Xinlei Pan, Animesh Garg, Animashree Anandkumar, Yuke Zhu

Evolution in nature illustrates that the creatures' biological structure and their sensorimotor skills adapt to the environmental changes for survival. Likewise, the ability to morph and acquire new skills can facilitate…

Bayesian OptimizationMORPH

A Multi-constraint and Multi-objective Allocation Model for Emergency Rescue in IoT Environment

2024-03-15 · Xinrun Xu, Zhanbiao Lian, Yurong Wu, Manying Lv 외

Emergency relief operations are essential in disaster aftermaths, necessitating effective resource allocation to minimize negative impacts and maximize benefits. In prolonged crises or extensive disasters, a systematic, …

Decision Making

EVINCE: Optimizing Multi-LLM Dialogues Using Conditional Statistics and Information Theory

2024-08-26 · Edward Y. Chang

EVINCE (Entropy and Variation IN Conditional Exchanges) is a novel framework for optimizing multi-LLM dialogues using conditional statistics and information theory. It addresses limitations in multi-agent debate (MAS) fr…

Decision MakingDiversityvalid

DIDI: Diffusion-Guided Diversity for Offline Behavioral Generation

2024-05-23 · Jinxin Liu, Xinghong Guo, Zifeng Zhuang, Donglin Wang

In this paper, we propose a novel approach called DIffusion-guided DIversity (DIDI) for offline behavioral generation. The goal of DIDI is to learn a diverse set of skills from a mixture of label-free offline data. We ac…

D4RLDecision MakingDiversity

Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development

2025-05-22 · Ming Shen, Raphael Shu, Anurag Pratik, James Gung 외

We have seen remarkable progress in large language models (LLMs) empowered multi-agent systems solving complex tasks necessitating cooperation among experts with diverse skills. However, optimizing LLM-based multi-agent …