paper-with-me

홈 › Papers

SPARTA ALIGNMENT: Collectively Aligning Multiple Language Models through Combat

2025-06-05 · Yuru Jiang, Wenxuan Ding, Shangbin Feng, Greg Durrett, Yulia Tsvetkov

We propose SPARTA ALIGNMENT, an algorithm to collectively align multiple LLMs through competition and combat. To complement a single model's lack of diversity in generation and biases in evaluation, multiple LLMs form a "sparta tribe" to compete against each other in fulfilling instructions while serving as judges for the competition of others. For each iteration, one instruction and two models are selected for a duel, the other models evaluate the two responses, and their evaluation scores are aggregated through a adapted elo-ranking based reputation system, where winners/losers of combat gain/lose weight in evaluating others. The peer-evaluated combat results then become preference pairs where the winning response is preferred over the losing one, and all models learn from these preferences at the end of each iteration. SPARTA ALIGNMENT enables the self-evolution of multiple LLMs in an iterative and collective competition process. Extensive experiments demonstrate that SPARTA ALIGNMENT outperforms initial models and 4 self-alignment baselines across 10 out of 12 tasks and datasets with 7.0% average improvement. Further analysis reveals that SPARTA ALIGNMENT generalizes more effectively to unseen tasks and leverages the expertise diversity of participating models to produce more logical, direct and informative outputs.

📄 PDF Abstract BibTeX arXiv:2506.04721

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Probabilistic Aggregation and Targeted Embedding Optimization for Collective Moral Reasoning in Large Language Models

2025-06-17 · Chenchen Yuan, Zheyu Zhang, Shuo Yang, Bardh Prenkaj 외

Large Language Models (LLMs) have shown impressive moral reasoning abilities. Yet they often diverge when confronted with complex, multi-factor moral dilemmas. To address these discrepancies, we propose a framework that …

Investigating Post-pretraining Representation Alignment for Cross-Lingual Question Answering

2021-09-24 · EMNLP (MRQA) 2021 11 · Fahim Faisal, Antonios Anastasopoulos

Human knowledge is collectively encoded in the roughly 6500 languages spoken around the world, but it is not distributed equally across languages. Hence, for information-seeking question answering (QA) systems to adequat…

Cross-Lingual Question AnsweringQuestion Answering

Unsupervised Hyperalignment for Multilingual Word Embeddings

2018-11-02 · Jean Alaux, Edouard Grave, Marco Cuturi, Armand Joulin

We consider the problem of aligning continuous word representations, learned in multiple languages, to a common space. It was recently shown that, in the case of two languages, it is possible to learn such a mapping with…

Multilingual Word EmbeddingsTranslationWord EmbeddingsWord Translation

SPARTAN: Sparse Hierarchical Memory for Parameter-Efficient Transformers

2022-11-29 · Ameet Deshpande, Md Arafat Sultan, Anthony Ferritto, Ashwin Kalyan 외

Fine-tuning pre-trained language models (PLMs) achieves impressive performance on a range of downstream tasks, and their sizes have consequently been getting bigger. Since a different copy of the model is required for ea…

Raspberry Pi 4

Unsupervised Hyper-alignment for Multilingual Word Embeddings

2019-05-01 · ICLR 2019 5 · Jean Alaux, Edouard Grave, Marco Cuturi, Armand Joulin

We consider the problem of aligning continuous word representations, learned in multiple languages, to a common space. It was recently shown that, in the case of two languages, it is possible to learn such a mapping with…

Multilingual Word EmbeddingsTranslationWord EmbeddingsWord Translation