paper-with-me

Papers

Typologically Informed Parameter Aggregation

2026-01-23 · Stef Accou, Wessel Poelman arxiv

Massively multilingual language models enable cross-lingual generalization but underperform on low-resource and unseen languages. While adapter-based fine-tuning offers a parameter-efficient solution, training language-specific adapters at scale remains costly. We introduce Typologically Informed Parameter Aggregation (TIPA), a training-free method that constructs proxy language adapters by aggregating existing ones, weighted by typological similarity. Integrated into the MAD-X framework, these proxies enable zero-shot cross-lingual transfer without additional training. We evaluate TIPA on five NLP tasks and over 230 languages. TIPA consistently outperforms or matches baselines such as English-only fine-tuning or selecting the typologically closest language adapter. We see the largest gains for languages lacking dedicated adapters. Our results demonstrate that typologically informed aggregation provides a viable alternative to language-specific modules without any training needed.

📄 PDF Abstract BibTeX arXiv:2601.16629

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-Shot Cross-Lingual Transfer

Similar Papers 제목 키워드 기반

A Principled Framework for Evaluating on Typologically Diverse Languages

2024-07-06 · Esther Ploeger, Wessel Poelman, Andreas Holck Høeg-Petersen, Anders Schlichtkrull 외

Beyond individual languages, multilingual natural language processing (NLP) research increasingly aims to develop models that perform well across languages generally. However, evaluating these systems on all the world's …

UCxn: Typologically Informed Annotation of Constructions Atop Universal Dependencies

2024-03-26 · Leonie Weissweiler, Nina Böbel, Kirian Guiller, Santiago Herrera 외

The Universal Dependencies (UD) project has created an invaluable collection of treebanks with contributions in over 140 languages. However, the UD annotations do not tell the full story. Grammatical constructions that c…

Fisher-Informed Parameterwise Aggregation for Federated Learning with Heterogeneous Data

2026-01-20 · Zhipeng Chang, Ting He, Wenrui Hao arxiv

Federated learning aggregates model updates from distributed clients, but standard first order methods such as FedAvg apply the same scalar weight to all parameters from each client. Under non-IID data, these uniformly w…

Image ClassificationFederated Learning

The ACQDIV Corpus Database and Aggregation Pipeline

2020-05-01 · LREC 2020 5 · Anna Jancso, Steven Moran, Sabine Stoll

We present the ACQDIV corpus database and aggregation pipeline, a tool developed as part of the European Research Council (ERC) funded project ACQDIV, which aims to identify the universal cognitive processes that allow c…

Language Acquisition

LingGym: How Far Are LLMs from Thinking Like Field Linguists?

2025-11-01 · Changbing Yang, Franklin Ma, Freda Shi, Jian Zhu arxiv

This paper introduces LingGym, a new benchmark that evaluates LLMs' capacity for meta-linguistic reasoning using Interlinear Glossed Text (IGT) and grammatical descriptions extracted from 18 typologically diverse referen…