paper-with-me

홈 › Papers

The Stem Cell Hypothesis: Dilemma behind Multi-Task Learning with Transformer Encoders

2021-09-14 · EMNLP 2021 11 · Han He, Jinho D. Choi

Multi-task learning with transformer encoders (MTL) has emerged as a powerful technique to improve performance on closely-related tasks for both accuracy and efficiency while a question still remains whether or not it would perform as well on tasks that are distinct in nature. We first present MTL results on five NLP tasks, POS, NER, DEP, CON, and SRL, and depict its deficiency over single-task learning. We then conduct an extensive pruning analysis to show that a certain set of attention heads get claimed by most tasks during MTL, who interfere with one another to fine-tune those heads for their own objectives. Based on this finding, we propose the Stem Cell Hypothesis to reveal the existence of attention heads naturally talented for many tasks that cannot be jointly trained to create adequate embeddings for all of those tasks. Finally, we design novel parameter-free probes to justify our hypothesis and demonstrate how attention heads are transformed across the five tasks during MTL through label analysis.

📄 PDF Abstract BibTeX arXiv:2109.06939

Code (1)

emorynlp/stem-cell-hypothesis 공식 구현 pytorch

Tasks

Multi-Task LearningNERPOS

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

The Prisoner's dilemma as a cancer model

2016-01-16

Tumor development is an evolutionary process in which a heterogeneous population of cells with differential growth capabilities compete for resources in order to gain a proliferative advantage. What are the minimal ingre…

model

How human-derived brain organoids are built differently from brain organoids derived from genetically-close relatives: A multi-scale hypothesis

2023-04-17 · Tao Zhang, Sarthak Gupta, Madeline A. Lancaster, J. M. Schwarz

How genes affect tissue scale organization remains a longstanding biological puzzle. As experimental efforts aim to quantify gene expression, chromatin organization, cellular structure, and tissue structure, computationa…

Emergent cooperation through mutual information maximization

2020-06-21 · Santiago Cuervo, Marco Alzate

With artificial intelligence systems becoming ubiquitous in our society, its designers will soon have to start to consider its social dimension, as many of these systems will have to interact among them to work efficient…

Deep Reinforcement Learning

Critically assessing atavism, an evolution-centered and deterministic hypothesis on cancer

2024-11-26 · Bertrand Daignan-Fornier, Thomas Pradeu

Cancer is most commonly viewed as resulting from somatic mutations enhancing proliferation and invasion. Some hypotheses further propose that these new capacities reveal a breakdown of multicellularity allowing cancer ce…

Consequentialist conditional cooperation in social dilemmas with imperfect information

2017-10-19 · ICLR 2018 1 · Alexander Peysakhovich, Adam Lerer

Social dilemmas, where mutual cooperation can lead to high payoffs but participants face incentives to cheat, are ubiquitous in multi-agent interaction. We wish to construct agents that cooperate with pure cooperators, a…

Deep Reinforcement LearningReinforcement Learning