paper-with-me

Papers

LIMIT-BERT : Linguistics Informed Multi-Task BERT

2020-11-01 · Findings of the Association for Computational Linguistics 2020 · Junru Zhou, Zhuosheng Zhang, Hai Zhao, Shuailiang Zhang

In this paper, we present Linguistics Informed Multi-Task BERT (LIMIT-BERT) for learning language representations across multiple linguistics tasks by Multi-Task Learning. LIMIT-BERT includes five key linguistics tasks: Part-Of-Speech (POS) tags, constituent and dependency syntactic parsing, span and dependency semantic role labeling (SRL). Different from recent Multi-Task Deep Neural Networks (MT-DNN), our LIMIT-BERT is fully linguistics motivated and thus is capable of adopting an improved masked training objective according to syntactic and semantic constituents. Besides, LIMIT-BERT takes a semi-supervised learning strategy to offer the same large amount of linguistics task data as that for the language model training. As a result, LIMIT-BERT not only improves linguistics tasks performance but also benefits from a regularization effect and linguistics information that leads to more general representations to help adapt to new tasks and domains. LIMIT-BERT outperforms the strong baseline Whole Word Masking BERT on both dependency and constituent syntactic/semantic parsing, GLUE benchmark, and SNLI task. Our practice on the proposed LIMIT-BERT also enables us to release a well pre-trained model for multi-purpose of natural language processing tasks once for all.

📄 PDF Abstract BibTeX

Code (1)

DoodleJZ/LIMIT-BERT 공식 구현 pytorch

Tasks

Language ModelingLanguage ModellingMulti-Task LearningPOSSemantic ParsingSemantic Role Labeling

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Attention 설명 없음
Adam 설명 없음
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Residual Connection 설명 없음
WordPiece 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…

Similar Papers 제목 키워드 기반

LIMIT-BERT : Linguistic Informed Multi-Task BERT

2019-10-31 · Junru Zhou, Zhuosheng Zhang, Hai Zhao, Shuailiang Zhang

In this paper, we present a Linguistic Informed Multi-Task BERT (LIMIT-BERT) for learning language representations across multiple linguistic tasks by Multi-Task Learning (MTL). LIMIT-BERT includes five key linguistic sy…

Multi-Task LearningPOSSemantic ParsingSemantic Role Labeling

Syntax-informed Question Answering with Heterogeneous Graph Transformer

2022-04-01 · Fangyi Zhu, Lok You Tan, See-Kiong Ng, Stéphane Bressan

Large neural language models are steadily contributing state-of-the-art performance to question answering and other natural language and information processing tasks. These models are expensive to train. We propose to ev…

Language ModelingLanguage ModellingQuestion Answering

Psycholinguistics meets Continual Learning: Measuring Catastrophic Forgetting in Visual Question Answering

2019-06-10 · ACL 2019 7 · Claudio Greco, Barbara Plank, Raquel Fernández, Raffaella Bernardi

We study the issue of catastrophic forgetting in the context of neural multimodal approaches to Visual Question Answering (VQA). Motivated by evidence from psycholinguistics, we devise a set of linguistically-informed VQ…

Continual LearningQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Enhancing Chinese Pre-trained Language Model via Heterogeneous Linguistics Graph

2022-05-01 · ACL 2022 5 · Yanzeng Li, Jiangxia Cao, Xin Cong, Zhenyu Zhang 외

Chinese pre-trained language models usually exploit contextual character information to learn representations, while ignoring the linguistics knowledge, e.g., word and sentence information. Hence, we propose a task-free …

Language ModelingLanguage ModellingSentence

LingML: Linguistic-Informed Machine Learning for Enhanced Fake News Detection

2024-05-07 · Jasraj Singh, Fang Liu, Hong Xu, Bee Chin Ng 외

Nowadays, Information spreads at an unprecedented pace in social media and discerning truth from misinformation and fake news has become an acute societal challenge. Machine learning (ML) models have been employed to ide…

Fake News DetectionMisinformation