paper-with-me

홈 › Papers

Judicious Selection of Training Data in Assisting Language for Multilingual Neural NER

2018-07-01 · ACL 2018 7 · Rudra Murthy, Anoop Kunchukuttan, Pushpak Bhattacharyya

Multilingual learning for Neural Named Entity Recognition (NNER) involves jointly training a neural network for multiple languages. Typically, the goal is improving the NER performance of one of the languages (the primary language) using the other assisting languages. We show that the divergence in the tag distributions of the common named entities between the primary and assisting languages can reduce the effectiveness of multilingual learning. To alleviate this problem, we propose a metric based on symmetric KL divergence to filter out the highly divergent training instances in the assisting language. We empirically show that our data selection strategy improves NER performance in many languages, including those with very limited training data.

📄 PDF Abstract BibTeX

Code (1)

murthyrudra/NeuralNER 공식 구현 pytorch

Tasks

Domain AdaptationMachine Translationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NERTAG

Similar Papers 제목 키워드 기반

The impact of responding to patient messages with large language model assistance

2023-10-26 · Shan Chen, Marco Guevara, Shalini Moningi, Frank Hoebers 외

Documentation burden is a major contributor to clinician burnout, which is rising nationally and is an urgent threat to our ability to care for patients. Artificial intelligence (AI) chatbots, such as ChatGPT, could redu…

ChatbotDecision MakingLanguage ModelingLanguage Modelling+1

Pre-training via Leveraging Assisting Languages and Data Selection for Neural Machine Translation

2020-01-23 · Haiyue Song, Raj Dabre, Zhuoyuan Mao, Fei Cheng 외

Sequence-to-sequence (S2S) pre-training using large monolingual data is known to improve performance for various S2S NLP tasks in low-resource settings. However, large monolingual corpora might not always be available fo…

Machine TranslationNMTTranslation

NARMADA: Need and Available Resource Managing Assistant for Disasters and Adversities

2020-05-27 · WS 2020 7 · Kaustubh Hiware, Ritam Dutt, Sayan Sinha, Sohan Patro 외

Although a lot of research has been done on utilising Online Social Media during disasters, there exists no system for a specific task that is critical in a post-disaster scenario -- identifying resource-needs and resour…

Information RetrievalManagementRetrieval

Multilingual training set selection for ASR in under-resourced Malian languages

2021-08-13 · Ewald van der Westhuizen, Trideba Padhi, Thomas Niesler

We present first speech recognition systems for the two severely under-resourced Malian languages Bambara and Maasina Fulfulde. These systems will be used by the United Nations as part of a monitoring system to inform an…

Humanitarianspeech-recognitionSpeech Recognition

Action Shapley: A Training Data Selection Metric for World Model in Reinforcement Learning

2026-01-15 · Rajat Ghosh, Debojyoti Dutta arxiv

Numerous offline and model-based reinforcement learning systems incorporate world models to emulate the inherent environments. A world model is particularly important in scenarios where direct interactions with the real …

Computational EfficiencyReinforcement Learning