paper-with-me

Papers XLM-R

“XLM-R” 태그가 달린 논문 221편 · 필터 해제

Targeted Multilingual Adaptation for Low-resource Language Families

2024-05-20 · C. M. Downey, Terra Blevins, Dhwani Serai, Dwija Parikh 외

The "massively-multilingual" training of multilingual models is known to limit their utility in any one language, and they perform particularly poorly on low-resource languages. However, there is evidence that low-resour…

XLM-R

Zero-Shot Tokenizer Transfer

2024-05-13 · Benjamin Minixhofer, Edoardo Maria Ponti, Ivan Vulić

Language models (LMs) are bound to their tokenizer, which maps raw text to a sequence of vocabulary items (tokens). This restricts their flexibility: for example, LMs trained primarily on English may still perform well i…

XLM-R

Software Mention Recognition with a Three-Stage Framework Based on BERTology Models at SOMD 2024

2024-04-23 · Thuy Nguyen Thi, Anh Nguyen Viet, Thin Dang Van, Ngan Nguyen Luu Thuy

This paper describes our systems for the sub-task I in the Software Mention Detection in Scholarly Publications shared-task. We propose three approaches leveraging different pre-trained language models (BERT, SciBERT, an…

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+3

Adapting Mental Health Prediction Tasks for Cross-lingual Learning via Meta-Training and In-context Learning with Large Language Model

2024-04-13 · Zita Lifelo, Huansheng Ning, Sahraoui Dhelim

Timely identification is essential for the efficient handling of mental health illnesses such as depression. However, the current research fails to adequately address the prediction of mental health conditions from socia…

Cross-Lingual TransferIn-Context LearningLanguage ModelingLanguage Modelling+3

MaiNLP at SemEval-2024 Task 1: Analyzing Source Language Selection in Cross-Lingual Textual Relatedness

2024-04-03 · Shijia Zhou, Huangyan Shan, Barbara Plank, Robert Litschko

This paper presents our system developed for the SemEval-2024 Task 1: Semantic Textual Relatedness (STR), on Track C: Cross-lingual. The task aims to detect semantic relatedness of two sentences in a given target languag…

Cross-Lingual TransferData AugmentationMachine TranslationXLM-R+1

Cross-Lingual Transfer Robustness to Lower-Resource Languages on Adversarial Datasets

2024-03-29 · Shadi Manafi, Nikhil Krishnaswamy

Multilingual Language Models (MLLMs) exhibit robust cross-lingual transfer capabilities, or the ability to leverage information acquired in a source language and apply it to a target language. These capabilities find pra…

Cross-Lingual Transfernamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+2

Solution for Emotion Prediction Competition of Workshop on Emotionally and Culturally Intelligent AI

2024-03-26 · Shengdong Xu, Zhouyang Chi, Yang Yang

This report provide a detailed description of the method that we explored and proposed in the WECIA Emotion Prediction Competition (EPC), which predicts a person's emotion through an artistic work with a comment. The dat…

DiversityXLM-R

CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification

2024-03-18 · Korbinian Randl, John Pavlopoulos, Aron Henriksson, Tony Lindgren

Contaminated or adulterated food poses a substantial risk to human health. Given sets of labeled web texts for training, Machine Learning and Natural Language Processing can be applied to automatically detect such risks.…

Conformal PredictionIn-Context LearningXLM-R

Machines Do See Color: A Guideline to Classify Different Forms of Racist Discourse in Large Corpora

2024-01-17 · Diana Davila Gordillo, Joan Timoneda, Sebastian Vallejo Vera

Current methods to identify and classify racist language in text rely on small-n qualitative approaches or large-n approaches focusing exclusively on overt forms of racist discourse. This article provides a step-by-step …

text-classificationText ClassificationXLM-R

LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization

2024-01-11 · Muhammad Farid Adilazuarda, Samuel Cahyawijaya, Alham Fikri Aji, Genta Indra Winata 외

Pretrained language models (PLMs) have become remarkably adept at task and language generalization. Nonetheless, they often fail when faced with unseen languages. In this work, we present LinguAlchemy, a regularization m…

intent-classificationIntent ClassificationLanguage ModellingNews Classification+1

Hate Speech and Offensive Content Detection in Indo-Aryan Languages: A Battle of LSTM and Transformers

2023-12-09 · Nikhil Narayan, Mrutyunjay Biswal, Pramod Goyal, Abhranta Panigrahi

Social media platforms serve as accessible outlets for individuals to express their thoughts and experiences, resulting in an influx of user-generated data spanning all age groups. While these platforms enable free expre…

Hate Speech DetectionModel SelectionXLM-R

A Text-to-Text Model for Multilingual Offensive Language Identification

2023-12-06 · Tharindu Ranasinghe, Marcos Zampieri

The ubiquity of offensive content on social media is a growing cause for concern among companies and government organizations. Recently, transformer-based models such as BERT, XLNET, and XLM-R have achieved state-of-the-…

DecoderLanguage IdentificationXLM-R

KBioXLM: A Knowledge-anchored Biomedical Multilingual Pretrained Language Model

2023-11-20 · Lei Geng, Xu Yan, Ziqiang Cao, Juntao Li 외

Most biomedical pretrained language models are monolingual and cannot handle the growing cross-lingual requirements. The scarcity of non-English domain corpora, not to mention parallel data, poses a significant hurdle in…

Language ModelingLanguage ModellingRelationRelation Prediction+1

MELA: Multilingual Evaluation of Linguistic Acceptability

2023-11-15 · Ziyin Zhang, Yikang Liu, Weifang Huang, Junyu Mao 외

In this work, we present the largest benchmark to date on linguistic acceptability: Multilingual Evaluation of Linguistic Acceptability -- MELA, with 46K samples covering 10 languages from a diverse set of language famil…

Code GenerationCross-Lingual TransferLinguistic AcceptabilityMulti-Task Learning+1

Zero-Shot Cross-Lingual Sentiment Classification under Distribution Shift: an Exploratory Study

2023-11-11 · Maarten De Raedt, Semere Kiros Bitew, Fréderic Godin, Thomas Demeester 외

The brittleness of finetuned language model performance on out-of-distribution (OOD) test samples in unseen domains has been well-studied for English, yet is unexplored for multi-lingual models. Therefore, we study gener…

Cross-Lingual Sentiment ClassificationCross-Lingual TransferLanguage ModellingSentiment Analysis+3

Counterfactually Probing Language Identity in Multilingual Models

2023-10-29 · Anirudh Srinivasan, Venkata S Govindarajan, Kyle Mahowald

Techniques in causal analysis of language models illuminate how linguistic information is organized in LLMs. We use one such technique, AlterRep, a method of counterfactual probing, to explore the internal structure of m…

counterfactualLanguage ModelingLanguage ModellingMasked Language Modeling+1

Lost in Translation, Found in Spans: Identifying Claims in Multilingual Social Media

2023-10-27 · Shubham Mittal, Megha Sundriyal, Preslav Nakov

Claim span identification (CSI) is an important step in fact-checking pipelines, aiming to identify text segments that contain a checkworthy claim or assertion in a social media post. Despite its importance to journalist…

Cross-Lingual TransferFact CheckingXLM-R

Improving Cross-Lingual Transfer through Subtree-Aware Word Reordering

2023-10-20 · Ofir Arviv, Dmitry Nikolaev, Taelin Karidi, Omri Abend

Despite the impressive growth of the abilities of multilingual language models, such as XLM-R and mT5, it has been shown that they still face difficulties when tackling typologically-distant languages, particularly in th…

Cross-Lingual TransferPOSXLM-R

MedAI Dialog Corpus (MEDIC): Zero-Shot Classification of Doctor and AI Responses in Health Consultations

2023-10-19 · Olumide E. Ojo, Olaronke O. Adebanji, Alexander Gelbukh, Hiram Calvo 외

Zero-shot classification enables text to be classified into classes not seen during training. In this study, we examine the efficacy of zero-shot learning models in classifying healthcare consultation responses from Doct…

Classificationtext-classificationText ClassificationXLM-R+2

ViSoBERT: A Pre-Trained Language Model for Vietnamese Social Media Text Processing

2023-10-17 · Quoc-Nam Nguyen, Thang Chau Phan, Duc-Vu Nguyen, Kiet Van Nguyen

English and Chinese, known as resource-rich languages, have witnessed the strong development of transformer-based language models for natural language processing tasks. Although Vietnam has approximately 100M people spea…

Language ModelingLanguage ModellingVietnamese Hate Speech DetectionVietnamese Language Models+2
← 이전 21–40 / 221 다음 →