Increasing the Quality and Quantity of Source Language Data for Unsupervised Cross-Lingual POS Tagging
Code (0)
등록된 구현이 없습니다.
Tasks
Cross-Lingual POS TaggingPOSPOS TaggingSimilar Papers 제목 키워드 기반
Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens
As low-resourced languages are increasingly incorporated into NLP research, there is an emphasis on collecting large-scale datasets. But in prioritizing quantity over quality, we risk 1) building language technologies th…
Machine TranslationQuality versus Quantity: Building Catalan-English MT Resources
In this work, we make the case of quality over quantity when training a MT system for a medium-to-low-resource language pair, namely Catalan-English. We compile our training corpus out of existing resources of varying qu…
Cross-Lingual TransferTransfer LearningTranslationIncreasing Coverage and Precision of Textual Information in Multilingual Knowledge Graphs
Recent work in Natural Language Processing and Computer Vision has been using textual information -- e.g., entity names and descriptions -- available in knowledge graphs to ground neural models to high-quality structured…
Entity LinkingKnowledge Graph CompletionKnowledge GraphsMachine Translation+1Localising In-Domain Adaptation of Transformer-Based Biomedical Language Models
In the era of digital healthcare, the huge volumes of textual information generated every day in hospitals constitute an essential but underused asset that could be exploited with task-specific, fine-tuned biomedical lan…
Domain AdaptationMachine TranslationManagementEnd-to-End Speech Translation for Low-Resource Languages Using Weakly Labeled Data
The scarcity of high-quality annotated data presents a significant challenge in developing effective end-to-end speech-to-text translation (ST) systems, particularly for low-resource languages. This paper explores the hy…
SentenceSpeech-to-TextSpeech-to-Text TranslationTranslation