paper-with-me

홈 › Papers

Increasing the Quality and Quantity of Source Language Data for Unsupervised Cross-Lingual POS Tagging

2013-10-01 · IJCNLP 2013 10 · Long Duong, Paul Cook, Steven Bird, Pavel Pecina
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual POS TaggingPOSPOS Tagging

Similar Papers 제목 키워드 기반

Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens

2025-11-05 · Hellina Hailu Nigatu, Bethelhem Yemane Mamo, Bontu Fufa Balcha, Debora Taye Tesfaye 외 arxiv

As low-resourced languages are increasingly incorporated into NLP research, there is an emphasis on collecting large-scale datasets. But in prioritizing quantity over quality, we risk 1) building language technologies th…

Machine Translation

Quality versus Quantity: Building Catalan-English MT Resources

2022-06-01 · SIGUL (LREC) 2022 6 · Ona de Gibert Bonet, Ksenia Kharitonova, Blanca Calvo Figueras, Jordi Armengol-Estapé 외

In this work, we make the case of quality over quantity when training a MT system for a medium-to-low-resource language pair, namely Catalan-English. We compile our training corpus out of existing resources of varying qu…

Cross-Lingual TransferTransfer LearningTranslation

Increasing Coverage and Precision of Textual Information in Multilingual Knowledge Graphs

2023-11-27 · Simone Conia, Min Li, Daniel Lee, Umar Farooq Minhas 외

Recent work in Natural Language Processing and Computer Vision has been using textual information -- e.g., entity names and descriptions -- available in knowledge graphs to ground neural models to high-quality structured…

Entity LinkingKnowledge Graph CompletionKnowledge GraphsMachine Translation+1

Localising In-Domain Adaptation of Transformer-Based Biomedical Language Models

2022-12-20 · Tommaso Mario Buonocore, Claudio Crema, Alberto Redolfi, Riccardo Bellazzi 외

In the era of digital healthcare, the huge volumes of textual information generated every day in hospitals constitute an essential but underused asset that could be exploited with task-specific, fine-tuned biomedical lan…

Domain AdaptationMachine TranslationManagement

End-to-End Speech Translation for Low-Resource Languages Using Weakly Labeled Data

2025-06-19 · Aishwarya Pothula, Bhavana Akkiraju, Srihari Bandarupalli, Charan D 외

The scarcity of high-quality annotated data presents a significant challenge in developing effective end-to-end speech-to-text translation (ST) systems, particularly for low-resource languages. This paper explores the hy…

SentenceSpeech-to-TextSpeech-to-Text TranslationTranslation