paper-with-me

홈 › Papers

Multilingual Central Repository: a Cross-lingual Framework for Developing Wordnets

2021-07-01 · Xavier Gómez Guinovart, Itziar Gonzalez-Dios, Antoni Oliver, German Rigau

Language resources are necessary for language processing,but building them is costly, involves many researches from different areas and needs constant updating. In this paper, we describe the crosslingual framework used for developing the Multilingual Central Repository (MCR), a multilingual knowledge base that includes wordnets of Basque, Catalan, English, Galician, Portuguese, Spanish and the following ontologies: Base Concepts, Top Ontology, WordNet Domains and Suggested Upper Merged Ontology. We present the story of MCR, its state in 2017 and the developed tools.

📄 PDF Abstract BibTeX arXiv:2107.00333

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multilingual Central Repository version 3.0

2012-05-01 · LREC 2012 5 · Aitor Gonzalez-Agirre, Egoitz Laparra, German Rigau

This paper describes the upgrading process of the Multilingual Central Repository (MCR). The new MCR uses WordNet 3.0 as Interlingual-Index (ILI). Now, the current version of the MCR integrates in the same EuroWordNet fr…

M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation

2024-10-28 · Jiaheng Liu, Ken Deng, Congnan Liu, Jian Yang 외

Repository-level code completion has drawn great attention in software engineering, and several benchmark datasets have been introduced. However, existing repository-level code completion benchmarks usually focus on a li…

Code Completion

Resources for Multilingual Hate Speech Detection

2022-07-01 · NAACL (WOAH) 2022 7 · Ayme Arango Monnar, Jorge Perez, Barbara Poblete, Magdalena Saldaña 외

Most of the published approaches and resources for hate speech detection are tailored for the English language. In consequence, cross-lingual and cross-cultural perspectives lack some essential resources.The lack of dive…

DiversityHate Speech Detection

ML2B: Multi-Lingual ML Benchmark For AutoML

2025-09-26 · Ekaterina Trofimova, Zosia Shamina, Maria Selifanova, Artem Zaitsev 외 arxiv

Large language models (LLMs) have recently demonstrated strong capabilities in generating machine learning (ML) code, enabling end-to-end pipeline construction from natural language instructions. However, existing benchm…

Representation LearningCode Generation

Exploring Polyglot Harmony: On Multilingual Data Allocation for Large Language Models Pretraining

2025-09-19 · Ping Guo, Yubing Ren, Binbin Liu, Fengze Liu 외 arxiv

Large language models (LLMs) have become integral to a wide range of applications worldwide, driving an unprecedented global demand for effective multilingual capabilities. Central to achieving robust multilingual perfor…