paper-with-me

Papers

Overview of GUA-SPA at IberLEF 2023: Guarani-Spanish Code Switching Analysis

2023-09-12 · Luis Chiruzzo, Marvin Agüero-Torales, Gustavo Giménez-Lugo, Aldo Alvarez, Yliana Rodríguez, Santiago Góngora, Thamar Solorio

We present the first shared task for detecting and analyzing code-switching in Guarani and Spanish, GUA-SPA at IberLEF 2023. The challenge consisted of three tasks: identifying the language of a token, NER, and a novel task of classifying the way a Spanish span is used in the code-switched context. We annotated a corpus of 1500 texts extracted from news articles and tweets, around 25 thousand tokens, with the information for the tasks. Three teams took part in the evaluation phase, obtaining in general good results for Task 1, and more mixed results for Tasks 2 and 3.

📄 PDF Abstract BibTeX arXiv:2309.06163

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesNERSingle Particle Analysis

Similar Papers 제목 키워드 기반

On the logistical difficulties and findings of Jopara Sentiment Analysis

2021-05-06 · NAACL (CALCS) 2021 6 · Marvin M. Agüero-Torales, David Vilares, Antonio G. López-Herrera

This paper addresses the problem of sentiment analysis for Jopara, a code-switching language between Guarani and Spanish. We first collect a corpus of Guarani-dominant tweets and discuss on the difficulties of finding qu…

BIG-bench Machine LearningSentiment Analysis

Development of a Guarani - Spanish Parallel Corpus

2020-05-01 · LREC 2020 5 · Luis Chiruzzo, Pedro Amarilla, Adolfo R{\'\i}os, Gustavo Gim{\'e}nez Lugo

This paper presents the development of a Guarani - Spanish parallel corpus with sentence-level alignment. The Guarani sentences of the corpus use the Jopara Guarani dialect, the dialect of Guarani spoken in Paraguay, whi…

Sentence

Overview of ADoBo at IberLEF 2025: Automatic Detection of Anglicisms in Spanish

2025-07-29 · Elena Alvarez-Mellado, Jordi Porta-Zamorano, Constantine Lignos, Julio Gonzalo arxiv

This paper summarizes the main findings of ADoBo 2025, the shared task on anglicism identification in Spanish proposed in the context of IberLEF 2025. Participants of ADoBo 2025 were asked to detect English lexical borro…

Jojajovai: A Parallel Guarani-Spanish Corpus for MT Benchmarking

2022-06-01 · LREC 2022 6 · Luis Chiruzzo, Santiago Góngora, Aldo Alvarez, Gustavo Giménez-Lugo 외

This work presents a parallel corpus of Guarani-Spanish text aligned at sentence level. The corpus contains about 30,000 sentence pairs, and is structured as a collection of subsets from different sources, further split …

BenchmarkingSentenceTranslation

Overview of CAPITEL Shared Tasks at IberLEF 2020: Named Entity Recognition and Universal Dependencies Parsing

2020-11-11 · Jordi Porta-Zamorano, Luis Espinosa-Anke

We present the results of the CAPITEL-EVAL shared task, held in the context of the IberLEF 2020 competition series. CAPITEL-EVAL consisted on two subtasks: (1) Named Entity Recognition and Classification and (2) Universa…

ArticlesDependency Parsingnamed-entity-recognitionNamed Entity Recognition+1