paper-with-me

홈 › Papers

Evaluating Language Tools for Fifteen EU-official Under-resourced Languages

2020-10-23 · LREC 2020 5 · Diego Alves, Gaurish Thakkar, Marko Tadić

This article presents the results of the evaluation campaign of language tools available for fifteen EU-official under-resourced languages. The evaluation was conducted within the MSC ITN CLEOPATRA action that aims at building the cross-lingual event-centric knowledge processing on top of the application of linguistic processing chains (LPCs) for at least 24 EU-official languages. In this campaign, we concentrated on three existing NLP platforms (Stanford CoreNLP, NLP Cube, UDPipe) that all provide models for under-resourced languages and in this first run we covered 15 under-resourced languages for which the models were available. We present the design of the evaluation campaign and present the results as well as discuss them. We considered the difference between reported and our tested results within a single percentage point as being within the limits of acceptable tolerance and thus consider this result as reproducible. However, for a number of languages, the results are below what was reported in the literature, and in some cases, our testing results are even better than the ones reported previously. Particularly problematic was the evaluation of NERC systems. One of the reasons is the absence of universally or cross-lingually applicable named entities classification scheme that would serve the NERC task in different languages analogous to the Universal Dependency scheme in parsing task. To build such a scheme has become one of our the future research directions.

📄 PDF Abstract BibTeX arXiv:2010.12428

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MCP Server Architecture Patterns for LLM-Integrated Applications

2026-06-29 · Carson Rodrigues, Oysturn Vas arxiv

The Model Context Protocol (MCP), introduced by Anthropic in November 2024, defines a standardized interface for connecting large language models (LLMs) to external tools, data sources, and services. Within months of rel…

Evaluating a Multi-sense Definition Generation Model for Multiple Languages

2020-06-12 · Arman Kabiri, Paul Cook

Most prior work on definition modeling has not accounted for polysemy, or has done so by considering definition modeling for a target word in a given context. In contrast, in this study, we propose a context-agnostic app…

Word Embeddings

Evaluating the Effectiveness of Large Language Models in Solving Simple Programming Tasks: A User-Centered Study

2025-07-05 · Kai Deng arxiv

As large language models (LLMs) become more common in educational tools and programming environments, questions arise about how these systems should interact with users. This study investigates how different interaction …

How many words does ChatGPT know? The answer is ChatWords

2023-09-28 · Gonzalo Martínez, Javier Conde, Pedro Reviriego, Elena Merino-Gómez 외

The introduction of ChatGPT has put Artificial Intelligence (AI) Natural Language Processing (NLP) in the spotlight. ChatGPT adoption has been exponential with millions of users experimenting with it in a myriad of tasks…

ArxEval: Evaluating Retrieval and Generation in Language Models for Scientific Literature

2025-01-17 · Aarush Sinha, Viraj Virk, Dipshikha Chakraborty, P. S. Sreeja

Language Models [LMs] are now playing an increasingly large role in information generation and synthesis; the representation of scientific knowledge in these systems needs to be highly accurate. A prime challenge is hall…

HallucinationRetrieval