paper-with-me

Papers

REPA: Russian Error Types Annotation for Evaluating Text Generation and Judgment Capabilities

2025-03-17 · Alexander Pugachev, Alena Fenogenova, Vladislav Mikhailov, Ekaterina Artemova

Recent advances in large language models (LLMs) have introduced the novel paradigm of using LLMs as judges, where an LLM evaluates and scores the outputs of another LLM, which often correlates highly with human preferences. However, the use of LLM-as-a-judge has been primarily studied in English. In this paper, we evaluate this framework in Russian by introducing the Russian Error tyPes Annotation dataset (REPA), a dataset of 1k user queries and 2k LLM-generated responses. Human annotators labeled each response pair expressing their preferences across ten specific error types, as well as selecting an overall preference. We rank six generative LLMs across the error types using three rating systems based on human preferences. We also evaluate responses using eight LLM judges in zero-shot and few-shot settings. We describe the results of analyzing the judges and position and length biases. Our findings reveal a notable gap between LLM judge performance in Russian and English. However, rankings based on human and LLM preferences show partial alignment, suggesting that while current LLM judges struggle with fine-grained evaluation in Russian, there is potential for improvement.

📄 PDF Abstract BibTeX arXiv:2503.13102

Code (0)

등록된 구현이 없습니다.

Tasks

2kText Generation

Similar Papers 제목 키워드 기반

Entity Linking over Nested Named Entities for Russian

2022-06-01 · LREC 2022 6 · Natalia Loukachevitch, Pavel Braslavski, Vladimir Ivanov, Tatiana Batura 외

In this paper, we describe entity linking annotation over nested named entities in the recently released Russian NEREL dataset for information extraction. The NEREL collection is currently the largest Russian dataset ann…

Entity Linking

Making Sign Language Corpora Comparable: A Study of Palm-Up and Throw-Away in Polish Sign Language, German Sign Language, and Russian Sign Language

2022-06-01 · SignLang (LREC) 2022 6 · Anna Kuder

This paper is primarily devoted to describing the preparation phase of a large-scale comparative study based on naturalistic linguistic data drawn from multiple sign language corpora. To provide an example, I am using my…

All

Learning Nested Named Entity Recognition from Flat Annotations

2026-02-28 · Igor Rozhkov, Natalia Loukachevitch arxiv

Nested named entity recognition identifies entities contained within other entities, but requires expensive multi-level annotation. While flat NER corpora exist abundantly, nested resources remain scarce. We investigate …

Nested Named Entity Recognition

CoRuSS - a New Prosodically Annotated Corpus of Russian Spontaneous Speech

2016-05-01 · LREC 2016 5 · Tatiana Kachkovskaia, Daniil Kocharov, Pavel Skrelin, Nina Volskaya

This paper describes speech data recording, processing and annotation of a new speech corpus CoRuSS (Corpus of Russian Spontaneous Speech), which is based on connected communicative speech recorded from 60 native Russian…

NEREL-BIO: A Dataset of Biomedical Abstracts Annotated with Nested Named Entities

2022-10-21 · Natalia Loukachevitch, Suresh Manandhar, Elina Baral, Igor Rozhkov 외

This paper describes NEREL-BIO -- an annotation scheme and corpus of PubMed abstracts in Russian and smaller number of abstracts in English. NEREL-BIO extends the general domain dataset NEREL by introducing domain-specif…

Machine Reading ComprehensionReading Comprehension