paper-with-me

홈 › Papers

Attributivity and Subjectivity in Contemporary Written Czech

2021-12-01 · Quasy (SyntaxFest) 2021 12 · Miroslav Kubát, Radek Čech, Xinying Chen
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SYN2015: Representative Corpus of Contemporary Written Czech

2016-05-01 · LREC 2016 5 · Michal K{\v{r}}en, V{\'a}clav Cvr{\v{c}}ek, Tom{\'a}{\v{s}} {\v{C}}apka, Anna {\v{C}}erm{\'a}kov{\'a} 외

The paper concentrates on the design, composition and annotation of SYN2015, a new 100-million representative corpus of contemporary written Czech. SYN2015 is a sequel of the representative corpora of the SYN series that…

text-classificationText Classification

Czech Dataset for Cross-lingual Subjectivity Classification

2022-04-29 · LREC 2022 6 · Pavel Přibáň, Josef Steinberger

In this paper, we introduce a new Czech subjectivity dataset of 10k manually annotated subjective and objective sentences from movie reviews and descriptions. Our prime motivation is to provide a reliable dataset that ca…

ClassificationSubjectivity Analysis

Benchmark of stylistic variation in LLM-generated texts

2025-09-12 · Jiří Milička, Anna Marklová, Václav Cvrček arxiv

This study investigates the register variation in texts written by humans and comparable texts produced by large language models (LLMs). Biber's multidimensional analysis (MDA) is applied to a sample of human-written tex…

Different Time, Different Language: Revisiting the Bias Against Non-Native Speakers in GPT Detectors

2026-02-05 · Adnan Al Ali, Jindřich Helcl, Jindřich Libovický arxiv

LLM-based assistants have been widely popularised after the release of ChatGPT. Concerns have been raised about their misuse in academia, given the difficulty of distinguishing between human-written and generated text. T…

The SYN-series corpora of written Czech

2014-05-01 · LREC 2014 5 · Milena Hn{\'a}tkov{\'a}, Michal K{\v{r}}en, Pavel Proch{\'a}zka, Hana Skoumalov{\'a}

The paper overviews the SYN series of synchronic corpora of written Czech compiled within the framework of the Czech National Corpus project. It describes their design and processing with a focus on the annotation, i.e. …

LemmatizationMorphological Tagging