Attributivity and Subjectivity in Contemporary Written Czech
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
SYN2015: Representative Corpus of Contemporary Written Czech
The paper concentrates on the design, composition and annotation of SYN2015, a new 100-million representative corpus of contemporary written Czech. SYN2015 is a sequel of the representative corpora of the SYN series that…
text-classificationText ClassificationCzech Dataset for Cross-lingual Subjectivity Classification
In this paper, we introduce a new Czech subjectivity dataset of 10k manually annotated subjective and objective sentences from movie reviews and descriptions. Our prime motivation is to provide a reliable dataset that ca…
ClassificationSubjectivity AnalysisBenchmark of stylistic variation in LLM-generated texts
This study investigates the register variation in texts written by humans and comparable texts produced by large language models (LLMs). Biber's multidimensional analysis (MDA) is applied to a sample of human-written tex…
Different Time, Different Language: Revisiting the Bias Against Non-Native Speakers in GPT Detectors
LLM-based assistants have been widely popularised after the release of ChatGPT. Concerns have been raised about their misuse in academia, given the difficulty of distinguishing between human-written and generated text. T…
The SYN-series corpora of written Czech
The paper overviews the SYN series of synchronic corpora of written Czech compiled within the framework of the Czech National Corpus project. It describes their design and processing with a focus on the annotation, i.e. …
LemmatizationMorphological Tagging