P\'art\'elet: A Hungarian Corpus of Propaganda Texts from the Hungarian Socialist Era
In this paper, we present P{\'a}rt{\'e}let, a digitized Hungarian corpus of Communist propaganda texts. P{\'a}rt{\'e}let was the official journal of the governing party during the Hungarian socialism from 1956 to 1989, hence it represents the direct political agitation and propaganda of the dictatorial system in question. The paper has a dual purpose: first, to present a general review of the corpus compilation process and the basic statistical data of the corpus, and second, to demonstrate through two case studies what the dataset can be used for. We show that our corpus provides a unique opportunity for conducting research on Hungarian propaganda discourse, as well as analyzing changes of this discourse over a 35-year period of time with computer-assisted methods.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A Hungarian Sentiment Corpus Manually Annotated at Aspect Level
In this paper we present a Hungarian sentiment corpus manually annotated at aspect level. Our corpus consists of Hungarian opinion texts written about different types of products. The main aim of creating the corpus was …
Sentiment Analysistext annotationDetecting Uncertainty Cues in Hungarian Social Media Texts
In this paper, we aim at identifying uncertainty cues in Hungarian social media texts. We present our machine learning based uncertainty detector which is based on a rich features set including lexical, morphological, sy…
BIG-bench Machine LearningDomain AdaptationInstance SearchHunOr: A Hungarian---Russian Parallel Corpus
In this paper, we present HunOr, the first multi-domain Hungarian―Russian parallel corpus. Some of the corpus texts have been manually aligned and split into sentences, besides, named entities also have been annotated …
Cross-Lingual Information RetrievalInformation RetrievalMachine TranslationNamed Entity Recognition (NER)+4Identification and Analysis of Personification in Hungarian: The PerSECorp project
Despite the recent findings on the conceptual and linguistic organization of personification, we have relatively little knowledge about its lexical patterns and grammatical templates. It is especially true in the case of…
POSSyntax-based data augmentation for Hungarian-English machine translation
We train Transformer-based neural machine translation models for Hungarian-English and English-Hungarian using the Hunglish2 corpus. Our best models achieve a BLEU score of 40.0 on HungarianEnglish and 33.4 on English-Hu…
Data AugmentationMachine TranslationTranslation