paper-with-me

홈 › Papers

POLygraph: Polish Fake News Dataset

2024-07-01 · Daniel Dzienisiewicz, Filip Graliński, Piotr Jabłoński, Marek Kubis, Paweł Skórzewski, Piotr Wierzchoń

This paper presents the POLygraph dataset, a unique resource for fake news detection in Polish. The dataset, created by an interdisciplinary team, is composed of two parts: the "fake-or-not" dataset with 11,360 pairs of news articles (identified by their URLs) and corresponding labels, and the "fake-they-say" dataset with 5,082 news articles (identified by their URLs) and tweets commenting on them. Unlike existing datasets, POLygraph encompasses a variety of approaches from source literature, providing a comprehensive resource for fake news detection. The data was collected through manual annotation by expert and non-expert annotators. The project also developed a software tool that uses advanced machine learning techniques to analyze the data and determine content authenticity. The tool and dataset are expected to benefit various entities, from public sector institutions to publishers and fact-checking organizations. Further dataset exploration will foster fake news detection and potentially stimulate the implementation of similar models in other languages. The paper focuses on the creation and composition of the dataset, so it does not include a detailed evaluation of the software tool for content authenticity analysis, which is planned at a later stage of the project.

📄 PDF Abstract BibTeX arXiv:2407.01393

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesFact CheckingFake News Detection

Similar Papers 제목 키워드 기반

CrossNews-UA: A Cross-lingual News Semantic Similarity Benchmark for Ukrainian, Polish, Russian, and English

2025-10-22 · Daryna Dementieva, Evgeniya Sukhodolskaya, Alexander Fraser arxiv

In the era of social networks and rapid misinformation spread, news analysis remains a critical task. Detecting fake news across multiple languages, particularly beyond English, poses significant challenges. Cross-lingua…

Semantic Similarity

Machine Learning Approach to Fact-Checking in West Slavic Languages

2019-09-01 · RANLP 2019 9 · Pavel P{\v{r}}ib{\'a}{\v{n}}, Tom{\'a}{\v{s}} Hercig, Josef Steinberger

Fake news detection and closely-related fact-checking have recently attracted a lot of attention. Automatization of these tasks has been already studied for English. For other languages, only a few studies can be found (…

BIG-bench Machine LearningFact CheckingFake News Detection

Annotation-Scheme Reconstruction for "Fake News" and Japanese Fake News Dataset

2022-04-06 · Taichi Murayama, Shohei Hisada, Makoto Uehara, Shoko Wakamiya 외

Fake news provokes many societal problems; therefore, there has been extensive research on fake news detection tasks to counter it. Many fake news datasets were constructed as resources to facilitate this task. Contempor…

Fake News Detection

Annotation-Scheme Reconstruction for “Fake News” and Japanese Fake News Dataset

2022-06-01 · LREC 2022 6 · Taichi Murayama, Shohei Hisada, Makoto Uehara, Shoko Wakamiya 외

Fake news provokes many societal problems; therefore, there has been extensive research on fake news detection tasks to counter it. Many fake news datasets were constructed as resources to facilitate this task. Contempor…

Fake News Detection

Each Fake News is Fake in its Own Way: An Attribution Multi-Granularity Benchmark for Multimodal Fake News Detection

2024-12-19 · Hao Guo, Zihan Ma, Zhi Zeng, Minnan Luo 외

Social platforms, while facilitating access to information, have also become saturated with a plethora of fake news, resulting in negative consequences. Automatic multimodal fake news detection is a worthwhile pursuit. E…

Fake News Detection