paper-with-me

홈 › Papers

A Longitudinal, Multinational, and Multilingual Corpus of News Coverage of the Russo-Ukrainian War

2026-01-22 · Dikshya Mohanty, Taisiia Sabadyn, Jelwin Rodrigues, Chenlu Wang, Abhishek Kalugade, Ritwik Banerjee arxiv

We present DNIPRO, a corpus of 246K news articles from the Russo-Ukrainian war (Feb 2022 -- Aug 2024) spanning eleven outlets across five nation-states (Russia, Ukraine, U.S., U.K., China) and three languages. The corpus features comprehensive metadata and human-evaluated annotations for stance, sentiment, and topical framing, enabling systematic analysis of competing geopolitical narratives. It is uniquely suited for empirical studies of narrative divergence, media framing, and information warfare. Our exploratory analyses reveal how media outlets construct incompatible realities through divergent attribution and topical selection without direct refutation of opposing narratives. DNIPRO empowers empirical research on narrative evolution, cross-lingual information flow, and computational detection of implicit contradictions in fragmented information ecosystems.

📄 PDF Abstract BibTeX arXiv:2601.16309

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MultiNews: A Web collection of an Aligned Multimodal and Multilingual Corpus

2017-11-01 · WS 2017 11 · Haithem Afli, Pintu Lohar, Andy Way

Integrating Natural Language Processing (NLP) and computer vision is a promising effort. However, the applicability of these methods directly depends on the availability of a specific multimodal data that includes images…

ArticlesContent-Based Image RetrievalImage RetrievalMachine Translation+1

Multilingual Open Text Release 1: Public Domain News in 44 Languages

2022-01-14 · LREC 2022 6 · Chester Palen-Michel, June Kim, Constantine Lignos

We present Multilingual Open Text (MOT), a new multilingual corpus containing text in 44 languages, many of which have limited existing text resources for natural language processing. The first release of the corpus cont…

Articles

A Multilingual Simplified Language News Corpus

2022-06-01 · READI (LREC) 2022 6 · Renate Hauser, Jannis Vamvas, Sarah Ebling, Martin Volk

Simplified language news articles are being offered by specialized web portals in several countries. The thousands of articles that have been published over the years are a valuable resource for natural language processi…

ArticlesText Simplification

A Dataset for Multi-lingual Epidemiological Event Extraction

2020-05-01 · LREC 2020 5 · Stephen Mutuvi, Antoine Doucet, Ga{\"e}l Lejeune, Moses Odeo

This paper proposes a corpus for the development and evaluation of tools and techniques for identifying emerging infectious disease threats in online news text. The corpus can not only be used for information extraction,…

ArticlesEvent Extractiontext-classificationText Classification

Novelty in news search: a longitudinal study of the 2020 US elections

2022-11-09 · Roberto Ulloa, Mykola Makhortykh, Aleksandra Urman, Juhi Kulshrestha

The 2020 US elections news coverage was extensive, with new pieces of information generated rapidly. This evolving scenario presented an opportunity to study the performance of search engines in a context in which they h…