paper-with-me

Papers

Time-Aware Datasets are Adaptive Knowledgebases for the New Normal

2022-11-22 · Abhijit Suprem, Sanjyot Vaidya, Joao Eduardo Ferreira, Calton Pu

Recent advances in text classification and knowledge capture in language models have relied on availability of large-scale text datasets. However, language models are trained on static snapshots of knowledge and are limited when that knowledge evolves. This is especially critical for misinformation detection, where new types of misinformation continuously appear, replacing old campaigns. We propose time-aware misinformation datasets to capture time-critical phenomena. In this paper, we first present evidence of evolving misinformation and show that incorporating even simple time-awareness significantly improves classifier accuracy. Second, we present COVID-TAD, a large-scale COVID-19 misinformation da-taset spanning 25 months. It is the first large-scale misinformation dataset that contains multiple snapshots of a datastream and is orders of magnitude bigger than related misinformation datasets. We describe the collection and labeling pro-cess, as well as preliminary experiments.

📄 PDF Abstract BibTeX arXiv:2211.12508

Code (0)

등록된 구현이 없습니다.

Tasks

Misinformationtext-classificationText Classification

Similar Papers 제목 키워드 기반

A Foundry of Human Activities and Infrastructures

2017-10-31 · Robert B. Allen, Eunsang Yang, Tatsawan Timakum

Direct representation knowledgebases can enhance and even provide an alternative to document-centered digital libraries. Here we consider realist semantic modeling of everyday activities and infrastructures in such knowl…

LitSumm: Large language models for literature summarisation of non-coding RNAs

2023-11-06 · Andrew Green, Carlos Ribas, Nancy Ontiveros-Palacios, Sam Griffiths-Jones 외

Curation of literature in life sciences is a growing challenge. The continued increase in the rate of publication, coupled with the relatively fixed number of curators worldwide presents a major challenge to developers o…

Contrastive Entity Coreference and Disambiguation for Historical Texts

2024-06-21 · Abhishek Arora, Emily Silcock, Leander Heldring, Melissa Dell

Massive-scale historical document collections are crucial for social science research. Despite increasing digitization, these documents typically lack unique cross-document identifiers for individuals mentioned within th…

Articlescoreference-resolutionCoreference ResolutionCross Document Coreference Resolution+1

GraphSubDetector: Time Series Subsequence Anomaly Detection via Density-Aware Adaptive Graph Neural Network

2024-11-26 · Weiqi Chen, Zhiqiang Zhou, Qingsong Wen, Liang Sun

Time series subsequence anomaly detection is an important task in a large variety of real-world applications ranging from health monitoring to AIOps, and is challenging due to the following reasons: 1) how to effectively…

Anomaly DetectionGraph Neural NetworkTime Series

Expert Knowledge & Machine Understanding: Bridging Reactome's Ontology with LLM Semantic Embeddings

2026-08-28 · Susanna Bravi, Riccardo De Luca, Rosa Sicilia, Christine Nardini 외 arxiv

Biological knowledgebases like Reactome provide high-quality pathways that include biological elements' relationships and textual descriptions (metadata). The quality of such pathways is granted by manual curation, that …