paper-with-me

홈 › Papers

Can tweets predict article retractions? A comparison between human and LLM labelling

2024-03-25 · Er-Te Zheng, Hui-Zhen Fu, Mike Thelwall, Zhichao Fang

Quickly detecting problematic research articles is crucial to safeguarding the integrity of scientific research. This study explores whether Twitter mentions of retracted articles can signal potential problems with the articles prior to their retraction, potentially serving as an early warning system for scholars. To investigate this, we analysed a dataset of 4,354 Twitter mentions associated with 504 retracted articles. The effectiveness of Twitter mentions in predicting article retractions was evaluated by both manual and Large Language Model (LLM) labelling. Manual labelling results indicated that 25.7% of tweets signalled problems before retraction. Using the manual labelling results as the baseline, we found that LLMs (GPT-4o-mini, Gemini 1.5 Flash, and Claude-3.5-Haiku) outperformed lexicon-based sentiment analysis tools (e.g., TextBlob) in detecting potential problems, suggesting that automatic detection of problematic articles from social media using LLMs is technically feasible. Nevertheless, since only a small proportion of retracted articles (11.1%) were criticised on Twitter prior to retraction, such automatic systems would detect only a minority of problematic articles. Overall, this study offers insights into how social media data, coupled with emerging generative AI techniques, can support research integrity.

📄 PDF Abstract BibTeX arXiv:2403.16851

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesLanguage ModelingLanguage ModellingLarge Language ModelSentiment Analysis

Similar Papers 제목 키워드 기반

Linking Tweets with Monolingual and Cross-Lingual News using Transformed Word Embeddings

2017-10-25 · Aditya Mogadala, Dominik Jung, Achim Rettinger

Social media platforms have grown into an important medium to spread information about an event published by the traditional media, such as news articles. Grouping such diverse sources of information that discuss the sam…

ArticlesWord Embeddings

KnowBias: A Novel AI Method to Detect Polarity in Online Content

2019-05-02 · Aditya Saligrama

We propose a novel training and inference method for detecting political bias in long text content such as newspaper opinion articles. Obtaining long text data and annotations at sufficient scale for training is difficul…

ArticlesDomain AdaptationGeneral ClassificationSentence+2

KnowBias: Detecting Political Polarity in Long Text Content

2019-09-22 · Aditya Saligrama

We introduce a classification scheme for detecting political bias in long text content such as newspaper opinion articles. Obtaining long text data and annotations at sufficient scale for training is difficult, but it is…

ArticlesDomain AdaptationGeneral ClassificationSentence

Discriminating between standard Romanian and Moldavian tweets using filtered character ngrams

2020-12-01 · VarDial (COLING) 2020 12 · Andrea Ceolin, Hong Zhang

We applied word unigram models, character ngram models, and CNNs to the task of distinguishing tweets of two related dialects of Romanian (standard Romanian and Moldavian) for the VarDial 2020 RDI shared task (Gaman et a…

Articlestext-classificationText Classification

Retractions: Updating from Complex Information

2021-06-21 · Duarte Gonçalves, Jonathan Libgober, Jack Willis

We modify a canonical experimental design to identify the effectiveness of retractions. Comparing beliefs after retractions to beliefs (a) without the retracted information and (b) after equivalent new information, we fi…

Experimental DesignMisinformation