paper-with-me

홈 › Papers

Written and spoken corpus of real and fake social media postings about COVID-19

2023-10-06 · Ng Bee Chin, Ng Zhi Ee Nicole, Kyla Kwan, Lee Yong Han Dylann, Liu Fang, Xu Hong

This study investigates the linguistic traits of fake news and real news. There are two parts to this study: text data and speech data. The text data for this study consisted of 6420 COVID-19 related tweets re-filtered from Patwa et al. (2021). After cleaning, the dataset contained 3049 tweets, with 2161 labeled as 'real' and 888 as 'fake'. The speech data for this study was collected from TikTok, focusing on COVID-19 related videos. Research assistants fact-checked each video's content using credible sources and labeled them as 'Real', 'Fake', or 'Questionable', resulting in a dataset of 91 real entries and 109 fake entries from 200 TikTok videos with a total word count of 53,710 words. The data was analysed using the Linguistic Inquiry and Word Count (LIWC) software to detect patterns in linguistic data. The results indicate a set of linguistic features that distinguish fake news from real news in both written and speech data. This offers valuable insights into the role of language in shaping trust, social media interactions, and the propagation of fake news.

📄 PDF Abstract BibTeX arXiv:2310.04237

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Variation in Coreference Strategies across Genres and Production Media

2020-12-01 · COLING 2020 8 · Berfin Akta{\c{s}}, Manfred Stede

In response to (i) inconclusive results in the literature as to the properties of coreference chains in written versus spoken language, and (ii) a general lack of work on automatic coreference resolution on both spoken l…

coreference-resolutionCoreference Resolution

Parallel Corpus for Japanese Spoken-to-Written Style Conversion

2020-05-01 · LREC 2020 5 · Mana Ihori, Akihiko Takashima, Ryo Masumura

With the increase of automatic speech recognition (ASR) applications, spoken-to-written style conversion that transforms spoken-style text into written-style text is becoming an important technology to increase the reada…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Punctuation Restorationspeech-recognition+1

Cross-corpora experiments of automatic proficiency assessment and error detection for spoken English

2022-07-01 · NAACL (BEA) 2022 7 · Stefano Bannò, Marco Matassoni

The growing demand for learning English as a second language has led to an increasing interest in automatic approaches for assessing spoken language proficiency. One of the most significant challenges in this field is th…

The Role of User Profile for Fake News Detection

2019-04-30 · Kai Shu, Xinyi Zhou, Suhang Wang, Reza Zafarani 외

Consuming news from social media is becoming increasingly popular. Social media appeals to users due to its fast dissemination of information, low cost, and easy access. However, social media also enables the widespread …

Fake News DetectionFeature ImportanceNews Classification

On task effects in NLG corpus elicitation: a replication study using mixed effects modeling

2019-10-01 · WS 2019 10 · Emiel van Miltenburg, Merel van de Kerkhof, Ruud Koolen, Martijn Goudbeek 외

Task effects in NLG corpus elicitation recently started to receive more attention, but are usually not modeled statistically. We present a controlled replication of the study by Van Miltenburg et al. (2018b), contrasting…