paper-with-me

Papers

NTULM: Enriching Social Media Text Representations with Non-Textual Units

2022-10-29 · COLING (WNUT) 2022 10 · Jinning Li, Shubhanshu Mishra, Ahmed El-Kishky, Sneha Mehta, Vivek Kulkarni

On social media, additional context is often present in the form of annotations and meta-data such as the post's author, mentions, Hashtags, and hyperlinks. We refer to these annotations as Non-Textual Units (NTUs). We posit that NTUs provide social context beyond their textual semantics and leveraging these units can enrich social media text representations. In this work we construct an NTU-centric social heterogeneous network to co-embed NTUs. We then principally integrate these NTU embeddings into a large pretrained language model by fine-tuning with these additional units. This adds context to noisy short-text social media. Experiments show that utilizing NTU-augmented text representations significantly outperforms existing text-only baselines by 2-5\% relative points on many downstream tasks highlighting the importance of context to social media NLP. We also highlight that including NTU context into the initial layers of language model alongside text is better than using it after the text embedding is generated. Our work leads to the generation of holistic general purpose social media content embedding.

📄 PDF Abstract BibTeX arXiv:2210.16586

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Enriching GNNs with Text Contextual Representations for Detecting Disinformation Campaigns on Social Media

2024-10-24 · Bruno Croso Cunha da Silva, Thomas Palmeira Ferraz, Roseli de Deus Lopes

Disinformation on social media poses both societal and technical challenges, requiring robust detection systems. While previous studies have integrated textual information into propagation networks, they have yet to full…

Data AugmentationFake News Detection

PHEMEPlus: Enriching Social Media Rumour Verification with External Evidence

2022-07-28 · FEVER (ACL) 2022 5 · John Dougrez-Lewis, Elena Kochkina, M. Arana-Catania, Maria Liakata 외

Work on social media rumour verification utilises signals from posts, their propagation and users involved. Other lines of work target identifying and fact-checking claims based on information from Wikipedia, or trustwor…

ArticlesFact Checking

FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG

2024-08-06 · Sai Puppala, Ismail Hossain, Md Jahangir Alam, Sajedul Talukder

Our paper introduces a novel approach to social network information retrieval and user engagement through a personalized chatbot system empowered by Federated Learning GPT. The system is designed to seamlessly aggregate …

ChatbotFederated LearningInformation RetrievalRAG+1

Enriching the E2E dataset

2021-08-01 · INLG (ACL) 2021 8 · Thiago castro Ferreira, Helena Vaz, Brian Davis, Adriana Pagano

This study introduces an enriched version of the E2E dataset, one of the most popular language resources for data-to-text NLG. We extract intermediate representations for popular pipeline tasks such as discourse ordering…

Referring ExpressionReferring expression generation

Trustworthy Hate Speech Detection Through Visual Augmentation

2024-09-20 · Ziyuan Yang, Ming Yan, Yingyu Chen, Hui Wang 외

The surge of hate speech on social media platforms poses a significant challenge, with hate speech detection~(HSD) becoming increasingly critical. Current HSD methods focus on enriching contextual information to enhance …

Hate Speech Detection