paper-with-me

Papers

Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models

2025-02-03 · Yuyang Gong, Zhuo Chen, Miaokun Chen, Fengchang Yu, Wei Lu, XiaoFeng Wang, Xiaozhong Liu, Jiawei Liu

Retrieval-Augmented Generation (RAG) systems based on Large Language Models (LLMs) have become essential for tasks such as question answering and content generation. However, their increasing impact on public opinion and information dissemination has made them a critical focus for security research due to inherent vulnerabilities. Previous studies have predominantly addressed attacks targeting factual or single-query manipulations. In this paper, we address a more practical scenario: topic-oriented adversarial opinion manipulation attacks on RAG models, where LLMs are required to reason and synthesize multiple perspectives, rendering them particularly susceptible to systematic knowledge poisoning. Specifically, we propose Topic-FlipRAG, a two-stage manipulation attack pipeline that strategically crafts adversarial perturbations to influence opinions across related queries. This approach combines traditional adversarial ranking attack techniques and leverages the extensive internal relevant knowledge and reasoning capabilities of LLMs to execute semantic-level perturbations. Experiments show that the proposed attacks effectively shift the opinion of the model's outputs on specific topics, significantly impacting user information perception. Current mitigation methods cannot effectively defend against such attacks, highlighting the necessity for enhanced safeguards for RAG systems, and offering crucial insights for LLM security research.

📄 PDF Abstract BibTeX arXiv:2502.01386

Code (0)

등록된 구현이 없습니다.

Tasks

Question AnsweringRAGRetrievalRetrieval-augmented Generation

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Warmup With Linear Decay Linear Warmup With Linear Decay is a learning rate schedule in which we increase the learning rate linearly for $n$ updates and then linearly decay afterwards.
Weight Decay 설명 없음
WordPiece 설명 없음
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.

Similar Papers 제목 키워드 기반

A Disentangled Adversarial Neural Topic Model for Separating Opinions from Plots in User Reviews

2020-10-22 · NAACL 2021 4 · Gabriele Pergola, Lin Gui, Yulan He

The flexibility of the inference process in Variational Autoencoders (VAEs) has recently led to revising traditional probabilistic topic models giving rise to Neural Topic Models (NTMs). Although these approaches have ac…

DisentanglementSentiment AnalysisSentiment ClassificationTopic Models

Multi-dimensional extensions of the Hegselmann-Krause model

2022-04-18 · Giulia De Pasquale, Maria Elena Valcher

In this paper, we consider two multi-dimensional Hagselmann-Krause (HK) models for opinion dynamics. The two models describe how individuals adjust their opinions on multiple topics, based on the influence of their peers…

model

Public Opinion Field Effect Fusion in Representation Learning for Trending Topics Diffusion

2023-09-21 · NeurIPS 2023 11

Trending topic diffusion and prediction analysis is an important problem and has been well studied in social networks. Representation learning is an effective way to extract node embeddings, which can help for topic prop…

Identifying Opinion-Topics and Polarity of Parliamentary Debate Motions

2018-10-01 · WS 2018 10 · Gavin Abercrombie, Riza Theresa Batista-Navarro

Analysis of the topics mentioned and opinions expressed in parliamentary debate motions{--}or proposals{--}is difficult for human readers, but necessary for understanding and automatic processing of the content of the su…

Sentiment AnalysisTopic Classification

MTOS: A LLM-Driven Multi-topic Opinion Simulation Framework for Exploring Echo Chamber Dynamics

2025-10-14 · Dingyi Zuo, Hongjie Zhang, Jie Ou, Chaosheng Feng 외 arxiv

The polarization of opinions, information segregation, and cognitive biases on social media have attracted significant academic attention. In real-world networks, information often spans multiple interrelated topics, pos…