paper-with-me

홈 › Papers

Multi-Faceted Studies on Data Poisoning can Advance LLM Development

2025-02-20 · Pengfei He, Yue Xing, Han Xu, Zhen Xiang, Jiliang Tang

The lifecycle of large language models (LLMs) is far more complex than that of traditional machine learning models, involving multiple training stages, diverse data sources, and varied inference methods. While prior research on data poisoning attacks has primarily focused on the safety vulnerabilities of LLMs, these attacks face significant challenges in practice. Secure data collection, rigorous data cleaning, and the multistage nature of LLM training make it difficult to inject poisoned data or reliably influence LLM behavior as intended. Given these challenges, this position paper proposes rethinking the role of data poisoning and argue that multi-faceted studies on data poisoning can advance LLM development. From a threat perspective, practical strategies for data poisoning attacks can help evaluate and address real safety risks to LLMs. From a trustworthiness perspective, data poisoning can be leveraged to build more robust LLMs by uncovering and mitigating hidden biases, harmful outputs, and hallucinations. Moreover, from a mechanism perspective, data poisoning can provide valuable insights into LLMs, particularly the interplay between data and model behavior, driving a deeper understanding of their underlying mechanisms.

📄 PDF Abstract BibTeX arXiv:2502.14182

Code (1)

PengfeiHePower/awesome-LLM-data-poisoning 공식 구현

Tasks

Data Poisoning

Similar Papers 제목 키워드 기반

Multifaceted User Modeling in Recommendation: A Federated Foundation Models Approach

2024-12-22 · Chunxu Zhang, Guodong Long, Hongkuan Guo, Zhaojie Liu 외

Multifaceted user modeling aims to uncover fine-grained patterns and learn representations from user data, revealing their diverse interests and characteristics, such as profile, preference, and personality. Recent studi…

Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective

2023-11-30 · Dawen Zhang, Boming Xia, Yue Liu, Xiwei Xu 외

The advent of Generative AI has marked a significant milestone in artificial intelligence, demonstrating remarkable capabilities in generating realistic images, texts, and data patterns. However, these advancements come …

Data PoisoningMachine Unlearning

Explaining Vulnerabilities to Adversarial Machine Learning through Visual Analytics

2019-07-17 · Yuxin Ma, Tiankai Xie, Jundong Li, Ross Maciejewski

Machine learning models are currently being deployed in a variety of real-world applications where model predictions are used to make decisions about healthcare, bank loans, and numerous other critical tasks. As the depl…

BIG-bench Machine LearningData Poisoning

Manipulating Recommender Systems: A Survey of Poisoning Attacks and Countermeasures

2024-04-23 · Thanh Toan Nguyen, Quoc Viet Hung Nguyen, Thanh Tam Nguyen, Thanh Trung Huynh 외

Recommender systems have become an integral part of online services to help users locate specific information in a sea of data. However, existing studies show that some recommender systems are vulnerable to poisoning att…

Recommendation SystemsSurvey

Advancing Automated Speaking Assessment Leveraging Multifaceted Relevance and Grammar Information

2025-06-19 · Hao-Chien Lu, Jhen-Ke Lin, Hong-Yun Lin, Chung-Chun Wang 외

Current automated speaking assessment (ASA) systems for use in multi-aspect evaluations often fail to make full use of content relevance, overlooking image or exemplar cues, and employ superficial grammar analysis that l…