paper-with-me

홈 › Papers

Robustness in Large Language Models: A Survey of Mitigation Strategies and Evaluation Metrics

2025-05-24 · Pankaj Kumar, Subhankar Mishra

Large Language Models (LLMs) have emerged as a promising cornerstone for the development of natural language processing (NLP) and artificial intelligence (AI). However, ensuring the robustness of LLMs remains a critical challenge. To address these challenges and advance the field, this survey provides a comprehensive overview of current studies in this area. First, we systematically examine the nature of robustness in LLMs, including its conceptual foundations, the importance of consistent performance across diverse inputs, and the implications of failure modes in real-world applications. Next, we analyze the sources of non-robustness, categorizing intrinsic model limitations, data-driven vulnerabilities, and external adversarial factors that compromise reliability. Following this, we review state-of-the-art mitigation strategies, and then we discuss widely adopted benchmarks, emerging metrics, and persistent gaps in assessing real-world reliability. Finally, we synthesize findings from existing surveys and interdisciplinary studies to highlight trends, unresolved issues, and pathways for future research.

📄 PDF Abstract BibTeX arXiv:2505.18658

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Measure and Improve Robustness in NLP Models: A Survey

2021-12-15 · NAACL 2022 7 · Xuezhi Wang, Haohan Wang, Diyi Yang

As NLP models achieved state-of-the-art performances over benchmarks and gained wide applications, it has been increasingly important to ensure the safe deployment of these models in the real world, e.g., making sure the…

Survey

Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions

2025-09-29 · Smita Khapre, Melkamu Abay Mersha, Hassan Shakil, Jonali Baruah 외 arxiv

The evolution of digital communication systems and the designs of online platforms have inadvertently facilitated the subconscious propagation of toxic behavior. Giving rise to reactive responses to toxic behavior. Toxic…

Language Generation Models Can Cause Harm: So What Can We Do About It? An Actionable Survey

2022-10-14 · Sachin Kumar, Vidhisha Balachandran, Lucille Njoo, Antonios Anastasopoulos 외

Recent advances in the capacity of large language models to generate human-like text have resulted in their increased adoption in user-facing settings. In parallel, these improvements have prompted a heated discourse aro…

Language ModelingLanguage ModellingSurveyText Generation

Safety of Embodied Navigation: A Survey

2025-08-07 · Zixia Wang, Jia Hu, Ronghui Mu arxiv

As large language models (LLMs) continue to advance and gain influence, the development of embodied AI has accelerated, drawing significant attention, particularly in navigation scenarios. Embodied navigation requires an…

A Survey on Proactive Defense Strategies Against Misinformation in Large Language Models

2025-07-05 · Shuliang Liu, Hongyi Liu, Aiwei Liu, Bingchen Duan 외

The widespread deployment of large language models (LLMs) across critical domains has amplified the societal risks posed by algorithmically generated misinformation. Unlike traditional false content, LLM-generated misinf…

Misinformation