paper-with-me

홈 › Papers

Shh, don't say that! Domain Certification in LLMs

2025-02-26 · Cornelius Emde, Alasdair Paren, Preetham Arvind, Maxime Kayser, Tom Rainforth, Thomas Lukasiewicz, Bernard Ghanem, Philip H. S. Torr, Adel Bibi

Large language models (LLMs) are often deployed to perform constrained tasks, with narrow domains. For example, customer support bots can be built on top of LLMs, relying on their broad language understanding and capabilities to enhance performance. However, these LLMs are adversarially susceptible, potentially generating outputs outside the intended domain. To formalize, assess, and mitigate this risk, we introduce domain certification; a guarantee that accurately characterizes the out-of-domain behavior of language models. We then propose a simple yet effective approach, which we call VALID that provides adversarial bounds as a certificate. Finally, we evaluate our method across a diverse set of datasets, demonstrating that it yields meaningful certificates, which bound the probability of out-of-domain samples tightly with minimum penalty to refusal behavior.

📄 PDF Abstract BibTeX arXiv:2502.19320

Code (0)

등록된 구현이 없습니다.

Tasks

valid

Methods 이 논문이 사용한 방법론

customer support 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Expert-Guided Prompting and Retrieval-Augmented Generation for Emergency Medical Service Question Answering

2025-11-14 · Xueren Ge, Sahil Murtaza, Anthony Cortez, Homa Alemzadeh arxiv

Large language models (LLMs) have shown promise in medical question answering, yet they often overlook the domain-specific expertise that professionals depend on, such as the clinical subject areas (e.g., trauma, airway)…

Question Answering

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

2026-04-22 · Gustav Keppler, Ghada Elbez, Veit Hagenmeyer arxiv

The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against industry standards. We introduceCyberCertBench, a new suite of Multip…

Question Answering

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

2026-06-29 · Mirko Perkusich, Danyllo Albuquerque, João Paiva, Robson Vilar 외 arxiv

Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt these tools to prepare for certifications such as Professional Scrum Mast…

Harnessing the Power of Large Language Models for Software Testing Education: A Focus on ISTQB Syllabus

2025-10-25 · Tuan-Phong Ngo, Bao-Ngoc Duong, Tuan-Anh Hoang, Joshua Dwight 외 arxiv

Software testing is a critical component in the software engineering field and is important for software engineering education. Thus, it is vital for academia to continuously improve and update educational methods to ref…

Certification of Speaker Recognition Models to Additive Perturbations

2024-04-29 · Dmitrii Korzh, Elvir Karimov, Mikhail Pautov, Oleg Y. Rogov 외

Speaker recognition technology is applied to various tasks, from personal virtual assistants to secure access systems. However, the robustness of these systems against adversarial attacks, particularly to additive pertur…

Few-Shot LearningSpeaker Recognition