paper-with-me

Papers

Is On-Device AI Broken and Exploitable? Assessing the Trust and Ethics in Small Language Models

2024-06-08 · Kalyan Nakka, Jimmy Dani, Nitesh Saxena

In this paper, we present a very first study to investigate trust and ethical implications of on-device artificial intelligence (AI), focusing on small language models (SLMs) amenable for personal devices like smartphones. While on-device SLMs promise enhanced privacy, reduced latency, and improved user experience compared to cloud-based services, we posit that they might also introduce significant risks and vulnerabilities compared to their on-server counterparts. As part of our trust assessment study, we conduct a systematic evaluation of the state-of-the-art on-devices SLMs, contrasted to their on-server counterparts, based on a well-established trustworthiness measurement framework. Our results show on-device SLMs to be significantly less trustworthy, specifically demonstrating more stereotypical, unfair and privacy-breaching behavior. Informed by these findings, we then perform our ethics assessment study using a dataset of unethical questions, that depicts harmful scenarios. Our results illustrate the lacking ethical safeguards in on-device SLMs, emphasizing their capabilities of generating harmful content. Further, the broken safeguards and exploitable nature of on-device SLMs is demonstrated using potentially unethical vanilla prompts, to which the on-device SLMs answer with valid responses without any filters and without the need for any jailbreaking or prompt engineering. These responses can be abused for various harmful and unethical scenarios like: societal harm, illegal activities, hate, self-harm, exploitable phishing content and many others, all of which indicates the severe vulnerability and exploitability of these on-device SLMs.

📄 PDF Abstract BibTeX arXiv:2406.05364

Code (0)

등록된 구현이 없습니다.

Tasks

EthicsPrompt Engineeringvalid

Similar Papers 제목 키워드 기반

Can We Trust AI Agents? A Case Study of an LLM-Based Multi-Agent System for Ethical AI

2024-10-25 · José Antonio Siqueira de Cerqueira, Mamia Agbese, Rebekah Rousi, Nannan Xi 외

AI-based systems, including Large Language Models (LLM), impact millions by supporting diverse tasks but face issues like misinformation, bias, and misuse. AI ethics is crucial as new technologies and concerns emerge, bu…

Bias DetectionEthicsFairnessMisinformation

AraTrust: An Evaluation of Trustworthiness for LLMs in Arabic

2024-03-14 · Emad A. Alghamdi, Reem I. Masoud, Deema Alnuhait, Afnan Y. Alomairi 외

The swift progress and widespread acceptance of artificial intelligence (AI) systems highlight a pressing requirement to comprehend both the capabilities and potential risks associated with AI. Given the linguistic compl…

EthicsMultiple-choice

Antisocial Analagous Behavior, Alignment and Human Impact of Google AI Systems: Evaluating through the lens of modified Antisocial Behavior Criteria by Human Interaction, Independent LLM Analysis, and AI Self-Reflection

2024-03-21 · Alan D. Ogilvie

Google AI systems exhibit patterns mirroring antisocial personality disorder (ASPD), consistent across models from Bard on PaLM to Gemini Advanced, meeting 5 out of 7 ASPD modified criteria. These patterns, along with co…

Ethics

Beyond Bias and Compliance: Towards Individual Agency and Plurality of Ethics in AI

2023-02-23 · Thomas Krendl Gilbert, Megan Welle Brozek, Andrew Brozek

AI ethics is an emerging field with multiple, competing narratives about how to best solve the problem of building human values into machines. Two major approaches are focused on bias and compliance, respectively. But ne…

Ethics

Semantic Chain-of-Trust: Autonomous Trust Orchestration for Collaborator Selection via Hypergraph-Aided Agentic AI

2025-07-31 · Botao Zhu, Xianbin Wang, Dusit Niyato arxiv

The effective completion of tasks in collaborative systems hinges on task-specific trust evaluations of potential devices for distributed collaboration. Due to independent operation of devices involved, dynamic evolution…