paper-with-me

Papers

A Survey of Safety and Trustworthiness of Deep Neural Networks: Verification, Testing, Adversarial Attack and Defence, and Interpretability

2018-12-18 · Xiaowei Huang, Daniel Kroening, Wenjie Ruan, James Sharp, Youcheng Sun, Emese Thamo, Min Wu, Xinping Yi

In the past few years, significant progress has been made on deep neural networks (DNNs) in achieving human-level performance on several long-standing tasks. With the broader deployment of DNNs on various applications, the concerns over their safety and trustworthiness have been raised in public, especially after the widely reported fatal incidents involving self-driving cars. Research to address these concerns is particularly active, with a significant number of papers released in the past few years. This survey paper conducts a review of the current research effort into making DNNs safe and trustworthy, by focusing on four aspects: verification, testing, adversarial attack and defence, and interpretability. In total, we survey 202 papers, most of which were published after 2017.

📄 PDF Abstract BibTeX arXiv:1812.08342

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial AttackSelf-Driving CarsSurvey

Similar Papers 제목 키워드 기반

A Survey of Safety and Trustworthiness of Large Language Models through the Lens of Verification and Validation

2023-05-19 · Xiaowei Huang, Wenjie Ruan, Wei Huang, Gaojie Jin 외

Large Language Models (LLMs) have exploded a new heatwave of AI for their ability to engage end-users in human-level conversations with detailed and articulate answers across many knowledge domains. In response to their …

Towards Trustworthy GUI Agents: A Survey

2025-03-30 · Yucheng Shi, Wenhao Yu, Wenlin Yao, Wenhu Chen 외

GUI agents, powered by large foundation models, can interact with digital interfaces, enabling various applications in web automation, mobile navigation, and software testing. However, their increasing autonomy has raise…

Decision MakingSequential Decision Makingsoftware testingSurvey

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

2026-05-17 · Jinhu Qi, Muzhi Li, Jiahong Liu, Yuqin Shu 외 arxiv

Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks autonomously, but their multi-step trajectories introduce new failu…

From Automation to Collaboration: Human-in-the-Loop Methods for Safe and Trustworthy NLP

2026-05-24 · Most. Sharmin Sultana Samu, MD. Tanvir Ahmed Seum, Md. Rakibul Islam arxiv

Large language models are widely deployed in high-stakes NLP tasks, yet risks such as bias, hallucination, adversarial vulnerability and unreliable generalization remain. Probe-based auditing reveals inconsistencies in m…

Text Generation

Deep Learning-Based Autonomous Driving Systems: A Survey of Attacks and Defenses

2021-04-05 · Yao Deng, Tiehua Zhang, Guannan Lou, Xi Zheng 외

The rapid development of artificial intelligence, especially deep learning technology, has advanced autonomous driving systems (ADSs) by providing precise control decisions to counterpart almost any driving event, spanni…

Anomaly DetectionAutonomous DrivingDeep Learning