paper-with-me

Papers

Robust Deep Learning Ensemble against Deception

2020-09-14 · Wenqi Wei, Ling Liu

Deep neural network (DNN) models are known to be vulnerable to maliciously crafted adversarial examples and to out-of-distribution inputs drawn sufficiently far away from the training data. How to protect a machine learning model against deception of both types of destructive inputs remains an open challenge. This paper presents XEnsemble, a diversity ensemble verification methodology for enhancing the adversarial robustness of DNN models against deception caused by either adversarial examples or out-of-distribution inputs. XEnsemble by design has three unique capabilities. First, XEnsemble builds diverse input denoising verifiers by leveraging different data cleaning techniques. Second, XEnsemble develops a disagreement-diversity ensemble learning methodology for guarding the output of the prediction model against deception. Third, XEnsemble provides a suite of algorithms to combine input verification and output verification to protect the DNN prediction models from both adversarial examples and out of distribution inputs. Evaluated using eleven popular adversarial attacks and two representative out-of-distribution datasets, we show that XEnsemble achieves a high defense success rate against adversarial examples and a high detection success rate against out-of-distribution data inputs, and outperforms existing representative defense methods with respect to robustness and defensibility.

📄 PDF Abstract BibTeX arXiv:2009.06589

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial RobustnessDeep LearningDenoisingDiversityEnsemble Learning

Similar Papers 제목 키워드 기반

Deep Neural Network Ensembles against Deception: Ensemble Diversity, Accuracy and Robustness

2019-08-29 · Ling Liu, Wenqi Wei, Ka-Ho Chow, Margaret Loper 외

Ensemble learning is a methodology that integrates multiple DNN learners for improving prediction performance of individual learners. Diversity is greater when the errors of the ensemble prediction is more uniformly dist…

DiversityEnsemble Learning

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

2026-04-15 · Erik Nordby, Tasha Pais, Aviel Parrack arxiv

Linear probes can detect when language models produce outputs they "know" are wrong, a capability relevant to both deception and reward hacking. However, single-layer probes are fragile: the best layer varies across mode…

Cognitive Inception: Agentic Reasoning against Visual Deceptions by Injecting Skepticism

2025-11-21 · Yinjie Zhao, Heng Zhao, Bihan Wen, Joey Tianyi Zhou arxiv

As the development of AI-generated contents (AIGC), multi-modal Large Language Models (LLM) struggle to identify generated visual inputs from real ones. Such shortcoming causes vulnerability against visual deceptions, wh…

Automatic Long-Term Deception Detection in Group Interaction Videos

2019-05-15 · Chongyang Bai, Maksim Bolonkin, Judee Burgoon, Chao Chen 외

Most work on automated deception detection (ADD) in video has two restrictions: (i) it focuses on a video of one person, and (ii) it focuses on a single act of deception in a one or two minute video. In this paper, we pr…

Deception Detection

Adaptive Deception Framework with Behavioral Analysis for Enhanced Cybersecurity Defense

2025-10-02 · Basil Abdullah AL-Zahrani arxiv

This paper presents CADL (Cognitive-Adaptive Deception Layer), an adaptive deception framework achieving 99.88% detection rate with 0.13% false positive rate on the CICIDS2017 dataset. The framework employs ensemble mach…

Intrusion Detection