Deep Detector Health Management under Adversarial Campaigns
Machine learning models are vulnerable to adversarial inputs that induce seemingly unjustifiable errors. As automated classifiers are increasingly used in industrial control systems and machinery, these adversarial errors could grow to be a serious problem. Despite numerous studies over the past few years, the field of adversarial ML is still considered alchemy, with no practical unbroken defenses demonstrated to date, leaving PHM practitioners with few meaningful ways of addressing the problem. We introduce turbidity detection as a practical superset of the adversarial input detection problem, coping with adversarial campaigns rather than statistically invisible one-offs. This perspective is coupled with ROC-theoretic design guidance that prescribes an inexpensive domain adaptation layer at the output of a deep learning model during an attack campaign. The result aims to approximate the Bayes optimal mitigation that ameliorates the detection model's degraded health. A proactively reactive type of prognostics is achieved via Monte Carlo simulation of various adversarial campaign scenarios, by sampling from the model's own turbidity distribution to quickly deploy the correct mitigation during a real-world campaign.
Code (0)
등록된 구현이 없습니다.
Tasks
Domain AdaptationManagementSimilar Papers 제목 키워드 기반
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
Large Language Models (LLMs) have raised increasing concerns about their misuse in generating hate speech. Among all the efforts to address this issue, hate speech detectors play a crucial role. However, the effectivenes…
Adversarial AttackBenchmarkingHate Speech DetectionFake It Until You Break It: On the Adversarial Robustness of AI-generated Image Detectors
While generative AI (GenAI) offers countless possibilities for creative and productive tasks, artificially generated media can be misused for fraud, manipulation, scams, misinformation campaigns, and more. To mitigate th…
Adversarial RobustnessMisinformationThe Effectiveness of Digital Interventions on COVID-19 Attitudes and Beliefs
During the course of the COVID-19 pandemic, a common strategy for public health organizations around the world has been to launch interventions via advertising campaigns on social media. Despite this ubiquity, little has…
IBMMS Decision Support Tool For Management of Bank Telemarketing Campaigns
Although direct marketing is a good method for banks to utilize in the face of global competition and the financial crisis, it has been shown to exhibit poor performance. However, there are some drawbacks to direct campa…
ManagementMarketingA deep adversarial approach based on multi-sensor fusion for remaining useful life prognostics
Multi-sensor systems are proliferating the asset management industry and by proxy, the structural health management community. Asset managers are beginning to require a prognostics and health management system to predict…
Asset ManagementManagementSensor FusionVariational Inference