paper-with-me

홈 › Papers

Universal Regression with Adversarial Responses

2022-03-09 · Moïse Blanchard, Patrick Jaillet

We provide algorithms for regression with adversarial responses under large classes of non-i.i.d. instance sequences, on general separable metric spaces, with provably minimal assumptions. We also give characterizations of learnability in this regression context. We consider universal consistency which asks for strong consistency of a learner without restrictions on the value responses. Our analysis shows that such an objective is achievable for a significantly larger class of instance sequences than stationary processes, and unveils a fundamental dichotomy between value spaces: whether finite-horizon mean estimation is achievable or not. We further provide optimistically universal learning rules, i.e., such that if they fail to achieve universal consistency, any other algorithms will fail as well. For unbounded losses, we propose a mild integrability condition under which there exist algorithms for adversarial regression under large classes of non-i.i.d. instance sequences. In addition, our analysis also provides a learning rule for mean estimation in general metric spaces that is consistent under adversarial responses without any moment conditions on the sequence, a result of independent interest.

📄 PDF Abstract BibTeX arXiv:2203.05067

Code (0)

등록된 구현이 없습니다.

Tasks

regression

Similar Papers 제목 키워드 기반

Universal Adversarial Attack on Deep Learning Based Prognostics

2021-09-15 · Arghya Basak, Pradeep Rathore, Sri Harsha Nistala, Sagar Srinivas 외

Deep learning-based time series models are being extensively utilized in engineering and manufacturing industries for process control and optimization, asset monitoring, diagnostic and predictive maintenance. These model…

Adversarial AttackDeep LearningDiagnosticregression+3

Universal Jailbreak Backdoors from Poisoned Human Feedback

2023-11-24 · Javier Rando, Florian Tramèr

Reinforcement Learning from Human Feedback (RLHF) is used to align large language models to produce helpful and harmless responses. Yet, prior work showed these models can be jailbroken by finding adversarial prompts tha…

Backdoor Attack

White-box Multimodal Jailbreaks Against Large Vision-Language Models

2024-05-28 · Ruofan Wang, Xingjun Ma, Hanxu Zhou, Chuanjun Ji 외

Recent advancements in Large Vision-Language Models (VLMs) have underscored their superiority in various multimodal tasks. However, the adversarial robustness of VLMs has not been fully explored. Existing methods mainly …

Adversarial RobustnessAdversarial Text

On the Universal Adversarial Perturbations for Efficient Data-free Adversarial Detection

2023-06-27 · Songyang Gao, Shihan Dou, Qi Zhang, Xuanjing Huang 외

Detecting adversarial samples that are carefully crafted to fool the model is a critical step to socially-secure applications. However, existing adversarial detection methods require access to sufficient training data, w…

text-classificationText Classification

Universal Adversarial Triggers Are Not Universal

2024-04-24 · Nicholas Meade, Arkil Patel, Siva Reddy

Recent work has developed optimization procedures to find token sequences, called adversarial triggers, which can elicit unsafe responses from aligned language models. These triggers are believed to be universally transf…