paper-with-me

홈 › Papers

With False Friends Like These, Who Can Have Self-Knowledge?

2020-09-28 · Lue Tao, Songcan Chen

Adversarial examples arise from excessive sensitivity of a model. Commonly studied adversarial examples are malicious inputs, crafted by an adversary from correctly classified examples, to induce misclassification. This paper studies an intriguing, yet far overlooked consequence of the excessive sensitivity, that is, a misclassified example can be easily perturbed to help the model to produce correct output. Such perturbed examples look harmless, but actually can be maliciously utilized by a false friend to make the model self-satisfied. Thus we name them hypocritical examples. With false friends like these, a poorly performed model could behave like a state-of-the-art one. Once a deployer trusts the hypocritical performance and uses the "well-performed" model in real-world applications, potential security concerns appear even in benign environments. In this paper, we formalize the hypocritical risk for the first time and propose a defense method specialized for hypocritical examples by minimizing the tradeoff between natural risk and an upper bound of hypocritical risk. Moreover, our theoretical analysis reveals connections between adversarial risk and hypocritical risk. Extensive experiments verify the theoretical results and the effectiveness of our proposed methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Sensitivity

Similar Papers 제목 키워드 기반

Is Cross-Lingual Transfer in Bilingual Models Human-Like? A Study with Overlapping Word Forms in Dutch and English

2026-04-08 · Iza Škrjanec, Irene Elisabeth Winther, Vera Demberg, Stefan L. Frank arxiv

Bilingual speakers show cross-lingual activation during reading, especially for words with shared surface form. Cognates (friends) typically lead to facilitation, whereas interlingual homographs (false friends) cause int…

Cross-Lingual Transfer

Automatically Building a Multilingual Lexicon of False Friends With No Supervision

2020-05-01 · LREC 2020 5 · Ana Sabina Uban, Liviu P. Dinu

Cognate words, defined as words in different languages which derive from a common etymon, can be useful for language learners, who can leverage the orthographical similarity of cognates to more easily understand a text i…

Cross-Lingual Word EmbeddingsLanguage AcquisitionWord Embeddings

A High Coverage Method for Automatic False Friends Detection for Spanish and Portuguese

2018-08-01 · COLING 2018 8 · Santiago Castro, Jairo Bonanata, Aiala Ros{\'a}

False friends are words in two languages that look or sound similar, but have different meanings. They are a common source of confusion among language learners. Methods to detect them automatically do exist, however they…

A Computational Approach to Measuring the Semantic Divergence of Cognates

2020-12-02 · Ana-Sabina Uban, Alina-Maria Ciobanu, Liviu P. Dinu

Meaning is the foundation stone of intercultural communication. Languages are continuously changing, and words shift their meanings for various reasons. Semantic divergence in related languages is a key concern of histor…

Cross-Lingual Word EmbeddingsSemantic SimilaritySemantic Textual SimilarityTranslation+1

The Complexity of Manipulation of k-Coalitional Games on Graphs

2024-08-14 · Hodaya Barr, Yohai Trabelsi, Sarit Kraus, Liam Roditty 외

In many settings, there is an organizer who would like to divide a set of agents into $k$ coalitions, and cares about the friendships within each coalition. Specifically, the organizer might want to maximize utilitarian …