paper-with-me

홈 › Papers

Disentangling the Linguistic Competence of Privacy-Preserving BERT

2023-10-17 · Stefan Arnold, Nils Kemmerzell, Annika Schreiner

Differential Privacy (DP) has been tailored to address the unique challenges of text-to-text privatization. However, text-to-text privatization is known for degrading the performance of language models when trained on perturbed text. Employing a series of interpretation techniques on the internal representations extracted from BERT trained on perturbed pre-text, we intend to disentangle at the linguistic level the distortion induced by differential privacy. Experimental results from a representational similarity analysis indicate that the overall similarity of internal representations is substantially reduced. Using probing tasks to unpack this dissimilarity, we find evidence that text-to-text privatization affects the linguistic competence across several formalisms, encoding localized properties of words while falling short at encoding the contextual relationships between spans of words.

📄 PDF Abstract BibTeX arXiv:2310.11363

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy Preserving

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…
Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Adam 설명 없음
WordPiece 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

On the Nature of BERT: Correlating Fine-Tuning and Linguistic Competence

2022-10-01 · COLING 2022 10 · Federica Merendi, Felice Dell’Orletta, Giulia Venturi

Several studies in the literature on the interpretation of Neural Language Models (NLM) focus on the linguistic generalization abilities of pre-trained models. However, little attention is paid to how the linguistic know…

Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?

2025-09-02 · Jaime Collado-Montañez, L. Alfonso Ureña-López, Arturo Montejo-Ráez arxiv

Large Language Models offer impressive language capabilities but suffer from well-known limitations, including hallucinations, biases, privacy concerns, and high computational costs. These issues are largely driven by th…

You Are What You Say: Exploiting Linguistic Content for VoicePrivacy Attacks

2025-06-11 · Ünal Ege Gaznepoglu, Anna Leschanowsky, Ahmad Aloradi, Prachi Singh 외

Speaker anonymization systems hide the identity of speakers while preserving other information such as linguistic content and emotions. To evaluate their privacy benefits, attacks in the form of automatic speaker verific…

Language ModelingLanguage ModellingSpeaker anonymizationSpeaker Verification

How Do BERT Embeddings Organize Linguistic Knowledge?

2021-06-01 · NAACL (DeeLIO) 2021 6 · Giovanni Puccetti, Alessio Miaschi, Felice Dell’Orletta

Several studies investigated the linguistic information implicitly encoded in Neural Language Models. Most of these works focused on quantifying the amount and type of information available within their internal represen…

Sentence

Dissociating language and thought in large language models

2023-01-16 · Kyle Mahowald, Anna A. Ivanova, Idan A. Blank, Nancy Kanwisher 외

Large Language Models (LLMs) have come closest among all models to date to mastering human language, yet opinions about their linguistic and cognitive capabilities remain split. Here, we evaluate LLMs using a distinction…