paper-with-me

Papers

Competence-Based Analysis of Language Models

2023-03-01 · Adam Davies, Jize Jiang, ChengXiang Zhai

Despite the recent successes of large, pretrained neural language models (LLMs), comparatively little is known about the representations of linguistic structure they learn during pretraining, which can lead to unexpected behaviors in response to prompt variation or distribution shift. To better understand these models and behaviors, we introduce a general model analysis framework to study LLMs with respect to their representation and use of human-interpretable linguistic properties. Our framework, CALM (Competence-based Analysis of Language Models), is designed to investigate LLM competence in the context of specific tasks by intervening on models' internal representations of different linguistic properties using causal probing, and measuring models' alignment under these interventions with a given ground-truth causal model of the task. We also develop a new approach for performing causal probing interventions using gradient-based adversarial attacks, which can target a broader range of properties and representations than prior techniques. Finally, we carry out a case study of CALM using these interventions to analyze and compare LLM competence across a variety of lexical inference tasks, showing that CALM can be used to explain behaviors across these tasks.

📄 PDF Abstract BibTeX arXiv:2303.00333

Code (0)

등록된 구현이 없습니다.

Tasks

Models Alignment

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Residual Connection 설명 없음
Adam 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…

Similar Papers 제목 키워드 기반

A Theory for Emergence of Complex Skills in Language Models

2023-07-29 · Sanjeev Arora, Anirudh Goyal

A major driver of AI products today is the fact that new skills emerge in language models when their parameter set and training corpora are scaled up. This phenomenon is poorly understood, and a mechanistic explanation v…

Inductive Bias

Dissociating language and thought in large language models

2023-01-16 · Kyle Mahowald, Anna A. Ivanova, Idan A. Blank, Nancy Kanwisher 외

Large Language Models (LLMs) have come closest among all models to date to mastering human language, yet opinions about their linguistic and cognitive capabilities remain split. Here, we evaluate LLMs using a distinction…

Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean

2025-06-02 · Sungho Kim, Nayeon Kim, Taehee Jeon, SangKeun Lee

We introduce the $\underline{Ko}rean \underline{G}rammar \underline{E}valuation Bench\underline{M}ark (KoGEM)$, designed to assess the linguistic competence of LLMs and humans in Korean. KoGEM consists of 1.5k multiple-c…

Multiple-choice

Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion

2026-04-20 · Chen Zhang, Jiuheng Lin, Zhiyuan Liao, Yansong Feng arxiv

Adapting large language models (LLMs) to low-resource languages (LRLs) is constrained by the scarcity of task data and computational resources. Although Proxy Tuning offers a logit-level strategy for introducing scaling …

Continual Pretraining

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation

2025-08-22 · Arka Mukherjee, Shreya Ghosh arxiv

As Vision-Language Models (VLMs) achieve widespread deployment across diverse cultural contexts, ensuring their cultural competence becomes critical for responsible AI systems. While prior work has evaluated cultural awa…

Semantic SimilarityObject RecognitionStory Generation