paper-with-me

Papers

Contrastive-Difference CKA Reveals Concept-Specific Structural Alignment Across Language Model Architectures

2026-06-15 · Xueping Gao arxiv

Do different LLM architectures encode high-level concepts in structurally compatible ways? We systematically characterize a geometric-functional universality dissociation: across multiple concept domains and architectural families, moderate geometric convergence coexists with near-perfect functional transfer. Using contrastive-difference CKA (CKA_Delta), a training-free diagnostic that computes kernel alignment on per-sample contrastive differences, we isolate concept-specific convergence from generic similarity -- achieving significant discrimination where standard CKA cannot. The dissociation replicates across all six concept domains we test (five with p <= 0.017 geometric discrimination and safety as a converging-functional trend, p = 0.08), including two non-instruction concepts (code-vs-NL, reasoning-vs-recall) validated without system prompts; a single 70B--70B pair provides an observational note that universality may strengthen with scale, requiring replication with additional >=70B models. We position CKA_Delta as a practical regime classifier and architectural outlier detector (Gemma: d = 1.08, AUC = 0.79) rather than an absolute transfer-accuracy predictor, providing a training-free diagnostic for cross-architecture concept monitoring.

📄 PDF Abstract BibTeX arXiv:2606.16897

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Machine Unlearning in Hyperbolic vs. Euclidean Multimodal Contrastive Learning: Adapting Alignment Calibration to MERU

2025-03-19 · Àlex Pujol Vidal, Sergio Escalera, Kamal Nasrollahi, Thomas B. Moeslund

Machine unlearning methods have become increasingly important for selective concept removal in large pre-trained models. While recent work has explored unlearning in Euclidean contrastive vision-language models, the effe…

Contrastive LearningMachine Unlearning

Mechanistic Interpretability of Structure-Aware Numerical Reasoning in LLaMA 3.1 8B

2026-08-19 · Rahul Chowdhury, Timothy A Rupprecht, Senhao Cao, Jiahao Liu 외 arxiv

Recent work has shown that large language models (LLMs) exhibit strong numerical sequence modeling capabilities and show promise in time-series prediction. While LLMs display in-context learning capabilities, the mechani…

Contrastive Concept Importance: Explaining Pairwise Class Decisions Through Automatically Extracted Concept Representations

2026-07-30 · Roel Visser, Isaac Roberts, Barbara Hammer arxiv

Concept-based explanations are a prevalent way to explain the decisions of complex black-box methods through semantically meaningful, human-interpretable concepts. To attribute the contribution of such concepts to a mode…

Studying Cultural Differences in Emoji Usage across the East and the West

2019-04-04 · Sharath Chandra Guntuku, Mingyang Li, Louis Tay, Lyle H. Ungar

Global acceptance of Emojis suggests a cross-cultural, normative use of Emojis. Meanwhile, nuances in Emoji use across cultures may also exist due to linguistic differences in expressing emotions and diversity in concept…

Cultural Vocal Bursts Intensity PredictionDiversity

Pseudo Contrastive Learning for Diagram Comprehension in Multimodal Models

2026-02-27 · Hiroshi Sasaki arxiv

Recent multimodal models such as Contrastive Language-Image Pre-training (CLIP) have shown remarkable ability to align visual and linguistic representations. However, domains where small visual differences carry large se…

Visual Question AnsweringContrastive LearningImage-text matching