paper-with-me

홈 › Papers

Entanglement as Memory: Mechanistic Interpretability of Quantum Language Models

2026-03-27 · Nathan Roll arxiv

Quantum language models have shown competitive performance on sequential tasks, yet whether trained quantum circuits exploit genuinely quantum resources -- or merely embed classical computation in quantum hardware -- remains unknown. Prior work has evaluated these models through endpoint metrics alone, without examining the memory strategies they actually learn internally. We introduce the first mechanistic interpretability study of quantum language models, combining causal gate ablation, entanglement tracking, and density-matrix interchange interventions on a controlled long-range dependency task. We find that single-qubit models are exactly classically simulable and converge to the same geometric strategy as matched classical baselines, while two-qubit models with entangling gates learn a representationally distinct strategy that encodes context in inter-qubit entanglement -- confirmed by three independent causal tests (p < 0.0001, d = 0.89). On real quantum hardware, only the classical geometric strategy survives device noise; the entanglement strategy degrades to chance. These findings open mechanistic interpretability as a tool for the science of quantum language models and reveal a noise-expressivity tradeoff governing which learned strategies survive deployment.

📄 PDF Abstract BibTeX arXiv:2603.26494

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Quantum Language Model with Entanglement Embedding for Question Answering

2020-08-23 · Yi-Wei Chen, Yu Pan, Daoyi Dong

Quantum Language Models (QLMs) in which words are modelled as quantum superposition of sememes have demonstrated a high level of model transparency and good post-hoc interpretability. Nevertheless, in the current literat…

Language ModelingLanguage ModellingQuestion Answering

QiNN-QJ: A Quantum-inspired Neural Network with Quantum Jump for Multimodal Sentiment Analysis

2025-10-31 · Yiwei Chen, Kehuan Yan, Yu Pan, Daoyi Dong arxiv

Quantum theory provides non-classical principles, such as superposition and entanglement, that inspires promising paradigms in machine learning. However, most existing quantum-inspired fusion models rely solely on unitar…

Multimodal Sentiment Analysis

Causal Intervention Framework for Variational Auto Encoder Mechanistic Interpretability

2025-05-06 · Dip Roy

Mechanistic interpretability of deep learning models has emerged as a crucial research direction for understanding the functioning of neural networks. While significant progress has been made in interpreting discriminati…

DisentanglementSpecificity

QIXAI: A Quantum-Inspired Framework for Enhancing Classical and Quantum Model Transparency and Understanding

2024-10-21 · John M. Willis

The impressive performance of deep learning models, particularly Convolutional Neural Networks (CNNs), is often hindered by their lack of interpretability, rendering them "black boxes." This opacity raises concerns in cr…

Feature ImportanceTime Series Analysis

Feature Entanglement-based Quantum Multimodal Fusion Neural Network

2026-01-09 · Yu Wu, Qianli Zhou, Jie Geng, Xinyang Deng 외 arxiv

Multimodal learning aims to enhance perceptual and decision-making capabilities by integrating information from diverse sources. However, classical deep learning approaches face a critical trade-off between the high accu…