paper-with-me

홈 › Papers

Layer-Order Inversion: Rethinking Latent Multi-Hop Reasoning in Large Language Models

2026-01-07 · Xukai Liu, Ye Liu, Jipeng Zhang, Yanghai Zhang, Kai Zhang, Qi Liu arxiv

Large language models (LLMs) perform well on multi-hop reasoning, yet how they internally compose multiple facts remains unclear. Recent work proposes \emph{hop-aligned circuit hypothesis}, suggesting that bridge entities are computed sequentially across layers before later-hop answers. Through systematic analyses on real-world multi-hop queries, we show that this hop-aligned assumption does not generalize: later-hop answer entities can become decodable earlier than bridge entities, a phenomenon we call \emph{layer-order inversion}, which strengthens with total hops. To explain this behavior, we propose a \emph{probabilistic recall-and-extract} framework that models multi-hop reasoning as broad probabilistic recall in shallow MLP layers followed by selective extraction in deeper attention layers. This framework is empirically validated through systematic probing analyses, reinterpreting prior layer-wise decoding evidence, explaining chain-of-thought gains, and providing a mechanistic diagnosis of multi-hop failures despite correct single-hop knowledge. Code is available at https://github.com/laquabe/Layer-Order-Inversion.

📄 PDF Abstract BibTeX arXiv:2601.03542

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Spatially-Adaptive Multilayer Selection for GAN Inversion and Editing

2022-06-16 · CVPR 2022 1 · Gaurav Parmar, Yijun Li, Jingwan Lu, Richard Zhang 외

Existing GAN inversion and editing methods work well for aligned objects with a clean background, such as portraits and animal faces, but often struggle for more difficult categories with complex scene layouts and object…

LSAP: Rethinking Inversion Fidelity, Perception and Editability in GAN Latent Space

2022-09-26 · Pu Cao, Lu Yang, Dongxu Liu, Zhiwei Liu 외

As the methods evolve, inversion is mainly divided into two steps. The first step is Image Embedding, in which an encoder or optimization process embeds images to get the corresponding latent codes. Afterward, the second…

Inverting Deep Generative models, One layer at a time

2019-06-18 · NeurIPS 2019 12 · Qi Lei, Ajil Jalal, Inderjit S. Dhillon, Alexandros G. Dimakis

We study the problem of inverting a deep generative model with ReLU activations. Inversion corresponds to finding a latent code vector that explains observed measurements as much as possible. In most prior works this is …

EEdit: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing

2025-03-13 · Zexuan Yan, Yue Ma, Chang Zou, Wenteng Chen 외

Inversion-based image editing is rapidly gaining momentum while suffering from significant computation overhead, hindering its application in real-time interactive scenarios. In this paper, we rethink that the redundancy…

E2Style: Improve the Efficiency and Effectiveness of StyleGAN Inversion

2021-04-15 · Tianyi Wei, Dongdong Chen, Wenbo Zhou, Jing Liao 외

This paper studies the problem of StyleGAN inversion, which plays an essential role in enabling the pretrained StyleGAN to be used for real image editing tasks. The goal of StyleGAN inversion is to find the exact latent …

Face Parsing