paper-with-me

홈 › Papers

Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacement

2026-05-14 · Injin Kong, Hyoungjoon Lee, Yohan Jo arxiv

Continuous diffusion language models lag behind autoregressive transformers, partly because diffusion is applied in spaces poorly suited to language denoising and token recovery. We propose DiHAL, a geometry-guided diffusion-transformer hybrid that asks where diffusion should enter a pretrained transformer. DiHAL scores layers with geometry-based proxies, selects a diffusion-friendly hidden-state interface, and replaces the lower transformer prefix with a diffusion bridge while retaining the upper layers and original LM head. By reconstructing the selected-layer hidden state rather than tokens, DiHAL avoids direct continuous-to-discrete recovery. Experiments on 8B-scale backbones show that the geometry score predicts effective shallow insertion layers under a fixed bridge-training protocol and that hidden-state recovery improves over continuous diffusion baselines in a diagnostic comparison matching the diffusion/recovery training budget. These results suggest that hidden-state geometry helps identify where diffusion-based replacement is feasible inside pretrained language models.

📄 PDF Abstract BibTeX arXiv:2605.14368

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Layers to Networks: Comparing Neural Representations via Diffusion Geometry

2026-05-15 · Atharva Khandait, Jan E. Gerken arxiv

Diffusion geometry is a manifold learning framework that uses random walks defined by Markov transition matrices to characterize the geometry of a dataset at multiple scales. We use diffusion geometry for neural represen…

RoomDreamer: Text-Driven 3D Indoor Scene Synthesis with Coherent Geometry and Texture

2023-05-18 · Liangchen Song, Liangliang Cao, Hongyu Xu, Kai Kang 외

The techniques for 3D indoor scene capturing are widely used, but the meshes produced leave much to be desired. In this paper, we propose "RoomDreamer", which leverages powerful natural language to synthesize a new room …

Image GenerationIndoor Scene Synthesis

GeoBlock: Inferring Block Granularity from Dependency Geometry in Diffusion Language Models

2026-03-04 · Lipeng Wan, Junjie Ma, Jianhui Gu, Zeyang Liu 외 arxiv

Block diffusion enables efficient parallel refinement in diffusion language models, but its decoding behavior depends critically on block size. Existing block-sizing strategies rely on fixed rules or heuristic signals an…

Meta-Learned Basis Adaptation for Parametric Linear PDEs

2026-04-10 · Vikas Dwivedi, Monica Sigovan, Bruno Sixou arxiv

We propose a hybrid physics-informed framework for solving families of parametric linear partial differential equations (PDEs) by combining a meta-learned predictor with a least-squares corrector. The predictor, termed \…

HUG-VAS: A Hierarchical NURBS-Based Generative Model for Aortic Geometry Synthesis and Controllable Editing

2025-07-15 · Pan Du, Mingqi Xu, Xiaozhi Zhu, Jian-Xun Wang

Accurate characterization of vascular geometry is essential for cardiovascular diagnosis and treatment planning. Traditional statistical shape modeling (SSM) methods rely on linear assumptions, limiting their expressivit…

Denoising