paper-with-me

홈 › Papers

PISA: Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers

2026-02-01 · Haopeng Li, Shitong Shao, Wenliang Zhong, Zikai Zhou, Lichen Bai, Hui Xiong, Zeke Xie arxiv

Diffusion Transformers are fundamental for video and image generation, but their efficiency is bottlenecked by the quadratic complexity of attention. While block sparse attention accelerates computation by attending only critical key-value blocks, it suffers from degradation at high sparsity by discarding context. In this work, we discover that attention scores of non-critical blocks exhibit distributional stability, allowing them to be approximated accurately and efficiently rather than discarded, which is essentially important for sparse attention design. Motivated by this key insight, we propose PISA, a training-free Piecewise Sparse Attention that covers the full attention span with sub-quadratic complexity. Unlike the conventional keep-or-drop paradigm that directly drop the non-critical block information, PISA introduces a novel exact-or-approximate strategy: it maintains exact computation for critical blocks while efficiently approximating the remainder through block-wise Taylor expansion. This design allows PISA to serve as a faithful proxy to full attention, effectively bridging the gap between speed and quality. Experimental results demonstrate that PISA achieves 1.91 times and 2.57 times speedups on Wan2.1-14B and Hunyuan-Video, respectively, while consistently maintaining the highest quality among sparse attention methods. Notably, even for image generation on FLUX, PISA achieves a 1.2 times acceleration without compromising visual quality. Code is available at: https://github.com/xie-lab-ml/piecewise-sparse-attention.

📄 PDF Abstract BibTeX arXiv:2602.01077

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

WISER: A Semantic Approach for Expert Finding in Academia based on Entity Linking

2018-05-10 · Paolo Cifariello, Paolo Ferragina, Marco Ponza

We present WISER, a new semantic search engine for expert finding in academia. Our system is unsupervised and it jointly combines classical language modeling techniques, based on text evidences, with the Wikipedia Knowle…

Entity LinkingLanguage ModelingLanguage Modelling

PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization

2025-11-13 · Runpeng Geng, Yanting Wang, Chenlong Yin, Minhao Cheng 외 arxiv

Long context LLMs are vulnerable to prompt injection, where an attacker can inject an instruction in a long context to induce an LLM to generate an attacker-desired output. Existing prompt injection defenses are designed…

Pixel-level and Semantic-level Adjustable Super-resolution: A Dual-LoRA Approach

2024-12-04 · CVPR 2025 1 · Lingchen Sun, Rongyuan Wu, Zhiyuan Ma, Shuaizheng Liu 외

Diffusion prior-based methods have shown impressive results in real-world image super-resolution (SR). However, most existing methods entangle pixel-level and semantic-level SR objectives in the training process, struggl…

Image Super-ResolutionSuper-Resolution

Widely Interpretable Semantic Representation: Frameless Meaning Representation for Broader Applicability

2023-09-12 · Lydia Feng, Gregor Williamson, Han He, Jinho D. Choi

This paper presents a novel semantic representation, WISeR, that overcomes challenges for Abstract Meaning Representation (AMR). Despite its strengths, AMR is not easily applied to languages or domains without predefined…

Abstract Meaning Representation

Widely Interpretable Semantic Representation: Frameless Meaning Representation for Broader Applicability

2021-09-17 · ACL ARR September 2021 9 · Anonymous

This paper presents a semantic representation called WISeR that overcomes challenges for Abstract Meaning Representation (AMR). Despite its richness and exapandability, AMR is not easily applied to languages or domains w…

Abstract Meaning Representation