paper-with-me

Papers

MixSA: Training-free Reference-based Sketch Extraction via Mixture-of-Self-Attention

2025-01-01 · Rui Yang, XiaoJun Wu, Shengfeng He

Current sketch extraction methods either require extensive training or fail to capture a wide range of artistic styles, limiting their practical applicability and versatility. We introduce Mixture-of-Self-Attention (MixSA), a training-free sketch extraction method that leverages strong diffusion priors for enhanced sketch perception. At its core, MixSA employs a mixture-of-self-attention technique, which manipulates self-attention layers by substituting the keys and values with those from reference sketches. This allows for the seamless integration of brushstroke elements into initial outline images, offering precise control over texture density and enabling interpolation between styles to create novel, unseen styles. By aligning brushstroke styles with the texture and contours of colored images, particularly in late decoder layers handling local textures, MixSA addresses the common issue of color averaging by adjusting initial outlines. Evaluated with various perceptual metrics, MixSA demonstrates superior performance in sketch quality, flexibility, and applicability. This approach not only overcomes the limitations of existing methods but also empowers users to generate diverse, high-fidelity sketches that more accurately reflect a wide range of artistic expressions.

📄 PDF Abstract BibTeX arXiv:2501.00816

Code (0)

등록된 구현이 없습니다.

Tasks

Decoder

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Semi-supervised reference-based sketch extraction using a contrastive learning framework

2024-07-19 · Chang Wook Seo, Amirsaman Ashtari, Junyong Noh

Sketches reflect the drawing style of individual artists; therefore, it is important to consider their unique styles when extracting sketches from color images for various applications. Unfortunately, most existing sketc…

Contrastive Learning

MixSarc: A Bangla-English Code-Mixed Corpus for Implicit Meaning Identification

2026-02-25 · Kazi Samin Yasar Alam, Md Tanbir Chowdhury, Tamim Ahmed, Ajwad Abrar 외 arxiv

Bangla-English code-mixing is widespread across South Asian social media, yet resources for implicit meaning identification in this setting remain scarce. Existing sentiment and sarcasm models largely focus on monolingua…

Humor Detection

Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch Generation

2025-10-18 · Rui Yang, Huining Li, Yiyi Long, Xiaojun Wu 외 arxiv

Generating sketches guided by reference styles requires precise transfer of stroke attributes, such as line thickness, deformation, and texture sparsity, while preserving semantic structure and content fidelity. To this …

Text to Sketch Generation with Multi-Styles

2025-11-06 · Tengjie Li, Shikui Tu, Lei Xu arxiv

Recent advances in vision-language models have facilitated progress in sketch generation. However, existing specialized methods primarily focus on generic synthesis and lack mechanisms for precise control over sketch sty…

Style Transfer

SketchingReality: From Freehand Scene Sketches To Photorealistic Images

2026-02-16 · Ahmed Bourouis, Mikhail Bessmeltsev, Yulia Gryaditskaya arxiv

Recent years have witnessed remarkable progress in generative AI, with natural language emerging as the most common conditioning input. As underlying models grow more powerful, researchers are exploring increasingly dive…