paper-with-me

홈 › Papers

$\textit{BlockFormer}$ : Transformer-based inference from interaction maps

2026-05-20 · Eloïse Touron, Pedro L. C. Rodrigues, Julyan Arbel, Nelle Varoquaux, Michael Arbel arxiv

Inference from interaction maps, such as centromere identification from genome-wide chromosome conformation capture techniques -- notably Hi-C -- can be formulated as a generic inverse problem: infer a set of parameters given a map summarizing pairwise interactions between entities through blocks of variable numbers and sizes. In this work, we introduce a data-driven approach that leverages shared structure between these maps, such as global alignment between localized patterns, while handling the variability in number and size of entities arising in real-world data. Our approach relies on a transformer architecture capable of handling such variability and a custom simulator to generate abundant, yet computationally cheap synthetic data for training. Applied to the problem of centromere localization, the method accurately recovers their genomic positions across a wide range of species of various genome sizes.

📄 PDF Abstract BibTeX arXiv:2605.21617

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SOFI: Multi-Scale Deformable Transformer for Camera Calibration with Enhanced Line Queries

2024-09-23 · Sebastian Janampa, Marios Pattichis

Camera calibration consists of estimating camera parameters such as the zenith vanishing point and horizon line. Estimating the camera parameters allows other tasks like 3D rendering, artificial reality effects, and obje…

Camera Calibration

Improving Mandarin Speech Recogntion with Block-augmented Transformer

2022-07-24 · Xiaoming Ren, Huifeng Zhu, Liuwei Wei, Minghui Wu 외

Recently Convolution-augmented Transformer (Conformer) has shown promising results in Automatic Speech Recognition (ASR), outperforming the previous best published Transformer Transducer. In this work, we believe that th…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderLanguage Modeling+3

TexTailor: Inference-Time Textual Guidance Tailoring for Multimodal Diffusion Transformers

2026-01-05 · Binglei Li, Mengping Yang, Zhiyu Tan, Junping Zhang 외 arxiv

Recent breakthroughs of transformer-based diffusion models, particularly with Multimodal Diffusion Transformers (MMDiT) driven models like FLUX and Qwen Image, have facilitated thrilling experiences in visual generation.…

Text-to-Image GenerationImage Editing

Trie-Aware Transformers for Generative Recommendation

2026-02-25 · Zhenxiang Xu, Jiawei Chen, Sirui Chen, Yong He 외 arxiv

Generative recommendation (GR) aligns with advances in generative AI by casting next-item prediction as token-level generation rather than score-based ranking. Most GR methods adopt a two-stage pipeline: (i) \textit{item…

HorNet: Efficient High-Order Spatial Interactions with Recursive Gated Convolutions

2022-07-28 · Yongming Rao, Wenliang Zhao, Yansong Tang, Jie zhou 외

Recent progress in vision Transformers exhibits great success in various tasks driven by the new spatial modeling mechanism based on dot-product self-attention. In this paper, we show that the key ingredients behind the …

Image ClassificationObject DetectionSemantic SegmentationVocal Bursts Intensity Prediction