paper-with-me

홈 › Papers

Deep spatial context: when attention-based models meet spatial regression

2024-01-18 · Paulina Tomaszewska, Elżbieta Sienkiewicz, Mai P. Hoang, Przemysław Biecek

We propose 'Deep spatial context' (DSCon) method, which serves for investigation of the attention-based vision models using the concept of spatial context. It was inspired by histopathologists, however, the method can be applied to various domains. The DSCon allows for a quantitative measure of the spatial context's role using three Spatial Context Measures: $SCM_{features}$, $SCM_{targets}$, $SCM_{residuals}$ to distinguish whether the spatial context is observable within the features of neighboring regions, their target values (attention scores) or residuals, respectively. It is achieved by integrating spatial regression into the pipeline. The DSCon helps to verify research questions. The experiments reveal that spatial relationships are much bigger in the case of the classification of tumor lesions than normal tissues. Moreover, it turns out that the larger the size of the neighborhood taken into account within spatial regression, the less valuable contextual information is. Furthermore, it is observed that the spatial context measure is the largest when considered within the feature space as opposed to the targets and residuals.

📄 PDF Abstract BibTeX arXiv:2401.10044

Code (1)

ptomaszewska/dscon 공식 구현 pytorch

Tasks

regression

Similar Papers 제목 키워드 기반

MEET: A Million-Scale Dataset for Fine-Grained Geospatial Scene Classification with Zoom-Free Remote Sensing Imagery

2025-03-14 · Yansheng Li, Yuning Wu, Gong Cheng, Chao Tao 외

Accurate fine-grained geospatial scene classification using remote sensing imagery is essential for a wide range of applications. However, existing approaches often rely on manually zooming remote sensing images at diffe…

ClassificationScene Classification

Attention meets Geometry: Geometry Guided Spatial-Temporal Attention for Consistent Self-Supervised Monocular Depth Estimation

2021-10-15 · Patrick Ruhkamp, Daoyi Gao, Hanzhi Chen, Nassir Navab 외

Inferring geometrically consistent dense 3D scenes across a tuple of temporally consecutive images remains challenging for self-supervised monocular depth prediction pipelines. This paper explores how the increasingly po…

Depth EstimationDepth PredictionMonocular Depth Estimation

ASoBO: Attentive Beamformer Selection for Distant Speaker Diarization in Meetings

2024-06-05 · Theo Mariotte, Anthony Larcher, Silvio Montresor, Jean-Hugh Thomas

Speaker Diarization (SD) aims at grouping speech segments that belong to the same speaker. This task is required in many speech-processing applications, such as rich meeting transcription. In this context, distant microp…

speaker-diarizationSpeaker Diarization

When Earth Foundation Models Meet Diffusion: An Application to Land Surface Temperature Super-Resolution

2026-04-18 · Yiheng Chen, Zihui Ma, Peishi Jiang, Yilong Dai 외 arxiv

Land surface temperature (LST) super-resolution is important for environmental monitoring. However, it remains challenging as coarse thermal observations severely underdetermine fine-scale structure. In this paper, we pr…

TAPTRv3: Spatial and Temporal Context Foster Robust Tracking of Any Point in Long Video

2024-11-27 · Jinyuan Qu, Hongyang Li, Shilong Liu, Tianhe Ren 외

In this paper, we present TAPTRv3, which is built upon TAPTRv2 to improve its point tracking robustness in long videos. TAPTRv2 is a simple DETR-like framework that can accurately track any point in real-world videos wit…

Point Tracking