paper-with-me

Papers

Deep Compositional Spatial Models

2019-06-06 · Andrew Zammit-Mangion, Tin Lok James Ng, Quan Vu, Maurizio Filippone

Spatial processes with nonstationary and anisotropic covariance structure are often used when modelling, analysing and predicting complex environmental phenomena. Such processes may often be expressed as ones that have stationary and isotropic covariance structure on a warped spatial domain. However, the warping function is generally difficult to fit and not constrained to be injective, often resulting in `space-folding.' Here, we propose modelling an injective warping function through a composition of multiple elemental injective functions in a deep-learning framework. We consider two cases; first, when these functions are known up to some weights that need to be estimated, and, second, when the weights in each layer are random. Inspired by recent methodological and technological advances in deep learning and deep Gaussian processes, we employ approximate Bayesian methods to make inference with these models using graphics processing units. Through simulation studies in one and two dimensions we show that the deep compositional spatial models are quick to fit, and are able to provide better predictions and uncertainty quantification than other deep stochastic models of similar complexity. We also show their remarkable capacity to model nonstationary, anisotropic spatial data using radiances from the MODIS instrument aboard the Aqua satellite.

📄 PDF Abstract BibTeX arXiv:1906.02840

Code (0)

등록된 구현이 없습니다.

Tasks

Gaussian ProcessesUncertainty Quantification

Similar Papers 제목 키워드 기반

MultihopSpatial: Multi-hop Compositional Spatial Reasoning Benchmark for Vision-Language Model

2026-03-19 · Youngwan Lee, Soojin Jang, Yoorhim Cho, Seunghwan Lee 외 arxiv

Spatial reasoning is foundational for Vision-Language Models (VLMs), particularly when deployed as Vision-Language-Action (VLA) agents in physical environments. However, existing benchmarks predominantly focus on element…

Reinforcement LearningSpatial ReasoningAnswer SelectionVisual Grounding

DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning

2025-11-04 · Lachlan McPheat, Navdeep Kaur, Robert Blackwell, Alessandra Russo 외 arxiv

We introduce DecompSR, decomposed spatial reasoning, a large benchmark dataset (over 5m datapoints) and generation framework designed to analyse compositional spatial reasoning ability. The generation of DecompSR allows …

Spatial Reasoning

SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence

2025-06-09 · Ziyang Gong, Wenhao Li, Oliver Ma, Songyuan Li 외

Multimodal Large Language Models (MLLMs) have achieved remarkable progress in various multimodal tasks. To pursue higher intelligence in space, MLLMs require integrating multiple atomic spatial capabilities to handle com…

CompAlign: Improving Compositional Text-to-Image Generation with a Complex Benchmark and Fine-Grained Feedback

2025-05-16 · Yixin Wan, Kai-Wei Chang

State-of-the-art T2I models are capable of generating high-resolution images given textual prompts. However, they still struggle with accurately depicting compositional scenes that specify multiple objects, attributes, a…

AttributeImage GenerationText to Image GenerationText-to-Image Generation

Inverse Compositional Spatial Transformer Networks

2016-12-12 · CVPR 2017 7 · Chen-Hsuan Lin, Simon Lucey

In this paper, we establish a theoretical connection between the classical Lucas & Kanade (LK) algorithm and the emerging topic of Spatial Transformer Networks (STNs). STNs are of interest to the vision and learning comm…

ClassificationGeneral Classification