paper-with-me

홈 › Papers

Spatial Adapter: Structured Spatial Decomposition and Closed-Form Covariance for Frozen Predictors

2026-05-12 · Wen-Ting Wang, Wei-Ying Wu, Hao-Yun Huang, Xuan-Chun Wang arxiv

We present the Spatial Adapter, a parameter-efficient post-hoc layer that equips any frozen first-stage predictor with a structured spatial representation of its residual field and an induced closed-form spatial covariance. The adapter operates as a cascade second stage on residuals, jointly learning a spatially regularized orthonormal basis and per-sample scores via a tractable mini-batch ADMM procedure, without modifying any first-stage parameter. Because the first-stage parameters are frozen, the adapter does not retrain the backbone; its role is to supply a compressed distributional summary of the residual field. Smoothness, sparsity, and orthogonality together turn a generic low-rank factorization into an identifiable spatial representation whose induced residual covariance admits a closed-form low-rank-plus-noise estimator; the effective rank is determined data-adaptively by spectral thresholding, while the nominal rank K is an optimization-side upper bound only. This covariance enables kriging-style spatial prediction at unobserved locations, with plug-in uncertainty quantification as a secondary downstream use. Across synthetic data, Weather2K for spatial-holdout prediction, and GWHD patch grids as a basis-transferability diagnostic, the adapter recovers residual spatial structure when paired with frozen first stages from linear models to deep spatiotemporal and vision backbones; the added representation uses fewer than K(N+T) parameters alongside a compact residual-trend network.

📄 PDF Abstract BibTeX arXiv:2605.11394

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

From Building Blocks to Planning: Multi-Step Spatial Reasoning in LLMs with Reinforcement Learning

2025-12-31 · Amir Tahmasbi, Sadegh Majidi, Kazem Taram, Aniket Bera arxiv

Spatial reasoning in large language models (LLMs) has gained increasing attention due to applications in navigation and planning. Despite strong general language capabilities, LLMs still struggle with spatial transformat…

Reinforcement LearningSpatial Reasoning

AnimeAdapter: A Modular Adapter for Appearance-Consistent Anime Character Generation

2026-05-17 · Yixuan Han arxiv

We present a lightweight appearance adapter for Stable Diffusion that enables controllable and consistent anime character generation under diverse editing conditions. Instead of relying on large-scale vision-language mod…

Adapting Prithvi-EO for Fallow Detection for Food-Water Nexus: ViT-Adapter Necks and Parameter-Efficient Backbone tuning of Geospatial Foundation Model

2026-06-10 · Sk Muhammad Asif, Orhun Aydin arxiv

Understanding spatial distribution of fallow land is important for optimizing the food-water (FW) nexus, given fallowing's role in crop rotation and water conservation. Fallow is a low accuracy class in USDA Cropland Dat…

Object Detection

Structured Hyperedge Adaptation for Parameter-Efficient Fine-Tuning of Vision Transformers

2026-06-21 · Edwin Kwadwo Tenagyei, Lei Wang, Ugochukwu Ejike Akpudo, Jun Zhou 외 arxiv

Parameter-efficient fine-tuning (PEFT) has become a practical solution for adapting large pretrained vision transformers (ViTs) to downstream tasks while updating only a small subset of parameters. However, existing adap…

parameter-efficient fine-tuning

ViewSRD: 3D Visual Grounding via Structured Multi-View Decomposition

2025-07-15 · Ronggang Huang, Haoxin Yang, Yan Cai, Xuemiao Xu 외

3D visual grounding aims to identify and localize objects in a 3D space based on textual descriptions. However, existing methods struggle with disentangling targets from anchors in complex multi-anchor queries and resolv…

3D visual groundingVisual Grounding