paper-with-me

홈 › Papers

A Geometric View of SRC: Learning Representations for Stable Residual Inference

2026-05-28 · Vangelis P. Oikonomou arxiv

Reconstruction-based inference assigns a class by comparing class-wise reconstruction residuals; Sparse Representation Classification (SRC) is a canonical instance whose reliability depends on the geometry of the learned representation. We adopt a strict training-inference separation: SRC is used only as a fixed test-time rule and is never differentiated, unrolled, or optimized during training. In a span-level idealization based on class-conditional spans and their associated projection residuals, we formalize residual-ordering stability through a residual margin and characterize geometric obstructions -- span overlap, dominance, and near-overlap via small principal angles -- that can collapse this margin in worst-case directions. This span-level theory is primary: it specifies when the idealized residual family is well-separated, and it provides a conditional solver-level interpretation for practical residual approximations (e.g., OMP) insofar as they remain close to the span-level residual ordering. Under explicit coverage and separation assumptions, we derive a quantitative lower bound on the (idealized) residual margin. Guided by these targets, we propose geometry-shaping objectives that promote masked within-class self-expressiveness, discourage cross-class reconstruction pathways and inter-class span alignment, and prevent collapse -- without invoking SRC residuals or predictions during training. Experiments on images (COIL-100), text (TREC), and EEG connectivity evaluate all representations under identical fixed SRC/OMP inference and report residual margins and geometric diagnostics; cross-entropy is included only as a reference geometry under the same evaluation protocol.

📄 PDF Abstract BibTeX arXiv:2605.29673

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Text as Partial Constraint: Core-Residual Alignment for Robust Vision-Language Learning

2026-07-03 · Chengzhen Yu, Canran Xiao, Siyuan Ma, Yang Liu arxiv

Vision-language alignment powers open-vocabulary recognition, retrieval, and LVLM grounding, yet natural captions are often underspecified, making similarity brittle and overly confident under paraphrase and omitted deta…

Adversarial Robustness

PRISM: Feed-Forward Single-Image 3D Reconstruction via Geometric Warp-Residual Modeling

2026-06-24 · Zhijie Zheng, Xinhao Xiang, Jiawei Zhang arxiv

Reconstructing 3D scenes from a single image is a fundamental challenge in computer vision, with broad applications in virtual reality, robotics, and content creation. Recent methods achieve outstanding performance by le…

3D Reconstruction

Multimodal-Prior-Guided Importance Sampling for Hierarchical Gaussian Splatting in Sparse-View Novel View Synthesis

2026-03-03 · Kaiqiang Xiong, Zhanke Wang, Ronggang Wang arxiv

We present multimodal-prior-guided importance sampling as the central mechanism for hierarchical 3D Gaussian Splatting (3DGS) in sparse-view novel view synthesis. Our sampler fuses complementary cues { -- } photometric r…

Novel View Synthesis

Stable Single-Pixel Contrastive Learning for Semantic and Geometric Tasks

2025-12-04 · Leonid Pogorelyuk, Niels Bracher, Aaron Verkleeren, Lars Kühmichel 외 arxiv

We pilot a family of stable contrastive losses for learning pixel-level representations that jointly capture semantic and geometric information. Our approach maps each pixel of an image to an overcomplete descriptor that…

Contrastive Learning

VGGT-Edit: Feed-forward Native 3D Scene Editing with Residual Field Prediction

2026-05-14 · Kaixin Zhu, Yiwen Tang, Yifan Yang, Renrui Zhang 외 arxiv

High-quality 3D scene reconstruction has recently advanced toward generalizable feed-forward architectures, enabling the generation of complex environments in a single forward pass. However, despite their strong performa…

3D scene Editing