paper-with-me

Feature Upsampling

1개 벤치마크 · 논문 39편 · 이 태스크의 논문 보기 →

Benchmarks

ImageNet

결과 8개

Most implemented

Deep Image Prior

2017-11-29 · 구현 14개

On Point Affiliation in Feature Upsampling

2023-07-17 · 구현 2개

Papers

RaysUp: Ultra-light Universal Feature Upsampling via Geometry-Aware Ray Representation

2026-06-22 · Yuchuan Ding, Linfei Li, Lin Zhang, Ying Shen arxiv

Pre-trained Vision Foundation Models (VFMs) have become central to modern computer vision due to their powerful semantic representations and strong generalization ability. However, their patchified or pooled outputs are …

Feature Upsampling

ViT-Up: Faithful Feature Upsampling for Vision Transformers

2026-06-12 · Krispin Wandel, Jingchuan Wang, Hesheng Wang arxiv

Vision Transformers (ViTs) have become a dominant architecture for visual representation learning, providing exceptionally strong and broadly reusable backbone features. However, ViTs are commonly operated on relatively …

Semantic correspondenceRepresentation LearningSemantic SegmentationFeature Upsampling

Weighted Reverse Convolution for Feature Upsampling

2026-05-17 · Wentong Li, Zhiyuan Qi, Zichen Zhao, Kai Zhang 외 arxiv

Pre-trained vision foundation models (VFMs) provide strong semantic representations, yet their patch-level features are inherently coarse, limiting their effectiveness on tasks requiring fine-grained localization, dense …

Video Object SegmentationComputational EfficiencyFeature UpsamplingDepth Estimation

DINO Soars: DINOv3 for Open-Vocabulary Semantic Segmentation of Remote Sensing Imagery

2026-05-04 · Ryan Faulkenberry, Saurabh Prasad arxiv

The remote sensing (RS) domain suffers from a lack of densely labeled datasets, which are costly to obtain. Thus, models that can segment RS imagery well without supervised fine-tuning are valuable, but existing solution…

Open Vocabulary Semantic SegmentationFeature Upsampling

Frozen Vision Transformers for Dense Prediction on Small Datasets: A Case Study in Arrow Localization

2026-04-18 · Maxwell Shepherd arxiv

We present a system for automated detection, localization, and scoring of arrow punctures on 40\,cm indoor archery target faces, trained on only 48 annotated photographs (5{,}084 punctures). Our pipeline combines three c…

Feature Upsampling

HD-VGGT: High-Resolution Visual Geometry Transformer

2026-03-28 · Tianrun Chen, Yuanqi Hu, Yidong Han, Hanjie Xu 외 arxiv

High-resolution imagery is essential for accurate 3D reconstruction, as many geometric details only emerge at fine spatial scales. Recent feed-forward approaches, such as the Visual Geometry Grounded Transformer (VGGT), …

Feature Upsampling3D Reconstruction

전체 39편 보기 →