paper-with-me

Papers

BED-SAM2: Boundary-Enhanced-Depth SAM2 via Monocular Geometric Priors

2026-05-24 · Tyler Rust, Dara McNally, Kyle O'Donnell, Colin Kelly, Chandra Kambhamettu arxiv

Building upon the SAM2 vision foundation model for downstream segmentation, this study introduces Boundary Enhanced Depth (BED)-SAM2. The SAM2 Hiera encoder architecture is modified to directly encode monocular depth information from RGB images, thereby providing geometric cues that enhance object boundary delineation and facilitate the extraction of camouflaged object shapes. BED-SAM2 demonstrates competitive state-of-the-art performance across multiple salient and camouflaged object detection tasks with as few as five training epochs.

📄 PDF Abstract BibTeX arXiv:2605.24893

Code (0)

등록된 구현이 없습니다.

Tasks

Object Detection

Similar Papers 제목 키워드 기반

Enhanced Scale-aware Depth Estimation for Monocular Endoscopic Scenes with Geometric Modeling

2024-08-14 · Ruofeng Wei, Bin Li, Kai Chen, Yiyao Ma 외

Scale-aware monocular depth estimation poses a significant challenge in computer-aided endoscopic navigation. However, existing depth estimation methods that do not consider the geometric priors struggle to learn the abs…

Depth EstimationMonocular Depth Estimation

Rethinking Monocular Depth Embedding for Generalized Stereo Matching

2026-07-10 · Libo Lin, Shuangli Du, Minghua Zhao, Zhenzhen You 외 arxiv

Generally, monocular methods capture rich contextual priors but lack geometric precision, whereas stereo methods are geometrically accurate yet struggle in textureless and occluded regions. Several approaches attempt to …

Data Augmentation

MonoDGP: Monocular 3D Object Detection with Decoupled-Query and Geometry-Error Priors

2024-10-25 · CVPR 2025 1 · Fanqi Pu, Yifan Wang, Jiru Deng, Wenming Yang

Perspective projection has been extensively utilized in monocular 3D object detection methods. It introduces geometric priors from 2D bounding boxes and 3D object dimensions to reduce the uncertainty of depth estimation.…

3D Object DetectionDepth EstimationDepth PredictionMonocular 3D Object Detection+3

MonoInstance: Enhancing Monocular Priors via Multi-view Instance Alignment for Neural Rendering and Reconstruction

2025-03-24 · CVPR 2025 1 · Wenyuan Zhang, Yixiao Yang, Han Huang, Liang Han 외

Monocular depth priors have been widely adopted by neural rendering in multi-view based tasks such as 3D reconstruction and novel view synthesis. However, due to the inconsistent prediction on each view, how to more effe…

3D ReconstructionNeural RenderingNovel View Synthesis

HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos

2025-12-06 · Weitao Xiong, Zhiyuan Yuan, Jiahao Lu, Chengfeng Zhao 외 arxiv

Monocular dynamic video reconstruction faces significant challenges in dynamic human scenes due to geometric inconsistencies and resolution degradation issues. Existing methods lack 3D human structural understanding, pro…

Monocular Depth EstimationDynamic ReconstructionVideo Reconstruction