paper-with-me

Papers Visual Localization

“Visual Localization” 태그가 달린 논문 494편 · 필터 해제

Adversarial Attacks on Robot Localization Systems via Deep Feature Perturbation

2026-06-01 · Zhenyu Li, Tianyi Shang arxiv

Robot localization systems are critical for autonomous navigation and safety. Adversarial perturbations can mislead these systems, resulting in mislocalization, navigation errors, or unsafe interactions, especially in mi…

Visual Localization

iVGR: Internalizing Visually Grounded Reasoning for MLLMs with Reinforcement Learning

2026-05-29 · Chang-Bin Zhang, Yujie Zhong, Qiang Zhang, Kai Han arxiv

While visually grounded Chain-of-Thought (CoT) has emerged as a promising paradigm to enhance fine-grained perception in multimodal large language models (MLLMs), its efficacy during the inference phase remains underexpl…

Reinforcement LearningVisual LocalizationVisual Grounding

MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue

2026-05-20 · Anna Deichler, Jim O'Regan, Fethiye Irmak Dogan, Lubos Marcinek 외 arxiv

Grounding language in the physical world requires AI systems to interpret references that emerge dynamically during conversation. While current vision-language models (VLMs) excel at static image tasks, they struggle to …

Visual Localization

Depth2Pose: A Pose-Based Benchmark for Monocular Depth Estimation without Ground-Truth Depth

2026-05-19 · Viktor Kocur, Sithu Aung, Gabrielle Flood, Yaqing Ding 외 arxiv

Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted depth is increasingly used as an input signal for downstream tasks su…

Monocular Depth EstimationCamera Pose EstimationVisual Localization

Efficient Sparse-to-Dense Visual Localization via Compact Gaussian Scene Representation and Accelerated Dense Pose Estimation

2026-05-18 · Zizhuo Li, Songchu Deng, Linfeng Tang, Jiayi Ma arxiv

This letter presents LiteLoc, a novel and efficient localizer built on 3D Gaussian Splatting (3DGS). The previous state-of-the-art (SoTA) sparse-to-dense localizer, STDLoc, has shown remarkable localization capability bu…

Visual LocalizationPose Estimation

SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability

2026-05-15 · Amin Karimi Monsefi, Abolfazl Meyarian, Mridul Khurana, Shuheng Wang 외 arxiv

Animals are described as effectively camouflaged when they blend seamlessly with their surrounding, yet no standardized quantitative measure of this seamlessness exists. We address this gap by framing camouflage evaluati…

Visual Localization

PoseCompass: Intelligent Synthetic Pose Selection for Visual Localization

2026-05-12 · Yanan Zhou, Zhaoyan Qian, Yanli Li, Nan Yang 외 arxiv

In visual localization, Absolute Pose Regression (APR) enables real-time 6-DoF camera pose inference from single images, yet critically depends on fine-tuning data quality and coverage. While recent methods leverage 3D G…

Novel View SynthesisVisual LocalizationData Augmentation

Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization

2026-05-08 · Miso Lee, Sangeek Hyun, Yerim Jeon, Jae-Pil Heo arxiv

While Gaussian Splatting-based Feature Fields (GSFFs) have shown promise for visual localization, this paper highlights that photometrically optimized GSFFs are inherently ill-suited for 2D-3D matching. The volumetric ex…

Visual LocalizationPose Estimation

ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian Splatting

2026-05-06 · Yingdong Gu, Shaocheng Yan, Zhenjun Zhao, Yuan Kou 외 arxiv

Visual localization is a core technology for augmented reality and autonomous navigation. Recent methods combine the efficient rendering of 3D Gaussian Splatting (3DGS) with feature-based localization. These methods rely…

Visual Localization

MSACT: Multistage Spatial Alignment for Stable Low-Latency Fine Manipulation

2026-05-01 · Xianbo Cai, Hideyuki Ichiwara, Masaki Yoshikawa, Tetsuya Ogata arxiv

Real-world fine manipulation, particularly in bimanual manipulation, typically requires low-latency control and stable visual localization, while collecting large-scale data is costly and limited demonstrations may lead …

Visual LocalizationObject Tracking

Depth-Guided Privacy-Preserving Visual Localization Using 3D Sphere Clouds

2026-05-01 · Heejoon Moon, Jongwoo Lee, Jeonggon Kim, Je Hyeong Hong arxiv

The emergence of deep neural networks capable of revealing high-fidelity scene details from sparse 3D point clouds has raised significant privacy concerns in visual localization involving private maps. Lifting map points…

Camera Pose EstimationVisual LocalizationPoint Clouds

COMPASS: COmpact Multi-channel Prior-map And Scene Signature for Floor-Plan-Based Visual Localization

2026-04-28 · Muhammad Shaheer, Miguel Fernandez-Cortizas, Asier Bikandi-Noya, Holger Voos 외 arxiv

Architectural floor plans are widely available priors which contain not only geometry but also the semantic information of the environment, yet existing localization methods largely ignore this semantic information. To a…

Visual Localization

Revisiting Geometric Obfuscation with Dual Convergent Lines for Privacy-Preserving Image Queries in Visual Localization

2026-04-24 · Jeonggon Kim, Heejoon Moon, Je Hyeong Hong arxiv

Privacy-Preserving Image Queries (PPIQ) are an emerging mechanism for cloud-based visual localization, enabling pose estimation from obfuscated features instead of private images or raw keypoints. However, the main appro…

Visual LocalizationPose Estimation

Continual Hand-Eye Calibration for Open-world Robotic Manipulation

2026-04-17 · Fazeng Li, Gan Sun, Chenxi Liu, Yao He 외 arxiv

Hand-eye calibration through visual localization is a critical capability for robotic manipulation in open-world environments. However, most deep learning-based calibration models suffer from catastrophic forgetting when…

Visual LocalizationContinual Learning

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

2026-04-17 · Siddhant Bharadwaj, Ashish Vashist, Fahimul Aleem, Shruti Vyas arxiv

Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent advances in Vision-Language Models (VLMs) have demonstrated strong …

Multimodal ReasoningVisual LocalizationImage Matching

SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation

2026-04-15 · Songlin Du, Xiaoyong Lu, Yaping Yan, Guobao Xiao 외 arxiv

Local feature matching plays a critical role in understanding the correspondence between cross-view images. However, traditional methods are constrained by the inherent local nature of feature descriptors, limiting their…

Homography EstimationVisual LocalizationPose EstimationImage Matching

Seeing Through Touch: Tactile-Driven Visual Localization of Material Regions

2026-04-13 · Seongyu Kim, Seungwoo Lee, Hyeonggon Ryu, Joon Son Chung 외 arxiv

We address the problem of tactile localization, where the goal is to identify image regions that share the same material properties as a tactile input. Existing visuo-tactile methods rely on global alignment and thus fai…

Visual Localization

Privacy-Preserving Structureless Visual Localization via Image Obfuscation

2026-04-13 · Vojtech Panek, Patrik Beliansky, Zuzana Kukelova, Torsten Sattler arxiv

Visual localization is the task of estimating the camera pose of an image relative to a scene representation. In practice, visual localization systems are often cloud-based. Naturally, this raises privacy concerns in ter…

Visual Localization

AsymLoc: Towards Asymmetric Feature Matching for Efficient Visual Localization

2026-04-10 · Mohammad Omama, Gabriele Berton, Eric Foxlin, Yelin Kim arxiv

Precise and real-time visual localization is critical for applications like AR/VR and robotics, especially on resource-constrained edge devices such as smart glasses, where battery life and heat dissipation can be a prim…

Visual Localization

LSGS-Loc: Towards Robust 3DGS-Based Visual Localization for Large-Scale UAV Scenarios

2026-04-07 · Xiang Zhang, Tengfei Wang, Fang Xu, Xin Wang 외 arxiv

Visual localization in large-scale UAV scenarios is a critical capability for autonomous systems, yet it remains challenging due to geometric complexity and environmental variations. While 3D Gaussian Splatting (3DGS) ha…

Visual LocalizationPose Estimation
← 이전 21–40 / 494 다음 →