paper-with-me

홈 › Papers

CoVisPose: Co-Visibility Pose Transformer for Wide-Baseline Relative Pose Estimation in 360◦ Indoor Panoramas

2022-10-01 · European Conference on Computer Vision 2022 10 · Will Hutchcroft, Yuguang Li, Ivaylo Boyadzhiev, Zhiqiang Wan, HaiYan Wang, Sing Bing Kang

We present CoVisPose, a new end-to-end supervised learning method for relative camera pose estimation in wide baseline 360 indoor panoramas. To address the challenges of occlusion, perspective changes, and textureless or repetitive regions, we generate rich representations for direct pose regression by jointly learning dense bidirectional visual overlap, correspondence, and layout geometry. We estimate three image column-wise quantities: co-visibility (the probability that a given column’s image content is seen in the other panorama), angular correspondence (angular matching of columns across panoramas), and floor layout (the vertical floor-wall boundary angle). We learn these dense outputs by applying a transformer over the image-column feature sequences, which cover the full 360 field-of-view (FoV) from both panoramas. The resultant rich representation supports learning robust relative poses with an efficient 1D convolutional decoder. In addition to learned direct pose regression with scale, our network also supports pose estimation through a RANSAC-based rigid registration of the predicted corresponding layout boundary points. Our method is robust to extremely wide baselines with very low visual overlap, as well as significant occlusions. We improve upon the SOTA by a large margin, as demonstrated on a large-scale dataset of real homes, ZInD.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Camera Pose EstimationDecoderPose Estimationregression

Similar Papers 제목 키워드 기반

Graph-CoVis: GNN-based Multi-view Panorama Global Pose Estimation

2023-04-26 · Negar Nejatishahidin, Will Hutchcroft, Manjunath Narayana, Ivaylo Boyadzhiev 외

In this paper, we address the problem of wide-baseline camera pose estimation from a group of 360$^\circ$ panoramas under upright-camera assumption. Recent work has demonstrated the merit of deep-learning for end-to-end …

Camera Pose EstimationGraph Neural NetworkPose Estimation

Beyond SEO: A Transformer-Based Approach for Reinventing Web Content Optimisation

2025-07-03 · Florian Lüttgenau, Imar Colic, Gervasio Ramirez arxiv

The rise of generative AI search engines is disrupting traditional SEO, with Gartner predicting 25% reduction in conventional search usage by 2026. This necessitates new approaches for web content visibility in AI-driven…

MAP Visibility Estimation for Large-Scale Dynamic 3D Reconstruction

2014-06-01 · CVPR 2014 6 · Hanbyul Joo, Hyun Soo Park, Yaser Sheikh

Many traditional challenges in reconstructing 3D motion, such as matching across wide baselines and handling occlusion, reduce in significance as the number of unique viewpoints increases. However, to obtain this benefit…

3D Reconstruction

HAVEN: Hierarchical Adversary-aware Visibility-Enabled Navigation with Cover Utilization using Deep Transformer Q-Networks

2025-11-29 · Mihir Chauhan, Damon Conover, Aniket Bera arxiv

Autonomous navigation in partially observable environments requires agents to reason beyond immediate sensor input, exploit occlusion, and ensure safety while progressing toward a goal. These challenges arise in many rob…

Reinforcement Learning

Partial Skeleton Visibility for Action Recognition: A Constrained Field-of-View Approach

2026-07-01 · Yingjie Dai, Tianyang Xu, Yanglin Deng, Xiao-Jun Wu 외 arxiv

Skeleton-based action recognition has achieved remarkable success by exploiting joint coordinates and their topological connections, yet prevailing methods overwhelmingly assume complete and clean skeleton inputs. In rea…

Action UnderstandingAction Recognition