paper-with-me

홈 › Papers

Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams

2026-01-14 · Lachlan Holden, Feras Dayoub, Alberto Candela, David Harvey, Tat-Jun Chin arxiv

Accurate localisation in planetary robotics enables the advanced autonomy required to support the increased scale and scope of future missions. The successes of the Ingenuity helicopter and multiple planetary orbiters lay the groundwork for future missions that use ground-aerial robotic teams. In this paper, we consider rovers using machine learning to localise themselves in a local aerial map using limited field-of-view monocular ground-view RGB images as input. A key consideration for machine learning methods is that real space data with ground-truth position labels suitable for training is scarce. In this work, we propose a novel method of localising rovers in an aerial map using cross-view-localising dual-encoder deep neural networks. We leverage semantic segmentation with vision foundation models and high volume synthetic data to bridge the domain gap to real images. We also contribute a new cross-view dataset of real-world rover trajectories with corresponding ground-truth localisation data captured in a planetary analogue facility, plus a high volume dataset of analogous synthetic image pairs. Using particle filters for state estimation with the cross-view networks allows accurate position estimation over simple and complex trajectories based on sequences of ground-view images.

📄 PDF Abstract BibTeX arXiv:2601.09107

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Segmentation

Similar Papers 제목 키워드 기반

Single Image, Any Face: Generalisable 3D Face Generation

2024-09-25 · Wenqing Wang, Haosen Yang, Josef Kittler, Xiatian Zhu

The creation of 3D human face avatars from a single unconstrained image is a fundamental task that underlies numerous real-world vision and graphics applications. Despite the significant progress made in generative model…

Face Generation

DoDo Learning: DOmain-DemOgraphic Transfer in Language Models for Detecting Abuse Targeted at Public Figures

2023-07-31 · Angus R. Williams, Hannah Rose Kirk, Liam Burke, Yi-Ling Chung 외

Public figures receive a disproportionate amount of abuse on social media, impacting their active participation in public life. Automated systems can identify abuse at scale but labelling training data is expensive, comp…

text-classificationText Classification

There is more to graphs than meets the eye: Learning universal features with self-supervision

2023-05-31 · Laya Das, Sai Munikoti, Nrushad Joshi, Mahantesh Halappanavar

We study the problem of learning features through self-supervision that are generalisable to multiple graphs. State-of-the-art graph self-supervision restricts training to only one graph, resulting in graph-specific mode…

Node ClassificationRepresentation LearningSelf-Supervised Learning

Towards Generalisable Time Series Understanding Across Domains

2024-10-09 · Özgün Turgut, Philip Müller, Martin J. Menten, Daniel Rueckert

Recent breakthroughs in natural language processing and computer vision, driven by efficient pre-training on large datasets, have enabled foundation models to excel on a wide range of tasks. However, this potential has n…

BenchmarkingTime SeriesTime Series Analysis

MammoDG: Generalisable Deep Learning Breaks the Limits of Cross-Domain Multi-Center Breast Cancer Screening

2023-08-02 · Yijun Yang, Shujun Wang, Lihao Liu, Sarah Hickman 외

Breast cancer is a major cause of cancer death among women, emphasising the importance of early detection for improved treatment outcomes and quality of life. Mammography, the primary diagnostic imaging test, poses chall…

Decision MakingDiagnostic